HuggingFace

OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents

LLM agents are increasingly applied to open-ended everyday requests that span work, study, and life. These tasks are ...

AI 聚合
2026-08-06
HuggingFace

WorldCycle: Self-Verifiable Reinforcement Learning for Long-Horizon Video World Models

Interactive video world models are essential for long-horizon planning and exploration, yet they suffer from compound...

AI 聚合
2026-08-06
HuggingFace

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation

Leveraging pre-trained vision-language models (VLMs) to construct vision-language-action (VLA) models has emerged as ...

AI 聚合
2026-08-06
HuggingFace

Distill Where You Fail: Recovering Learning Signals of Negative RL-Groups from Adaptive Teacher Guidance

Reinforcement learning with verifiable rewards (RLVR) has become a standard paradigm for post-training large language...

AI 聚合
2026-08-06
HuggingFace

K-EXAONE 2.0 Technical Report

This technical report presents K-EXAONE 2.0, an open-weight multilingual foundation model developed by LG AI Research...

AI 聚合
2026-08-06
HuggingFace

Ego2Robot: Scalable Robot Data Synthesis from Egocentric Human Data

Learning generalizable robot manipulation policies requires large-scale and diverse demonstration data. Egocentric hu...

AI 聚合
2026-08-06
HuggingFace

Consistency-Driven Co-Evolution for Self-Supervised Cross-Representation Learning

As chart images, tabular data, and visualization code play increasingly important roles across diverse domains, cross...

AI 聚合
2026-08-06
HuggingFace

Agent Against Agent: An Agentic System for Automatic Prompt Injection Red Teaming

Prompt injection poses significant security risks to LLM agents. Efficient and effective red-teaming is therefore cri...

AI 聚合
2026-08-06
HuggingFace

AVE-Compass: Towards Holistic Evaluation for Audio-Video Editing Abilities

While instruction-based video editing has advanced rapidly, real-world videos contain tightly coupled audio and visua...

AI 聚合
2026-08-06
HuggingFace

GDPevo: Evaluating Agent Self-Evolution on Real Business Tasks

Agent self-evolution updates an agent's persistent state from prior experience and reuses it to solve related tasks m...

AI 聚合
2026-08-06
HuggingFace

FocusMem: Factorizing Content, Readout, and Trust in Latent GUI Memory

GUI agents must remember both useful experience from earlier tasks and unfinished progress in the current interaction...

AI 聚合
2026-08-06
HuggingFace

ABSeeker: Training Long-Horizon Search Agents via Answer-Backtracked Credit Assignment

Long-horizon search agents must make multiple sequential actions (steps) to search, retrieve, verify, and integrate e...

AI 聚合
2026-08-06
首页 上一页 第 77 / 212 页 下一页 末页