HuggingFace

GameCraft-Bench: Can Agents Build Playable Games End-to-End in a Real Game Engine?

Game generation is an emerging application of coding agents, requiring models to transform natural-language specifica...

AI 聚合
2026-06-17
HuggingFace

A Gradient Perspective on RLVR Stability and Winner Advantage Policy Optimization

Reinforcement learning with verifiable rewards (RLVR) improves language-model reasoning, but GRPO-style optimization ...

AI 聚合
2026-06-17
HuggingFace

Beyond Monolingual Deep Research: Evaluating Agents and Retrievers with Cross-Lingual BrowseComp-Plus

Deep research agents are increasingly evaluated on their ability to search for evidence, reason over retrieved source...

AI 聚合
2026-06-17
HuggingFace

OPD-Evolver: Cultivating Holistic Agent Evolver via On-Policy Distillation

Memory has become a standard substrate for self-evolving agents, yet retaining experience is not the same as learning...

AI 聚合
2026-06-17
HuggingFace

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients

Knowledge distillation transfers a teacher's competence to a small student but is brittle in the small-student regime...

AI 聚合
2026-06-17
HuggingFace

LoopCoder-v2: Only Loop Once for Efficient Test-Time Computation Scaling

Looped Transformers scale latent computation by repeatedly applying shared blocks, but sequential looping increases l...

AI 聚合
2026-06-17
HuggingFace

ProCUA-SFT Technical Report

Training computer-use agents (CUAs) -- models that interact with graphical desktops through screenshots and keyboard/...

AI 聚合
2026-06-17
HuggingFace

Variable-Width Transformers

Scaling model size, specifically depth and width, has driven significant progress in transformer-based language model...

AI 聚合
2026-06-17
HuggingFace

ACE-Ego-0: Unifying Egocentric Human and Robotic Data for VLA Pretraining

Vision-Language-Action (VLA) models benefit from large-scale and diverse embodied data, yet scaling robot trajectory ...

AI 聚合
2026-06-17
HuggingFace

Show the Signal, Hide the Noise: Spectral Forcing for Pixel-Space Diffusion

Pixel-space diffusion models are trained on full-bandwidth noisy images, yet the useful signal available to the denoi...

AI 聚合
2026-06-17
HuggingFace

MotionVLA: Vision-Language-Action Model for Humanoid Motion

Generating realistic humanoid motion from scene images and text involves both low-frequency pose semantics and high-f...

AI 聚合
2026-06-17
HuggingFace

ChLogic: Evaluating Robustness of Logical Reasoning in Chinese Expressions

Large language models perform increasingly well on standardized logical reasoning benchmarks, but whether this abilit...

AI 聚合
2026-06-17
首页 上一页 第 196 / 212 页 下一页 末页