arXiv

When LLMs Read Tables Carelessly: Measuring and Reducing Data Referencing Errors

While large language models (LLMs) perform well on table tasks, they still make data referencing errors (DREs), i.e.,...

AI 聚合
2026-07-01
arXiv

Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs

Metacognition is a critical component of intelligence that describes the ability to monitor and regulate one's own co...

AI 聚合
2026-07-01
arXiv

QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents

LLM agents increasingly act over long horizons, where a single trajectory can contain hundreds or thousands of action...

AI 聚合
2026-07-01
arXiv

Introspective Coupling: Self-Explanation Training Tracks Behavioral Change Despite Fixed Supervision

When does training language models (LMs) to generate explanations of their predictions yield faithful introspection, ...

AI 聚合
2026-07-01
HuggingFace

DOPD: Dual On-policy Distillation

On-policy distillation (OPD) offers superior capacity transfer by supervising student-sampled trajectories with dense...

AI 聚合
2026-07-01
HuggingFace

Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks

Would experience designing faster GPU kernels also help close in on a long-standing open mathematical conjecture? Lar...

AI 聚合
2026-07-01
HuggingFace

AVTok: 1D Unified Tokenization for Holistic Audio-Video Generation

Audio-video generation has recently gained unprecedented research attention, aiming to synthesize high-quality soundi...

AI 聚合
2026-07-01
HuggingFace

Orca: The World is in Your Mind

We introduce Orca, an initial instantiation of a general world foundation model. Orca learns a unified world latent s...

AI 聚合
2026-07-01
HuggingFace

MemLearner: Learning to Query Context memory for Video World Models

Video World Models are interactive video generation models that predict future world states based on user actions and...

AI 聚合
2026-07-01
HuggingFace

PolyFlow: Continuous Topology Embedding Flow Matching for Artist-style Mesh Generation

Autoregressive Transformers dominate high-quality mesh generation by producing artist-worthy topologies, yet their in...

AI 聚合
2026-07-01
HuggingFace

Xiaomi-GUI-0 Technical Report

Graphical user interface (GUI) agents build on vision-language models to complete user tasks end-to-end in real appli...

AI 聚合
2026-07-01
HuggingFace

BlockPilot: Instance-Adaptive Policy Learning for Diffusion-based Speculative Decoding

Speculative decoding accelerates inference by using a lightweight draft model to generate candidate tokens in paralle...

AI 聚合
2026-07-01
首页 上一页 第 163 / 215 页 下一页 末页