HuggingFace

RL-Index: Reinforcement Learning for Retrieval Index Reasoning

Retrieving external knowledge is essential for solving real-world tasks, yet it remains challenging when the relation...

AI 聚合
2026-06-25
HuggingFace

ShutterMuse: Capture-Time Photography Guidance with MLLMs

Real-world photography requires capture-time guidance for both camera framing and subject pose. Yet existing aestheti...

AI 聚合
2026-06-25
HuggingFace

MVTrack4Gen: Multi-View Point Tracking as Geometric Supervision for 4D Video Generation

Synthesizing a novel-view video from a monocular reference video along a target camera trajectory requires both geome...

AI 聚合
2026-06-25
HuggingFace

Constraint Tax in Open-Weight LLMs: An Empirical Study of Tool Calling Suppression Under Structured Output Constraints

Tool Calling and Structured Output are two core capabilities of modern Agent systems, yet their interaction under joi...

AI 聚合
2026-06-25
HuggingFace

ChartWalker: Benchmarking the Cross-Chart RAG Task

Cross-Chart Retrieval-Augmented Generation (RAG) is critical for complex multi-modal analytical tasks in scientific, ...

AI 聚合
2026-06-25
HuggingFace

AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning

Large language models are increasingly deployed as agents that reason over documents rather than answer from parametr...

AI 聚合
2026-06-25
HuggingFace

Semantic Browsing: Controllable Diversity for Image Generation

Modern text-to-image models excel in visual fidelity and prompt adherence. However, this strict adherence comes at th...

AI 聚合
2026-06-25
HuggingFace

Multi4D: High-Fidelity Dynamic Gaussian Splatting via Multi-Level Competitive Allocation

Dynamic 3D Gaussian splatting faces a fundamental tension between motion consistency and visual fidelity. Deformation...

AI 聚合
2026-06-25
HuggingFace

InSight: Self-Guided Skill Acquisition via Steerable VLAs

Vision-language-action (VLA) models can learn manipulation skills from demonstrations, but their capabilities are bou...

AI 聚合
2026-06-25
HuggingFace

Critique of Agent Model

What is an agent? What constitutes agency? With the rise of Large Language Model (LLM) systems marketed as ``coding a...

AI 聚合
2026-06-25
HuggingFace

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct

Scaling reinforcement learning for visual mathematical reasoning requires more than generating harder questions: as d...

AI 聚合
2026-06-25
HuggingFace

MEMPROBE: Probing Long-Term Agent Memory via Hidden User-State Recovery

Long-term memory promises LLM agents that grow more capable across sessions, maintaining an accurate, evolving unders...

AI 聚合
2026-06-25
首页 上一页 第 177 / 212 页 下一页 末页