HuggingFace

MPIE-Bench: Benchmarking Anatomically Plausible Multi-Person Interaction Editing

Text-to-image and personalized editing models now synthesize high-fidelity single-subject images with ease. Yet placi...

AI 聚合
2026-07-31
HuggingFace

Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation

Role-playing agents (RPAs) have become one of the most important consumer applications of large language models. User...

AI 聚合
2026-07-31
HuggingFace

AI Tour Meeting: Group Travel Planning by LLM Agents

This paper proposes AI Tour Meeting, a group travel planning framework powered by multiple Large Language Model (LLM)...

AI 聚合
2026-07-31
HuggingFace

Revisiting Lossy Verification in Speculative Decoding: Mechanisms, Trade-offs, and Failure Modes

Speculative Decoding (SD) accelerates large language model inference by allowing a lightweight draft model to propose...

AI 聚合
2026-07-31
HuggingFace

Explorative Modeling: Unlocking a Third Pretraining Axis and End-to-End Generation

The deep learning revolution, kicked off by AlexNet, taught us that end-to-end training beats decomposing a problem i...

AI 聚合
2026-07-31
HuggingFace

LEDGERMIND: Provenance-Constrained Multimodal Agentic Reasoning with a Structured Evidence Ledger

Multimodal agents for visual question answering increasingly operate as multi-step trajectories that interleave perce...

AI 聚合
2026-07-31
HuggingFace

INTACT: Isomorphic Intent-to-Action Learning for Search-Free World Models

Forward latent world models predict how actions change a scene, but recover actions for a desired change only through...

AI 聚合
2026-07-31
HuggingFace

Multi-Head Attention Residuals

Transformers propagate information across depth through a single additive residual stream: every sublayer reads only ...

AI 聚合
2026-07-31
HuggingFace

AMRD: Adaptive Multi-Teacher Relational Distillation for Lightweight Speech Emotion Recognition

On-device speech emotion recognition (SER) is critical for real-time applications, yet large self-supervised models t...

AI 聚合
2026-07-31
HuggingFace

ShadowDancer: Teaching Video World Models Any Action by Learning Unified Dynamics Representations from a Video and Its Shadow

We present ShadowDancer, a novel approach to any-action, frame-level control of interactive video world models. The o...

AI 聚合
2026-07-31
HuggingFace

Is Deep Research Reliable? Misleading Knowledge Induces False Conclusions

Deep Research agents extend LLM-based assistants into long-horizon workflows involving planning, retrieval, evidence ...

AI 聚合
2026-07-31
HuggingFace

Pedestrian Archetypes Extension -- More Pedestrian Models for Autonomous Vehicle Safety Testing

In our prior work, Pedestrian Archetypes, we defined pedestrian archetypes as collections of behaviors that uniquely ...

AI 聚合
2026-07-31
首页 上一页 第 94 / 215 页 下一页 末页