HuggingFace

Listening Forward: Next Patch Embedding Prediction Enables Scalable Audio Learners

Self-supervised learning (SSL) has driven substantial progress in audio representation learning, though existing meth...

AI 聚合
2026-08-21
HuggingFace

NARU: A Benchmark for NARrative Evolution and Cultural Nuance Understanding in Japanese Extreme Long Video

Long-form video understanding encompasses tasks that go beyond retrieving isolated events, including tracking an evol...

AI 聚合
2026-08-21
HuggingFace

Bounded Agents: Delegation Security for Multi-Agent AI Systems

LLM-based agents can act on behalf of a user to access cloud services, call tools, or invoke agents. At session start...

AI 聚合
2026-08-21
HuggingFace

LLMs Get Smarter from Targeted Synthetic Multilingual Data

Language-specific competency (LSC) is the phenomenon of a language model performing better or worse depending on the ...

AI 聚合
2026-08-21
HuggingFace

SPK: Eliciting Structured Prior Knowledge for Interpretable Out-of-Distribution Detection in Real-Time Object Detection

Object detectors often produce over-confident predictions for objects outside their training categories, leading to s...

AI 聚合
2026-08-21
HuggingFace

Towards Real-Time and Adaptable LiDAR Scene Completion

LiDAR scene completion is a key component of 3D perception in autonomous driving, where the scene must be completed i...

AI 聚合
2026-08-21
HuggingFace

Evaluating Music Context Preservation: A Multi-facet Framework for Music Editing Systems

Music editing plays a vital role in modern music production, with applications in film, broadcasting, and game develo...

AI 聚合
2026-08-21
HuggingFace

VA-Judger: Reward Modeling from Human Preference Feedback for Joint Video-Audio Generation

Using reinforcement learning to post-train joint video-audio generation models requires a reward signal. Existing met...

AI 聚合
2026-08-21
arXiv

DA-WAM: Decision-Aligned Future Latents for Driving World Models

Anticipating how scenes evolve under ego actions is fundamental to safe autonomous driving, yet the full potential of...

AI 聚合
2026-08-20
arXiv

Detecting Backdoors in Object Detection via Pre-NMS Prediction Distribution Shift

Object detection models deployed in safety-critical applications remain vulnerable to backdoor attacks that cause tar...

AI 聚合
2026-08-20
arXiv

Open-MOPD: Diagnosing and Fixing Capability Imbalance in Multi-Teacher On-Policy Distillation

Multi-teacher on-policy distillation (M-OPD) has emerged as a promising paradigm for consolidating domain-specialized...

AI 聚合
2026-08-20
arXiv

Discretizing Continuous Time Series for Imputation with Masked Diffusion Training

Time series imputation is a crucial area for reliable time series analysis, yet it remains challenging due to the com...

AI 聚合
2026-08-20
首页 上一页 第 39 / 212 页 下一页 末页