HuggingFace

SynCity 3000: Bootstrapping Scene-Scale 3D Diffusion

We present SynCity 3000, a framework for generating 3D scenes that are globally coherent while enabling fine-grained ...

AI 聚合
2026-07-08
HuggingFace

Speaker-Aware Temporal Aggregation Strategies on Segment Representations for Depression Detection in Dyadic Interaction: A Benchmark Study

Speech-based depression detection compresses features from short audio segments into one speaker-level decision, a st...

AI 聚合
2026-07-08
HuggingFace

Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process

Unified multi-modal models (UMMs) have shown promising interleaved text-image reasoning capabilities, yet effectively...

AI 聚合
2026-07-08
HuggingFace

Look Before You Leap: Distilling Tree Search into Action Evaluation for Frozen VLA Models

Vision-Language-Action (VLA) models acquire broad embodied capabilities through large-scale pretraining, yet their ge...

AI 聚合
2026-07-08
HuggingFace

Learning to Trigger: Reinforcement Learning at the Large Hadron Collider

High-throughput scientific facilities such as the Large Hadron Collider depend on real-time event filtering (triggeri...

AI 聚合
2026-07-08
HuggingFace

SeKV: Resolution-Adaptive KV Cache with Hierarchical Semantic Memory for Long-Context LLM Inference

Large language models increasingly operate over long contexts, where the KV cache becomes a dominant memory bottlenec...

AI 聚合
2026-07-08
HuggingFace

ACID: Action Consistency via Inverse Dynamics for Planning with World Models

Decision-time planning with action-conditioned world models has become a popular paradigm for embodied control. Howev...

AI 聚合
2026-07-08
HuggingFace

GaP: A Graph-as-Policy Multi-Agent Self-Learning Harness For Variational Automation Tasks

For robots to work reliably in commercial and industrial applications, can recent advances in agentic coding systems ...

AI 聚合
2026-07-08
arXiv

OptiAgent: End-to-End Optimization Modeling via Multi-Agent Iterative Refinement

We propose OptiAgent, a multi-agent framework that, given a natural language description of an Operations Research pr...

AI 聚合
2026-07-07
arXiv

Multiplayer Interactive World Models with Representation Autoencoders

We introduce the first multiplayer world model for highly dynamic environments governed by complex physical interacti...

AI 聚合
2026-07-07
arXiv

Selective Disclosure Watermarking for Large Language Models

Watermarking methods embed imperceptible and verifiable signals into text generated by large language models (LLMs). ...

AI 聚合
2026-07-07
arXiv

Graph Sparse Sampling: Breaking the Curse of the Horizon in Continuous MDP Planning

Planning under uncertainty in continuous domains is essential for autonomous systems, yet computationally demanding. ...

AI 聚合
2026-07-07
首页 上一页 第 149 / 215 页 下一页 末页