HuggingFace

Compute Globally, Materialize Locally: The Memory Contract of Sparse Event-KV

Long-horizon agents increasingly reuse their KV cache as memory: a serving system keeps a subset of cached entries an...

AI 聚合
2026-08-05
HuggingFace

Seeing or Knowing? Visual Context Sensitivity in Multimodal Large Language Models

Multimodal Large Language Models (MLLMs) achieve strong performance by integrating visual inputs with the rich priors...

AI 聚合
2026-08-05
HuggingFace

Loud or Silent? A Reusable Framework for Per-Modality Failure Analysis in Multimodal Clinical AI

Multimodal clinical models are usually judged on accuracy with every modality present, but deployment removes modalit...

AI 聚合
2026-08-05
HuggingFace

Zero-Mem: Zero-Token Memory Operations for LLM Agents

LLM agents need memory to act consistently over long interactions, yet many systems use additional LLM calls to opera...

AI 聚合
2026-08-05
HuggingFace

SG-WAM: Self-Guided World Modeling in Geometry-Aware Policy Space

World Action Models (WAMs) couple action generation with prediction of future states. Their effectiveness depends on ...

AI 聚合
2026-08-05
HuggingFace

To Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing

Large language models increasingly write and repair production code, yet evidence is mounting that their test-passing...

AI 聚合
2026-08-05
HuggingFace

Wnuan: Staged Post-Training for Question Answering over Proprietary Enterprise Knowledge

Enterprise question answering requires models to acquire proprietary knowledge without discarding general capabilitie...

AI 聚合
2026-08-05
HuggingFace

MemSFT: Mitigating Alignment Tax with an External Parametric Memory

Adapting Large Language Models (LLMs) to specialized domains often incurs an alignment tax, as fine-tuning on domain-...

AI 聚合
2026-08-05
HuggingFace

AgentStream: How Well Do Self-Evolving LLM Agents Perform Under Streaming Tasks?

Large language model (LLM) agents can self-evolve by continually improving from their own accumulated experience. How...

AI 聚合
2026-08-05
arXiv

Long-term Measurements: Towards a Longitudinal Understanding of Human-AI Interactions

Language models have taken on the role of a very new type of technology, by virtue of their "human-ness" and rapid in...

AI 聚合
2026-08-04
arXiv

DyFrDet: Towards Accurate Small Object Detection via Dynamic Frequency Suppression with Label Disambiguation

Despite the remarkable progress over the past decades, accurately identifying small objects remains challenging becau...

AI 聚合
2026-08-04
arXiv

SWE-Touch: Benchmarking Coding Agents When Users Touch the Code

Real-world software development requires coding agents to operate in shared workspaces where users may inspect and mo...

AI 聚合
2026-08-04
首页 上一页 第 85 / 215 页 下一页 末页