HuggingFace

Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation

MLLM-based segmentation faces a core segmentation trilemma: high segmentation performance, preserved dialogue ability...

AI 聚合
2026-08-06
HuggingFace

ChronoLens: Measuring Language Change Across Time, Languages, and Linguistic Levels

Historical language change affects morphology, syntax, semantics, and pragmatics, yet computational studies typically...

AI 聚合
2026-08-06
HuggingFace

RestoreKV: Recovering Full-Cache Behavior Under Aggressive Query-Agnostic KV Cache Eviction

Query-agnostic KV cache eviction compresses a context once and reuses the resulting cache for arbitrary future querie...

AI 聚合
2026-08-06
HuggingFace

CURV: Enhancing Chart Understanding Through Curriculum Visual Grounded Reasoning

Chart question answering (CQA) requires multimodal large language models (MLLMs) to integrate visual comprehension wi...

AI 聚合
2026-08-06
HuggingFace

ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning

On-policy training has emerged as a powerful post-training paradigm for improving the reasoning capabilities of large...

AI 聚合
2026-08-06
HuggingFace

LegalPincite: Multi-level Legal Information Retrieval Dataset

A common task in legal Information Retrieval (IR) is to find relevant legal sources from case-law collections. While ...

AI 聚合
2026-08-06
HuggingFace

ARCHead: Activation-Metric Residual Correction for Large Language Model Output Heads

Weight-only quantization substantially reduces the storage of large language model (LLM) transformer blocks, but prac...

AI 聚合
2026-08-06
HuggingFace

When Many Answers Are Valid, Voting Fails: Symbolic Verification for Best-of-K Causal Reasoning in LLMs

Self-consistency assumes the most frequent answer among sampled reasoning traces is the most reliable, but this can f...

AI 聚合
2026-08-06
HuggingFace

Know When to Stop: Segment-Level Credit Assignment for Reducing Overthinking

Reasoning language models frequently overthink: generating extended chains of behaviors such as hedging, approach aba...

AI 聚合
2026-08-06
arXiv

When and Where to Look: Adaptive Visual Evidence Scheduling for Efficient Long Video Understanding

Efficient long-video understanding requires vision--language models (VLMs) to reason over a small number of frames se...

AI 聚合
2026-08-05
arXiv

Equivariant Music Transformer

Humans recognize a musical passage even when it is shifted in time or transposed in pitch, indicating a notion of equ...

AI 聚合
2026-08-05
arXiv

The Transformer Revolution, Part 1: Dynamic Processing through Output- Weight Interconnections

This paper offers a new interpretation of the Transformer during inference. Against the "stochastic parrot" view that...

AI 聚合
2026-08-05
首页 上一页 第 81 / 215 页 下一页 末页