arXiv

A Mathematical Theory of Reusable Neural Bases for Network Compression

As large AI models become increasingly prevalent across a wide range of applications, memory cost has become a critic...

AI 聚合
2026-09-02
arXiv

Can LLMs Discover Scientific Laws in Real and Parallel Worlds?

Scientific equation discovery has long been central to scientific progress, proceeding through iterative cycles of hy...

AI 聚合
2026-09-02
arXiv

BS: Take the Hint - Interactive Multitracer PET/CT Lesion Segmentation with a Scribble-Conditioned ResEnc U-Net

Automated lesion segmentation in whole-body PET/CT is complicated by the variety of physiological tracer uptake patte...

AI 聚合
2026-09-02
arXiv

Retrieved but not ranked: surface-form bias in structural retrieval, from mathematics to agent trajectories

We evaluate embedding retrieval where surface form and meaning are pulled apart on purpose: retrieving items that sha...

AI 聚合
2026-09-02
arXiv

H3-World: Turning Language Understanding into World Control

We present H3-World, an efficient framework that turns the 33B MiniMax-H3 video generator into an interactive world m...

AI 聚合
2026-09-02
arXiv

From Confusion to Clarity: Confusion-Aware Retrieval and Knowledge Injection for Text Classification

Large language models (LLMs) struggle to classify text into taxonomies with many semantically similar labels, as the ...

AI 聚合
2026-09-02
arXiv

Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers

Vision-Language Models (VLMs) provide useful priors for interactive decision-making, but using them directly as polic...

AI 聚合
2026-09-02
arXiv

Scaling Near-Optimal SFT-RL Annotation Budget Allocation from Small to Large LLMs

How to divide a fixed annotation budget between supervised fine-tuning (SFT) and reinforcement learning (RL) during L...

AI 聚合
2026-09-02
arXiv

Designing Proactive Thought Partners for Writing

Writing involves diverse cognitive activities, from ideation to revision, and writers' needs vary across individuals ...

AI 聚合
2026-09-02
arXiv

Mechanism Design for Alignment and Control

We develop a framework for mechanism design with AI agents whose alignment (preferences) and capabilities (feasible a...

AI 聚合
2026-09-02
arXiv

The Rise of Verbal Reinforcement Learning

Natural language is emerging as a primary feedback channel for improving language agents, capable of conveying intent...

AI 聚合
2026-09-02
arXiv

CordisBench: Can Language Models Reason About Component Lifecycles in Dynamic Agent Harnesses?

Dynamic agent harnesses let language models change the software that shapes their own execution. This flexibility bri...

AI 聚合
2026-09-02
首页 上一页 第 9 / 212 页 下一页 末页