HuggingFace

HiLo-Token: Input-Adaptive High-Low Frequency Token Compression for Efficient Image Editing

Creative image editing tools, such as Photoshop's Remove or Generative Fill buttons, are central to everyday customer...

AI 聚合
2026-06-19
HuggingFace

When Does Trajectory-Level Supervision Permit Efficient Offline Reinforcement Learning?

Offline reinforcement learning is typically analyzed under process-level reward supervision, yet many sequential deci...

AI 聚合
2026-06-19
HuggingFace

Reinforcement Learning-Guided Retrieval with Soft Fusion for Robust Multimodal Imitation Learning under Missing Modalities

Robotic systems perceive the world through multiple input modalities -- including visual camera streams and natural l...

AI 聚合
2026-06-19
HuggingFace

Re-Centering Humans in LLM Personalization

Despite growing interest, most evaluations of large language models' (LLMs') personalization abilities have relied on...

AI 聚合
2026-06-19
HuggingFace

REVES: REvision and VErification--Augmented Training for Test-Time Scaling

Test-time scaling via sequential revision has emerged as a powerful paradigm for enhancing Large Language Model (LLM)...

AI 聚合
2026-06-19
arXiv

Mechanism-Guided Selective Unlearning for RLVR-Induced Reasoning

We propose MAST (Mechanism-Aligned Selective Targeting), a mechanism-guided method for unlearning RLVR-induced reason...

AI 聚合
2026-06-18
arXiv

STARE: Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability

Reinforcement Learning with Verifiable Rewards algorithms like GRPO have emerged as the dominant post-training paradi...

AI 聚合
2026-06-18
arXiv

TxBench-PP: Analyzing AI Agent Performance on Small-Molecule Preclinical Pharmacology

Artificial intelligence (AI) agents promise to accelerate drug discovery by compressing interpretation and decision-m...

AI 聚合
2026-06-18
arXiv

A Taxonomy of Mental Health and Technology Needs for Alzheimer's and Dementia Caregivers

Family members caring for individuals with Alzheimer's disease and related dementias (AD/ADRD) provide the foundation...

AI 聚合
2026-06-18
arXiv

OneCanvas: 3D Scene Understanding via Panoramic Reprojection

Existing approaches to 3D scene understanding in Vision-Language Models (VLMs) either rely on complex, model-specific...

AI 聚合
2026-06-18
arXiv

X+Slides: Benchmarking Audience-Conditioned Slide Generation

Automatically generating slide decks from source documents is an important application of large language models (LLMs...

AI 聚合
2026-06-18
arXiv

A Multi-Domain Benchmark for Detecting AI-Generated Text-Rich Images from GPT-Image-2

Text-rich images often contain privacy-sensitive, transactional, or decision-relevant information. As recent multimod...

AI 聚合
2026-06-18
首页 上一页 第 191 / 212 页 下一页 末页