HuggingFace

When Agents Commit Too Soon: Diagnosing Premature Commitment in LLM Agents

Long-horizon LLM agents can fail quietly: they settle on one reading of the evidence early, then spend the rest of th...

AI 聚合
2026-06-24
HuggingFace

ShotcreteDepth: A Bi-modal Dataset for Robust Robotic Depth Perception in Shotcrete Construction Environments

We introduce ShotcreteDepth, a bi-modal dataset from the construction domain that captures both an active shotcreting...

AI 聚合
2026-06-24
HuggingFace

Comparing Linear Probes with Mahalanobis Cosine Similarity

Linear probes are widely used in interpretability research and often compared by cosine similarity. The Mahalanobis c...

AI 聚合
2026-06-24
HuggingFace

Lift4D: Harmonizing Single-View 3D Estimation for 4D Reconstruction In-the-Wild

Reconstructing dynamic non-rigid objects from monocular video requires integrating visual cues from direct observatio...

AI 聚合
2026-06-24
HuggingFace

TROPT: An Open Framework for Unifying and Advancing Discrete Text Optimization

Discrete text-trigger optimization -- searching for text sequences that, when ingested by a model, steer it toward a ...

AI 聚合
2026-06-24
HuggingFace

Beyond Reward Engineering: A Data Recipe for Long-Context Reinforcement Learning

Long-context reasoning is an essential capability for large language models, particularly when they are deployed as a...

AI 聚合
2026-06-24
arXiv

Discovering Latent Groups for Robust Classification

Machine learning models exploit spurious correlations, achieving high average accuracy but failing disproportionately...

AI 聚合
2026-06-23
arXiv

Data Selection Through Iterative Self-Filtering for Vision-Language Settings

The availability of large amounts of clean data is paramount to training neural networks. However, at large scales, m...

AI 聚合
2026-06-23
arXiv

RECALL: Recovery Experience Collection for Active Lifelong Learning in Vision-Language-Action Models

Vision-Language-Action (VLA) models are commonly fine-tuned through passive imitation learning, where additional demo...

AI 聚合
2026-06-23
arXiv

DiT-Reward: Generative Representations for Text-to-Image Reward Modeling

Can representations learned for image generation also support the evaluation of generated images? We study text-to-im...

AI 聚合
2026-06-23
arXiv

AI-driven Optimisation of Quality of Recovery (QoR) in Remote Patient Monitoring

Remote patient monitoring depends on patient-reported data to capture the subjective dimension of recovery that devic...

AI 聚合
2026-06-23
arXiv

AI Exposure Scores: what they measure, what they miss, and what comes next

A set of exposure scores calculated in 2023 has become a central empirical input to the future of work debate. Produc...

AI 聚合
2026-06-23
首页 上一页 第 182 / 212 页 下一页 末页