HuggingFace

PointDiT: Pixel-Space Diffusion for Monocular Geometry Estimation

State-of-the-art single-image 3D reconstruction methods often rely on complex hybrid architectures and loss functions...

AI 聚合
2026-07-08
HuggingFace

When Classic Cache Policies Fail: Learning-Augmented Replacement for Semantic Retrieval Buffers

LLM agents increasingly rely on retrieval buffers to store and reuse past experience, yet the cache management polici...

AI 聚合
2026-07-08
HuggingFace

Gemma 4 Technical Report

We introduce Gemma 4, a new generation of open-weight, natively multimodal language models in the Gemma model family....

AI 聚合
2026-07-08
HuggingFace

PluraMath: Extending Mathematical Reasoning Evaluation Beyond High-Resource Languages

Mathematical reasoning has become a central task for evaluating and tuning reasoning Large Language Models (LLMs), ye...

AI 聚合
2026-07-08
HuggingFace

Image2Sim: Scaling Embodied Navigation via Generative Neural Simulator

Embodied navigation aims to build agents that interpret multimodal goals, reason in 3D space, and reach target destin...

AI 聚合
2026-07-08
HuggingFace

From Foundation to Application: Improving VLA Models in Practice

Despite recent progress of VLA foundation models, the disparity between laboratory conditions and real-world applicat...

AI 聚合
2026-07-08
HuggingFace

Bibby AI: An Editor-Native Agentic Platform for Academic Research, Writing, and Publishing

Academic output is produced across a fragmented toolchain: literature discovery in one application, reference managem...

AI 聚合
2026-07-08
HuggingFace

AI Wizards at EXIST 2026: Hierarchical Soft-Label Learning for Multimodal Sexism Identification in Memes

We present the AI Wizards submission to EXIST 2026 for multimodal sexism identification in memes. The task is compose...

AI 聚合
2026-07-08
HuggingFace

MANCE: Manifold Aware Concept Erasure

Concept erasure aims to remove a target concept from a representation while preserving the other information encoded ...

AI 聚合
2026-07-08
HuggingFace

Transition-Aware best-of-N sampling for Longitudinal Chest X-ray Reports

In longitudinal clinical practice, every chest X-ray is read in the context of the patients prior exam, and much of w...

AI 聚合
2026-07-08
HuggingFace

Taste-aware music retrieval from audio embeddings

Crossmodal correspondences between sound and taste are well established in psychology and neuroscience, but largely a...

AI 聚合
2026-07-08
HuggingFace

Unified Audio Intelligence Without Regressing on Text Intelligence

Audio intelligence involves understanding, reasoning about, and generating both audio and speech. In this work, we in...

AI 聚合
2026-07-08
首页 上一页 第 148 / 215 页 下一页 末页