HuggingFace

Push-Wiper: Toward General-Purpose Robotic Cleaning across Varied Stains and Surfaces with Segmented Pushing Trajectories

Viscous stains, characterized by high viscosity and complex rheological properties, remain a major challenge for robo...

AI 聚合
2026-08-05
HuggingFace

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent

We introduce Video-DeepResearch (Video-DR), extending multimodal agents from static images to continuous video stream...

AI 聚合
2026-08-05
HuggingFace

When Agents Learn to Be You: Benchmarking Privacy Leakage, Impersonation Risk, and Defenses in Persona Skills

Persona skills distill personal interaction histories into portable and executable artifacts for downstream agents. W...

AI 聚合
2026-08-05
HuggingFace

LLaDA MoE v2: Scaling Mixture-of-Experts Diffusion Language Models

Diffusion language models (dLLMs) offer an alternative to autoregressive (AR) language modeling, yet the scaling beha...

AI 聚合
2026-08-05
HuggingFace

PosterMELD: Multi-Agent Paper-to-Poster Generation for Controllable Design Diversity with Editable Print-Ready Outputs

Scientific poster construction compresses a long multimodal paper into a readable, editable canvas. Existing systems ...

AI 聚合
2026-08-05
HuggingFace

ContinualSkillBench: Can LLM Agents Truly Evolve Their Capabilities?

Modern agent frameworks equip large language models with external skill libraries to solve complex tasks. However, it...

AI 聚合
2026-08-05
HuggingFace

MiniWorld: Democratizing the Training of Video World Models from Scratch

Video world models predict future observations conditioned on historical observations and control signals, enabling l...

AI 聚合
2026-08-05
HuggingFace

Decoding Children's Gait Behavior

We introduce a new problem domain for human action recognition: the fine-grained analysis of children's gait behavior...

AI 聚合
2026-08-05
HuggingFace

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures

Existing evaluations often reduce agent failures to system-level outcomes, obscuring where the fault originated and w...

AI 聚合
2026-08-05
HuggingFace

A Frozen Pixel-Space Diffusion Model Can Guide Itself with Its Own Samples

Pixel-space diffusion models aim to learn an end-to-end generator directly over raw pixels. This is challenging becau...

AI 聚合
2026-08-05
HuggingFace

GEOID-Flood: A Large-Scale Multi-Modal Benchmark Dataset for Flood Segmentation

Geospatial foundation models aim to learn representations that transfer across regions and sensors, yet evaluating th...

AI 聚合
2026-08-05
HuggingFace

ICDAR 2026 Competition on Information Extraction from Atomic Layer Deposition/Etching (ALD/E) Scientific Figures

Scientific figure comprehension and reasoning using multimodal AI requires integrating visual perception with domain-...

AI 聚合
2026-08-05
首页 上一页 第 84 / 215 页 下一页 末页