HuggingFace

Vision Pretraining for Dense Spatial Perception

Dense spatial perception is essential for physical intelligence, where visual systems are expected to recover structu...

AI 聚合
2026-07-07
HuggingFace

UI-MOPD: Multi-Platform On-Policy Distillation for Continual GUI Agent Learning

Recent advances in multimodal foundation models and agent systems have driven GUI agents from single-platform task ex...

AI 聚合
2026-07-07
HuggingFace

PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space

3D reconstruction and generation are commonly tackled by separate paradigms: pixel-based regression for reconstructio...

AI 聚合
2026-07-07
HuggingFace

Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models

Predicting object dynamics (i.e., world modeling) is a fundamental challenge for robotic manipulation, and modeling d...

AI 聚合
2026-07-07
HuggingFace

PixCon: Clean-Positive Contrastive Learning for Foundation-Model Semi-Supervised Segmentation

Semi-supervised semantic segmentation (SSSS) has long turned on one question, which pseudo-labels to trust, and answe...

AI 聚合
2026-07-07
HuggingFace

Multiplayer Interactive World Models with Representation Autoencoders

We introduce the first multiplayer world model for highly dynamic environments governed by complex physical interacti...

AI 聚合
2026-07-07
HuggingFace

LLM-as-a-Verifier: A General-Purpose Verification Framework

Scaling pre-training, post-training, and test-time compute have become the central paradigms for improving the capabi...

AI 聚合
2026-07-07
HuggingFace

EVA-Client: A Unified Data Collection, Inference, and Deployment Framework for Embodied Policies on Real Robots

We present EVA-Client, an open-source framework for deployment, data collection, and evaluation of trained manipulati...

AI 聚合
2026-07-07
HuggingFace

EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments

Pretraining scaling laws reveal that model capability improves predictably with data and compute. But learning from r...

AI 聚合
2026-07-07
HuggingFace

InternVLA-A1.5: Unifying Understanding, Latent Foresight, and Action for Compositional Generalization

Unified models for robot manipulation aim to equip one policy with both the semantic priors of pretrained VLMs and th...

AI 聚合
2026-07-07
HuggingFace

GORGO: Online Tuning for Cross-Region Network-Aware LLM Serving

Increasingly, LLM inference services proxy client requests to engine replicas distributed globally. Load-balancing po...

AI 聚合
2026-07-07
HuggingFace

dOPSD: On-Policy Self-Distillation for Diffusion Language Models

Diffusion large language models (dLLMs) generate text by iteratively denoising a masked sequence, offering a parallel...

AI 聚合
2026-07-07
首页 上一页 第 151 / 215 页 下一页 末页