HuggingFace

Revisiting Local Context for Long-Horizon Streaming 3D Reconstruction

Streaming 3D reconstruction from extremely long videos requires estimating camera motion and scene geometry online un...

AI 聚合
2026-08-31
HuggingFace

J-Zero: Unified Challenger--Solver--Judge Co-Evolution from Zero Data

Self-evolving language models have recently emerged as a promising path toward superintelligence, with the advantage ...

AI 聚合
2026-08-31
HuggingFace

StarHarness: Evolving Harnesses with Stratified Search for Enterprise Environments

We present StarHarness, a framework for evolving environment-specific agent harnesses while keeping model weights fix...

AI 聚合
2026-08-31
HuggingFace

Paint What You See: Benchmarking Dexterous Visual Tool Use in Multimodal Agents

Evaluation is shifting from static QA toward agentic settings where models act through external tools. We identify a ...

AI 聚合
2026-08-31
HuggingFace

Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090

Language model pretraining has become almost synonymous with prohibitive cost, placing it out of reach for much of th...

AI 聚合
2026-08-31
HuggingFace

LayerRecall: A State-Conditioned Memory Router for Long-Horizon Consistency in Video Generation

Autoregressive video diffusion enables scalable long-video generation by producing chunks from a bounded recent conte...

AI 聚合
2026-08-31
HuggingFace

Ring Forcing: Towards Precise Long-Term Memory for Autoregressive Video Diffusion

Scaling video generation to long durations reveals a critical bottleneck: current models lack robust long-term memory...

AI 聚合
2026-08-31
HuggingFace

Fast Weight Attention for Continual Learning

Recurrent fast-weight memories and selective state-space models compress an expanding context into a fixed-size recur...

AI 聚合
2026-08-31
HuggingFace

Locate Anything in Videos: Rethinking Efficient Generative Spatio-Temporal Video Grounding

Spatio-temporal video grounding (STVG) requires models to identify when a referred event occurs and localize the targ...

AI 聚合
2026-08-31
HuggingFace

DART-SD: Diamond-topology Aware Retrieval and Tuning for Self-Distillation of Multi-Turn Tool-Calling Agents

Equipping Large Language Models (LLMs) with multi-turn tool-calling capabilities is essential for building autonomous...

AI 聚合
2026-08-31
HuggingFace

Agentic Artifact Creation: Systems, Evaluation, Principles, and Opportunities

Generative models can turn natural-language prompts into images, text, code, and other content, lowering the cost of ...

AI 聚合
2026-08-31
HuggingFace

Language Chain in Alignment: Cross-lingual Ranking Preference Optimization

The alignment of Large Language Models heavily relies on English-centric high-quality preference data, which often le...

AI 聚合
2026-08-31
首页 上一页 第 18 / 212 页 下一页 末页