HuggingFace

CoinVE-200K: A Large-Scale High-Quality Dataset for Compositional Instruction-Guided Video Editing

The quality and diversity of instruction-based video editing datasets are steadily improving, yet existing datasets m...

AI 聚合
2026-08-19
HuggingFace

EDITBRIDGE: Towards Faithful and Efficient Ultra-High-Resolution Image Editing

High-resolution image editing is increasingly demanded in professional workflows, yet existing diffusion-based models...

AI 聚合
2026-08-19
HuggingFace

Cross-Model Memory Transfer via Target-Side Reader Adaptation

Methods for improving knowledge use in large language models typically fall into two regimes. Non-parametric retrieva...

AI 聚合
2026-08-19
HuggingFace

Demystifying Agent Skills: Why They Work-Until They Don't

Skills have emerged as a practical and effective approach for enhancing LLM agents at inference time through structur...

AI 聚合
2026-08-19
HuggingFace

MoE-ViE: Mixture of Experts Vision Encoder for Efficient Image and Video Understanding

Vision encoders are a critical component of vision-language models, and scaling their capacity effectively improves p...

AI 聚合
2026-08-19
HuggingFace

PixRestore: Unified Image Restoration via Pixel Diffusion Transformer

Unified image restoration (UIR) aims to recover high-quality (HQ) content from low-quality (LQ) images with different...

AI 聚合
2026-08-19
HuggingFace

DiSCO: Defending text-to-image generation through distribution-guided contrastive prompt optimization

As text-to-image generative models advance, they raise critical safety concerns, particularly the generation of Not-S...

AI 聚合
2026-08-19
HuggingFace

V-RAE: Rethinking Video Latent Spaces for Generation

Latent video generation relies on autoencoders to define a compact space in which generative models operate. Although...

AI 聚合
2026-08-19
HuggingFace

Advancing Open and Reproducible Relational Learning: RelArena-α, TabPFN-Rel and RPI

This first release of Prior Labs in relational learning shows our continued commitment to open science. We open-sourc...

AI 聚合
2026-08-19
HuggingFace

Plausible but Not Valid: A Psychometric Audit of LLMs as Synthetic Survey Respondents

Large language models (LLMs) are increasingly used as synthetic survey respondents, but existing evaluations ask whet...

AI 聚合
2026-08-19
HuggingFace

Valid Per-Field Selective Risk Control for Document Extraction: Three Failure Modes, a Validity Ladder, and When Conditioning Pays

Per-field accept/review with selective risk at most alpha -- accept a field only if the error rate among accepted fie...

AI 聚合
2026-08-19
HuggingFace

StreamOPD: A Post-Training Recipe with Spatio-Temporal Cue Gating for Streaming Video Understanding

Streaming video understanding demands direct responses from the causally observed prefix of an unfolding video. Exist...

AI 聚合
2026-08-19
首页 上一页 第 45 / 212 页 下一页 末页