OPSD-V: On-Policy Self-Distillation for Post-Training Few-Step Autoregressive Video Generators
We propose OPSD-V, an on-policy self-distillation paradigm for post-training few-step autoregressive (AR) video diffu...
每天自动聚合 AI 领域最新动态
We propose OPSD-V, an on-policy self-distillation paradigm for post-training few-step autoregressive (AR) video diffu...
Reasoning has become a core capability for large models, especially when reliable decisions require understanding log...
Scientific ideas rarely start from a blank page. They inherit mechanisms, repair known limitations, and recombine pie...
Modern LLMs are increasingly deployed in long-context applications such as retrieval-augmented generation, repository...
Generating realistic 3D human motions in real-time within interactive applications is key for animation, simulation, ...
Inference-time scaling for text-to-image generation has progressed from simple Best-of-N (BoN) sampling to guided sea...
We introduce Vidu S1, a real-time interactive video generation model supporting voice control of digital characters. ...
Reinforcement learning (RL) has become the standard paradigm for enhancing the complex reasoning capabilities of larg...
Large language models (LLMs) increasingly act as integrated data-science agents, combining abstract reasoning with ad...
Magnetic resonance imaging (MRI) super-resolution is vital for improving diagnostic accessibility, yet most methods t...
The growing demand for image-to-video creation on mobile devices has increasingly focused on cinematic motion effects...
A key challenge in Arabic NLP is the scarcity of dialectal data relative to Modern Standard Arabic (MSA), causing LLM...