TuneJury: An Open Metric for Improving Music Generation Preference Alignment
We introduce TuneJury, an open, instance-level pairwise reward model for text-to-music that predicts a music preferen...
每天自动聚合 AI 领域最新动态
We introduce TuneJury, an open, instance-level pairwise reward model for text-to-music that predicts a music preferen...
As LLM agents are deployed in long-horizon sessions, context accumulation drives up inference costs. Existing approac...
Remote sensing vision-language models have advanced Earth observation understanding, but most existing work remains c...
Simple linear and frequency-domain models remain surprisingly competitive in long-horizon time-series forecasting, an...
Oppenheim and Lim (1981) showed that natural images stay recognizable when reconstructed from their Fourier phase alo...
Standard accuracy benchmarks are designed to test how closely large language models (LLMs) approach correct answers, ...
A content-moderation system can score well on every standard accuracy metric and still cause real harm, if its mistak...
DreamX-World 1.0 is a general-purpose interactive text/image-to-video world model for controllable long-horizon gener...
Diffusion transformers have demonstrated remarkable generative capabilities, yet the rich perceptual representations ...
As LLMs advance, post-training reinforcement learning (RL) increasingly relies on multi-dimensional rewards to cultiv...
In this paper, we introduce SP^3, a novel Plug-and-Play algorithm that accelerates maximum a posteriori image restora...
Long-form video generation requires recurring subjects to remain consistent across various shots, viewpoints, motions...