WARP: Weight-Space Analysis for Recovering Training Data Portfolios
Foundation models are routinely released to the public, yet the data recipes used to train them -- such as domain mix...
每天自动聚合 AI 领域最新动态
Foundation models are routinely released to the public, yet the data recipes used to train them -- such as domain mix...
Traffic matrices (TMs) capture network-wide origin-destination demand and are central to traffic engineering, yet acc...
Grid-based approaches to approximate nearest neighbor (ANN) search have been absent from modern scaling analyses. We ...
Diffusion transformers (DiTs) achieve state-of-the-art image and video generation, but their multi-step sampling and ...
Vision-Language-Action (VLA) models are fundamentally bottlenecked by the scarcity of expert demonstrations -- triple...
Whether pairing people with AI helps or hurts is usually reported as a single average effect. Using a real-money pred...
Software tests and code evolve together: a code change should be followed by new or updated tests that record the new...
Visual token pruning is a crucial strategy for accelerating VLMs by compressing redundant image patches, yet existing...
In this work, we focus on SE-RRMs, a symbol-equivariant instantiation of RRMs that exhibits improved extrapolation to...
Machine learning interatomic potentials (MLIPs) have become a hallmark of AI for scientific simulation. While efforts...
On-policy self-distillation (OPSD) has emerged as a practical method for training large language models (LLMs) to rea...
Long-form TV dramas present a formidable challenge for comprehensive video understanding, where deciphering complex s...