Sol-Attn: Accelerating Video Generation Inference via On-the-Fly Attention Sparsification
Diffusion transformers are essential for high-fidelity video generation, but long token sequences make attention a do...
每天自动聚合 AI 领域最新动态
Diffusion transformers are essential for high-fidelity video generation, but long token sequences make attention a do...
Large reasoning models (LRMs) generate long reasoning traces before producing final answers. While these traces may c...
Since Volta introduced Independent Thread Scheduling (ITS), NVIDIA GPUs have been widely assumed to handle warp diver...
Industrial Video Anomaly Detection (IVAD) aims to identify anomalous objects and events in an industrial process, whi...
Experiment trackers show how training is progressing, but changing a live run still usually requires trainer-specific...
Humans routinely communicate through abstractions of their bodies, including shadows, silhouettes, and reflections. Y...
For scale-invariant deep networks, Hyperball-style optimizers have shown strong performance in large-scale training b...
Enterprise AI agents are typically granted static credential sets at configuration time, holding every tool the role ...
Do learned audio embeddings encode structure that nobody told them to encode? We probe four large pretrained audio mo...
Generative AI is reshaping programming education, yet educators often infer students' AI-supported learning from clas...
Routing to select large language models (LLMs) with different cost-quality trade-offs has become a fundamental deploy...
Quantum state preparation is a key component of many quantum algorithms. Performing this step efficiently is essentia...