JanusMesh: Fast and Zero-Shot 3D Visual Illusion Generation via Cross-Space Denoising
Creating 3D visual illusions, a single 3D mesh that reveals entirely different semantics from various viewing angles,...
每天自动聚合 AI 领域最新动态
Creating 3D visual illusions, a single 3D mesh that reveals entirely different semantics from various viewing angles,...
Real-world spatial intelligence requires reasoning over a continuous and evolving 3D world, yet existing VLMs and too...
Current agentic robot systems can write executable Code-as-Policy programs, observe feedback, and revise behavior acr...
Advances in radiance fields have enabled photorealistic novel view synthesis. In several domains, large-scale real-wo...
Embodied foundation models are expected to benefit from data scaling like large language models, but face a much tigh...
Dexterous interaction with articulated objects is important for household, assistive, and humanoid manipulation, wher...
Conditional diffusion and flow models routinely fail to satisfy the very constraints that define their task. For inst...
Hybrid linear attention models offer an appealing path to faster long-context inference: they reduce the quadratic co...
Large Language Models (LLMs) have significantly advanced the automation of software engineering tasks. One prominent ...
LiveCodeBench (LCB) has recently become a widely adopted benchmark for evaluating large language models (LLMs) on cod...
FP4 training promises substantial reductions in memory and computation cost for LLM pretraining, yet current FP4 hard...
Scheduling policies in large-scale Automatic Speech Recognition (ASR) serving pipelines play a key role in determinin...