Adaptive Critical Token-Aware Retrieval for Repository-Level Code Generation
The repository-level code generation task requires synthesizing code that satisfies task requirements while remaining...
每天自动聚合 AI 领域最新动态
The repository-level code generation task requires synthesizing code that satisfies task requirements while remaining...
Evaluating software engineering agents on realistic benchmarks is costly, since each task may require multi-step code...
Traditional frameworks of political communication operate under linear, event-driven assumptions that treat voter per...
Generating professional scholarly content, such as peer reviews and rebuttals, requires an intricate synergy between ...
SAR-to-EO image translation aims to generate electro-optical (EO) imagery from synthetic aperture radar (SAR) observa...
Multimodal memory offers a scalable interface for long-video question answering, but existing methods often retrieve ...
A central bottleneck in multi-hop Question Answering (QA) is that the granularity at which a question is expressed of...
While unified multimodal models (UMMs) jointly perform visual understanding and generation within a single model, fun...
Prompt optimization can improve multi-agent LLM systems, but the prompts being optimized often serve two entangled ro...
Multimodal Large Language Models (MLLMs) are strong perceivers of images and video. We ask how far that reach extends...
AI tutors are most useful when they adapt to each student's strengths, weaknesses, and preferred guidance, but eviden...
We present H3-World, an efficient framework that turns the 33B MiniMax-H3 video generator into an interactive world m...