Self-Improvements in Modern Agentic Systems: A Survey
Self-improving autonomous agents are moving from research prototypes to deployed systems. The primary goal is control...
每天自动聚合 AI 领域最新动态
Self-improving autonomous agents are moving from research prototypes to deployed systems. The primary goal is control...
AI pentesting agents are increasingly credible as offensive security systems, but current benchmarks still provide li...
We present AffectFlow-DINO, a multi-task learning system for the 11th ABAW challenge that extends a standard determin...
Discrete denoising diffusion models (DDMs) have recently emerged as a compelling alternative to autoregressive (AR) m...
Autonomous UAV systems increasingly rely on multimodal large language models (MLLMs) to operate in complex real-world...
Interactive simulators have become powerful tools for training embodied agents and generating synthetic visual data, ...
Length-penalized reinforcement learning can shorten chain-of-thought reasoning while hiding an influence that drives ...
Languages with rich static semantics, such as Rust, provide stronger guarantees for AI-generated code, but their stri...
We propose the AIMO Interpretability Challenge, a competition on distinguishing robust from spurious reasoning in fro...
Serial verification gates are a core reliability primitive in LLM harnesses: a candidate answer is returned only if $...
Languages with rich static semantics, such as Rust, provide stronger guarantees for AI-generated code, but their stri...
Personal health management unfolds over repeated encounters, yet most health AI systems treat each request in isolati...