Can AI agents conduct open-ended AI research? Early evidence from two case studies
Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether agents can carr...
每天自动聚合 AI 领域最新动态
Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether agents can carr...
Most video editing systems still lack explicit layered video representations, limiting their ability to perform reali...
We introduce GPT-Red, an automated red-teaming agent that is trained to discover novel prompt injection attacks again...
Evaluating whether a vision-language model (VLM) can act through a physical body is challenging. The outcome of an ac...
Vision-language-action (VLA) models commonly adopt an LLM-centric V to L to A pathway, where visual observations are ...
Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether agents can carr...
Large language model (LLM) agents are increasingly expected to assist users in completing tasks. However, existing be...
Stealth, the discipline of achieving an objective without revealing your presence, capabilities, or collected intelli...
Training large language models (LLMs) to act in long-horizon games is a promising step toward generalist decision-mak...
Text-space optimization adapts large language models (LLMs) by editing external natural-language artifacts rather tha...
Large language model agents often encounter related yet distinct tasks that share reusable solution patterns. Yet sta...
Rubric-based reinforcement learning enriches language model training by evaluating model outputs against explicit cri...