arXiv

How Language Models Organize and Structure Moral Knowledge

How do large language models (LLMs) organize moral knowledge? Models detect moral content broadly, but detection is a...

AI 聚合
2026-08-28
arXiv

CLAP: Cross-Embodiment Video World Models are Zero-Shot Physical Simulators

State-of-the-art action-conditioned video models are typically restricted to a single robot embodiment, preventing th...

AI 聚合
2026-08-28
arXiv

Learning a Continuous Sepsis Severity Score Without Hour-by-Hour Supervision: A Two-Site Retrospective Study

Currently used sepsis severity indices rely on fixed variables and weights established decades ago, which are coarsel...

AI 聚合
2026-08-28
arXiv

Beyond F1: Evaluating Coverage and Failure Recovery in AI Model Security Scanners

Static scanners are increasingly used to identify executable or otherwise unsafe content in machine- learning artifac...

AI 聚合
2026-08-28
arXiv

Persona-Execution Separation: An Architecture Pattern for Evolving LLM Agents under Execution Audit

Large language model (LLM) agents in governed organizations must let the persona (instructions, tone, self-presentati...

AI 聚合
2026-08-28
arXiv

Mechanistic Reaction Prediction via Discrete Flow Matching on Graph-Structured Electron Occupation

Chemical reactions are fundamentally transformations in electron space, yet most machine learning approaches model th...

AI 聚合
2026-08-28
arXiv

RedEvoAgent: Automatic Red-Teaming Agent with Experience-Driven Skill Evolution

LLM-based agents are increasingly deployed in product-level execution harnesses, where jailbreaks can trigger harmful...

AI 聚合
2026-08-28
arXiv

From Static to Dynamic: Benchmarking Real-World Code Review with MCR-Bench

In real-world software development, code review typically involves iterative interactions between developers and revi...

AI 聚合
2026-08-28
arXiv

SWE-Prime: Fewer Trajectories, Better Performance

To improve large language models' ability to resolve real-world software issues, prior work has focused on constructi...

AI 聚合
2026-08-28
arXiv

WikiSkill: Compiling Agent Experience into Persistent Knowledge for Skill Evolution

Agent skills package specialized knowledge and workflows into reusable resources that extend AI agent capabilities. R...

AI 聚合
2026-08-28
HuggingFace

What Makes Good Agentic Data? An ACE Lens on Data Generation for LLM Agents

LLM agents increasingly rely on generated interaction data to learn how to interact with external environments. Agent...

AI 聚合
2026-08-28
HuggingFace

Magpie: Real-Time World Renderer for Interactive Games

Modern game development relies heavily on conventional graphics pipelines. High-quality visual content requires model...

AI 聚合
2026-08-28
首页 上一页 第 20 / 212 页 下一页 末页