VICBench: A Multi-Language Benchmark for Code Vulnerability Detection
Evaluating security vulnerability detection tools requires benchmark datasets with vulnerability-inducing commits (VI...
每天自动聚合 AI 领域最新动态
Evaluating security vulnerability detection tools requires benchmark datasets with vulnerability-inducing commits (VI...
Modernizing legacy Fortran is a problem of volume: the transformations are individually routine, but the codebases ca...
Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simu...
Multimodal Large Language Models (MLLMs) have been growing the capability for scientific writing and collaboration. F...
LLM agents increasingly rely on third-party skills, using natural-language descriptions for selection and instruction...
Background: Accurate segmentation of the Left Anterior Descending (LAD) artery in 3D free-breathing, non-contrast CT ...
Artificial intelligence tools for education and language support are increasingly framed as scalable responses to acc...
Agents deployed in enterprise settings must reason across structured APIs and document collections, yet existing benc...
Modern black-box Image-to-Video (I2V) models offer powerful capabilities in automated content creation, yet their lac...
Class activation mapping (CAM) is one of the most widely used visual explanation families in explainable artificial i...
Dynamic Master Logic (DML) provides a hierarchical framework for representing system behavior by linking functional o...
Safe offline RL typically assumes access to dense per-step cost annotations, but in practice supervisors provide only...