arXiv

Catching the Rug: Early Prediction of Fraudulent Memecoins on Solana via Machine Learning

The rapid proliferation of memecoins on blockchain platforms has increased the risk of fraudulent activities, particu...

AI 聚合
2026-08-21
arXiv

Break It Down, Pass It On: Cross-Task Skill Transfer in LLM Agents

Large language model (LLM) agents can induce skills from completed tasks and reuse them later to grow more capable wi...

AI 聚合
2026-08-21
arXiv

Inject, Align, Recover: Staged Post-Training for Retrieval-Free Document Knowledge Internalization

Large language models often fail to answer questions about a bounded document collection when the source documents ar...

AI 聚合
2026-08-21
arXiv

Phantom Gains: Auditing Self-Improvement Against a Measured Null

Whether a language model has improved itself is increasingly judged not by mean accuracy but by which individual prob...

AI 聚合
2026-08-21
arXiv

MidTool: Mid-training Data Synthesis for Agentic Tool Use

Mid-training is increasingly recognized as a critical stage for shaping the capabilities of large language models. Re...

AI 聚合
2026-08-21
arXiv

Pandora's AI Model Routing Box: Efficient Allocation with Costly Value Estimation

Heterogeneous AI systems composed of multiple models, architectures, harnesses, or inference-time settings can improv...

AI 聚合
2026-08-21
arXiv

AI4AI-Bench: Benchmarking LLM Agents in Algorithmic Design for Recursive Self-Improvement

Recursive self-improvement (RSI) asks whether an AI system can improve the process that produces AI systems, so that ...

AI 聚合
2026-08-21
arXiv

Inducing Task Models from Computer-Use Traces

Naturalistic computer-use traces, passively recorded screenshots and mouse or keyboard actions, are a valuable resour...

AI 聚合
2026-08-21
arXiv

An Agentic Approach for Active Data Collection, Travel Behavior Modeling, and Weather-Sensitive Demand Prediction

Travel behavior research increasingly combines digital data collection with predictive modeling, yet these stages are...

AI 聚合
2026-08-21
arXiv

G-CARL: Grounded Checklist-Aligned Reward Learning for Patient-Oriented Medical Report Interpretation

Personalized interpretation of medical reports has emerged as an increasingly important need among patients. Addressi...

AI 聚合
2026-08-21
HuggingFace

PolicyGuide: From Guarding One Action to Guiding the Whole Workflow for Policy-Compliant LLM Agents

Customer-service LLM agents must follow organizational policy when acting on a user's behalf. Compliance failures ari...

AI 聚合
2026-08-21
HuggingFace

EXIMO: VLM Guided Exploration of VLA Policies

How to efficiently finetune robot policies to learn new tasks on the fly? State of the art robotic manipulation polic...

AI 聚合
2026-08-21
首页 上一页 第 37 / 212 页 下一页 末页