arXiv

Natural Ungrokking: Asymmetric Control of Which Rules Survive Pretraining

Midway through an ordinary pretraining run, a small language model learns the pronoun-gender rule: cued with a girl's...

AI 聚合
2026-06-25
arXiv

The Unfireable Safety Kernel: Execution-Time AI Alignment for AI Agents and Other Escapable AI Systems

AI agents are granted access to tools, APIs, and other infrastructure, making them active principals in those systems...

AI 聚合
2026-06-25
arXiv

A welding penetration prediction model for laser welding process based on self-supervised learning using physics-informed neural networks

The laser welding full-penetration is of critical importance, as it constitutes one of the fundamental factors in ach...

AI 聚合
2026-06-25
arXiv

Model Forensics: Investigating Whether Concerning Behavior Reflects Misalignment

A central goal of safety research is determining whether a model is misaligned. Prior work has largely focused on det...

AI 聚合
2026-06-25
arXiv

A cross-process welding penetration status prediction algorithm based on unsupervised domain adaptation in laser and TIG welding

Supervised deep learning has been widely used for weld penetration state classification; however, its performance oft...

AI 聚合
2026-06-25
arXiv

Neglected Free Lunch from Post-training: Progress Advantage for LLM Agents

Process reward models enable fine-grained, step-level evaluation of LLMs, yet building them for agentic settings rema...

AI 聚合
2026-06-25
arXiv

On-Policy Self-Distillation with Sampled Demonstrations Reduces Output Diversity

On-policy self-distillation achieves strong pass@1 accuracy by using a single model as both teacher and student, with...

AI 聚合
2026-06-25
arXiv

Learning Action Priors for Cross-embodiment Robot Manipulation

Most Vision-Language-Action (VLA) models build on a Vision-Language Model (VLM) backbone by attaching an action modul...

AI 聚合
2026-06-25
HuggingFace

IV-CoT: Implicit Visual Chain-of-Thought for Structure-Aware Text-to-Image Generation

Unified multi-modal large language models (MLLMs) have achieved strong text-to-image generation quality, but still st...

AI 聚合
2026-06-25
HuggingFace

Beyond NL2Code: A Structured Survey of Multimodal Code Intelligence

While Large Language Models (LLMs) have substantially advanced text-to-code synthesis, many real programming tasks sp...

AI 聚合
2026-06-25
HuggingFace

Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models

Autoregressive video diffusion with causal diffusion transformers has emerged as a major paradigm for real-time strea...

AI 聚合
2026-06-25
HuggingFace

Improved Large Language Diffusion Models

Modern large language models are predominantly trained with autoregressive factorization and causal attention. We pre...

AI 聚合
2026-06-25
首页 上一页 第 175 / 212 页 下一页 末页