HuggingFace

One Success Isn't Reliability: Thinkingbox, a Sandbox and Benchmark for Agents in Stateful Business Workflows

Recent agent benchmarks increasingly ground evaluation in executable environments, from code repair to web navigation...

AI 聚合
2026-08-25
HuggingFace

EchoWM: Open and Enterable Omnimodal World Models

We present EchoWM, an omnimodal world model for enterable generative media that responds to continuous navigation whi...

AI 聚合
2026-08-25
HuggingFace

Beyond the Stability-Exploration Dilemma: Environmental Regularization for LLM Policy Optimization

Policy optimization (PO) for Large Language Models faces a stability--exploration trade-off, currently mediated by an...

AI 聚合
2026-08-25
HuggingFace

RIBOSPAN: A Long-Context RNA Foundation Model for Versatile RNA Modeling

Full-length RNAs, particularly messenger RNAs, often exceed the context lengths used to pretrain existing RNA foundat...

AI 聚合
2026-08-25
HuggingFace

From Generation to Simulation: How Far Are World Models from Being True Simulators?

With the rapid progress of diffusion models and large-scale video generation, generative world models are increasingl...

AI 聚合
2026-08-25
HuggingFace

Industrial-Instruction: An End-to-End Framework for Building Instruction-Tuning and Benchmark Datasets from Industrial Technical Reports

Industrial technical reports contain high-value knowledge for maintenance, troubleshooting, and product engineering, ...

AI 聚合
2026-08-25
HuggingFace

Towards a Densing Law for User Representation Learning at Billion-Scale Capacity

User representation learning in real-world industrial scenarios is commonly scaled by increasing user amount, behavio...

AI 聚合
2026-08-25
HuggingFace

Hybrid Quantum-inspired Kolmogorov-Arnold Networks for Privacy-Aware Federated Biosignal Learning

Electrocardiogram (ECG) recordings are sensitive biomedical data, limiting the ability of hospitals and wearable devi...

AI 聚合
2026-08-25
HuggingFace

WorldToken: Time-First Sequence Modeling for Robotic Imitation Learning

Robot policies receive heterogeneous observations at each decision step, yet sequence models differ in how they organ...

AI 聚合
2026-08-25
HuggingFace

Human-Centric Intelligence in the Era of Foundation Models: A Survey

Human-centric intelligence is evolving in the foundation-model era, with growing emphasis on scale, transferability, ...

AI 聚合
2026-08-25
HuggingFace

FlavourBench: Ranking Frontier Language Models with Executable Culinary Ground Truth

Open-ended language-model benchmarks usually inherit a judge: a human preference panel, another model, or a brittle e...

AI 聚合
2026-08-25
HuggingFace

ParaTempo: Efficient Parallel Reasoning via Temporal Confidence

Parallel reasoning improves the accuracy and robustness of large reasoning models by exploring multiple solution path...

AI 聚合
2026-08-25
首页 上一页 第 32 / 212 页 下一页 末页