Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments Paper • 2609.04148 • Published 4 days ago • 272
SolarWM: Open Data and Scalable Training for Long-Horizon Video World Models Paper • 2609.02886 • Published 5 days ago • 141
Rethinking On-Policy Distillation of Large Language Models II: One Training Example Paper • 2609.04172 • Published 4 days ago • 77
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning Paper • 2609.03430 • Published 4 days ago • 161
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 4 days ago • 226
HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness? Paper • 2609.01437 • Published 6 days ago • 241
DART-SD: Diamond-topology Aware Retrieval and Tuning for Self-Distillation of Multi-Turn Tool-Calling Agents Paper • 2608.18524 • Published 19 days ago • 91
SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers Paper • 2609.01343 • Published 6 days ago • 96
Scaling Large Reasoning Models beyond Human Supervision: A Path toward Superintelligence Paper • 2608.31075 • Published 7 days ago • 28
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement Paper • 2608.31046 • Published 7 days ago • 141
WeMM-Embedding: WeChat Multi-Modal Embedding Technical Report Paper • 2608.24053 • Published 13 days ago • 70
Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMs Paper • 2608.20492 • Published 18 days ago • 111
Agentic Game Development as a Verifiable Trajectory Data Engine for Scaling World Models Paper • 2608.25518 • Published 12 days ago • 196
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 12 days ago • 179
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 12 days ago • 179