HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness? Paper • 2609.01437 • Published 3 days ago • 166
DART-SD: Diamond-topology Aware Retrieval and Tuning for Self-Distillation of Multi-Turn Tool-Calling Agents Paper • 2608.18524 • Published 16 days ago • 91
SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers Paper • 2609.01343 • Published 3 days ago • 86
Scaling Large Reasoning Models beyond Human Supervision: A Path toward Superintelligence Paper • 2608.31075 • Published 4 days ago • 28
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement Paper • 2608.31046 • Published 4 days ago • 134
WeMM-Embedding: WeChat Multi-Modal Embedding Technical Report Paper • 2608.24053 • Published 10 days ago • 70
Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMs Paper • 2608.20492 • Published 15 days ago • 111
Agentic Game Development as a Verifiable Trajectory Data Engine for Scaling World Models Paper • 2608.25518 • Published 9 days ago • 195
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 9 days ago • 179
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 9 days ago • 179
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 9 days ago • 179