Pass the Baton: Trajectory-Relayed On-Policy Distillation Paper • 2607.26057 • Published 6 days ago • 31
Explorative Modeling: Unlocking a Third Pretraining Axis and End-to-End Generation Paper • 2607.27372 • Published 5 days ago • 13
Flux-OPD: On-Policy Distillation with Evolving Contexts Paper • 2607.28022 • Published 4 days ago • 40
MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training Paper • 2606.30406 • Published Jun 29 • 19
BM25 Wins at Scale: A Scaling Study of Retrieval-Augmented Generation Paradigms Paper • 2607.26497 • Published 4 days ago • 46
PhiZero: A World Model Built Around Physical Language Paper • 2607.28624 • Published 4 days ago • 158
Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering Paper • 2607.28568 • Published 4 days ago • 168
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published 4 days ago • 281
MemPrivacy: Privacy-Preserving Personalized Memory Management for Edge-Cloud Agents Paper • 2605.09530 • Published May 10 • 150