Hierarchical Self-Improvement: A Framework for Task-Specific Evolvable Agent Harnesses Paper • 2608.08466 • Published 15 days ago • 10
Co-RL: Unsupervised Reasoning Emerges from Diverse Cohort in Multi-agent RL Paper • 2608.17253 • Published 5 days ago • 92
FM-Bench: A Benchmark for Long-Horizon Management with Competing Agents Paper • 2608.18423 • Published 5 days ago • 16
Zetta ζ: An Efficient Closed-Loop Embodied Harness for Self-Evolving Physical Intelligence Paper • 2608.16590 • Published 7 days ago • 141
PixRestore: Unified Image Restoration via Pixel Diffusion Transformer Paper • 2608.16793 • Published 7 days ago • 2
MoE-ViE: Mixture of Experts Vision Encoder for Efficient Image and Video Understanding Paper • 2608.17402 • Published 6 days ago • 14
Dynamic Multi-Byte Prediction With Hierarchical Language Models Paper • 2608.15454 • Published 8 days ago • 18
FreeToken: Efficient Edge-Native MoE Serving with Bandwidth-Adaptive Execution Paper • 2608.16157 • Published 7 days ago • 80
ASI-Bench: At the Dawn of Artificial Superintelligence Paper • 2608.17271 • Published 6 days ago • 60
HarnessEval-W: Agentifying the Evaluation of Visual Worlds Paper • 2608.16859 • Published 7 days ago • 123
Improving the matrix multiplication exponent with modern optimization and AlphaEvolve Paper • 2608.16884 • Published 7 days ago • 17
Understanding Cognition-Induced Risks in Agentic AI Systems Paper • 2608.15304 • Published 9 days ago • 17
Multimodal Model Diffing for Feature Discovery and Control Paper • 2608.09928 • Published 14 days ago • 10
Second Thought: Reasoning in Parallel as LLM Agents Act and Observe Paper • 2608.13667 • Published 10 days ago • 16
Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning Paper • 2608.14290 • Published 10 days ago • 32
Agents Catching Agents: Shortcut Cascades and Benchmark Gaming in Clinical Multi-Agent Systems Paper • 2608.03744 • Published 19 days ago • 6