Rethinking On-Policy Distillation of Large Language Models II: One Training Example Paper • 2609.04172 • Published 1 day ago • 37
Safin-1: Safety from Within through Memory-Native State Evolution Paper • 2609.00092 • Published 4 days ago • 20
MARCH: Scaling Recurrent Memory with Content-Routed State Anchors Paper • 2608.12435 • Published 23 days ago • 1
Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning Paper • 2608.14290 • Published 21 days ago • 33
Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering Paper • 2607.28568 • Published Jul 30 • 186
Achieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified Scaling Paper • 2605.13301 • Published May 13 • 167