SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning Paper • 2607.14777 • Published 28 days ago • 106
Self-Improvements in Modern Agentic Systems: A Survey Paper • 2607.13104 • Published about 1 month ago • 33
Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning Paper • 2607.07690 • Published Jul 8 • 5