RSIAgent: Autonomous Exploration for Recursive Self-improvement in New Environments Paper • 2609.15364 • Published 1 day ago • 23
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 1 day ago • 251
NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept Prediction Paper • 2609.10715 • Published 7 days ago • 309
LoopArena: Benchmarking Models as Runtime Controllers for Loop Engineering Paper • 2608.28281 • Published 19 days ago • 106
CogEvol: Towards Efficient and Reliable Learning Environment Generation Paper • 2608.30968 • Published 16 days ago • 32
On the Design of Qwen3.8-Next Architecture: Evaluation, Efficiency, and Training Stability Paper • 2608.30320 • Published 16 days ago • 59
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement Paper • 2608.31046 • Published 16 days ago • 155
Harness-of-Harness: Multi-Day Autonomous Software Development with Continual Improvement Paper • 2609.01481 • Published 15 days ago • 19
Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving Paper • 2609.00111 • Published 16 days ago • 384
Cliff: Learning Process Rewards from the First Mistake Paper • 2609.02817 • Published 14 days ago • 21
WHALE: A Simple Recipe for Joint Harness-Weight Optimization Paper • 2609.00196 • Published 16 days ago • 36
HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness? Paper • 2609.01437 • Published 15 days ago • 268
Rethinking On-Policy Distillation of Large Language Models II: One Training Example Paper • 2609.04172 • Published 13 days ago • 97