NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness Paper • 2609.08183 • Published 4 days ago • 397
Measuring Language Transfer in Robot Policies: Adding Greek to a Cosmos3 Vision-Language-Action Policy Paper • 2609.07470 • Published 5 days ago • 19
RoboSPA: Can VLA Models Go Beyond Simple Scenes and Short-Horizon Tasks? Paper • 2609.05324 • Published 8 days ago • 25
Procedural Graphs: Self-Evolving Execution Structures for LLM Agents Paper • 2609.09153 • Published 4 days ago • 32
GE-Act 2.0: Pretraining and Scaling a World-Action Model for Robotic Manipulation Paper • 2609.05588 • Published 8 days ago • 51
Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation Paper • 2609.08798 • Published 4 days ago • 88
OpenWAM: An Open, Modular Exploration Towards Systematic World-Action Model Pretraining Paper • 2609.07398 • Published 5 days ago • 65
SimpleMemVLA: A Simple but Effective Native-Video Memory for Vision-Language-Action Models Paper • 2609.05533 • Published 10 days ago • 7
EmbodiedSkills: A Unified Framework for Orchestrating, Training, and Deploying VLA Agents Paper • 2609.01281 • Published 11 days ago • 14
τ^τ-Bench: An Environment for End-To-End, Realistic Agent Construction Paper • 2609.04611 • Published 8 days ago • 10
Dr. Claw: An AI Scientist Workspace for Vibe Research Paper • 2609.00365 • Published 12 days ago • 173
Rethinking On-Policy Distillation of Large Language Models II: One Training Example Paper • 2609.04172 • Published 9 days ago • 91
MULTI3IR: A Benchmark for Multi-perspective Multi-domain Multi-modal Information Retrieval Paper • 2608.30949 • Published 12 days ago • 9
Post-Training Language Models for Gold-Medal Performance in Coding Competitions Paper • 2609.02849 • Published 10 days ago • 11
Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills Paper • 2609.02749 • Published 10 days ago • 543
EM^2Mem: Event-Centric Multimodal Memory for Large Language Models Paper • 2609.00551 • Published 11 days ago • 14
DiagEvo: Diagnosis-Guided Self-Evolution via Hierarchical Error Memory Paper • 2609.00768 • Published 11 days ago • 22