Self-Improving-Coding-Agents/SI2CA-Training-Trajectories Viewer • Updated 1 day ago • 43.1k • 40
Self-Improving-Coding-Agents/SI2CA-Training-Trajectories Viewer • Updated 1 day ago • 43.1k • 40
SAS: Simple Attention Sparsification via End-to-End Optimization of Context Ranking Paper • 2609.13141 • Published 12 days ago • 65
Apodex 1.1: Scaling Agentic Intelligence for Complex Work Paper • 2608.23283 • Published about 1 month ago • 210
LEGO-RL: Harness-Native Reinforcement Learning for Coding Agents Paper • 2608.17393 • Published Aug 18 • 25
SWE-rebench V2: Language-Agnostic SWE Task Collection at Scale Paper • 2602.23866 • Published Feb 27 • 92
Guava: An Effective and Universal Harness for Embodied Manipulation Paper • 2606.18363 • Published Jun 16 • 28
Reinforcement Learning Elicits Contextual Learning of Unseen Language Translation Paper • 2606.06428 • Published Jun 4 • 25
EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL Paper • 2605.18703 • Published May 18 • 50
ClawEnvKit: Automatic Environment Generation for Claw-Like Agents Paper • 2604.18543 • Published Apr 20 • 31
OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks Paper • 2604.08539 • Published Apr 9 • 49
SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization Paper • 2604.02268 • Published Apr 2 • 100
HopChain: Multi-Hop Data Synthesis for Generalizable Vision-Language Reasoning Paper • 2603.17024 • Published Mar 17 • 111