SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Paper • 2609.20519 • Published 16 days ago • 138
PAWBench: How Far Are We from Probabilistically Aligned World Modeling? Paper • 2608.27345 • Published Aug 27 • 77