SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Paper • 2609.20519 • Published 10 days ago • 135
PAWBench: How Far Are We from Probabilistically Aligned World Modeling? Paper • 2608.27345 • Published about 1 month ago • 76