arxiv:2608.11216
taesiri PRO
taesiri
AI & ML interests
AGI ... one linear layer at a time
Recent Activity
submitted a paper about 7 hours ago
An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics upvoted a paper about 7 hours ago
Scaling Automatic Research Agents via World Models upvoted a paper about 7 hours ago
T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks