Running 215 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 215 Building and scaling RL environments for LLM training
view article Article What We Learned by Reproducing 2,200 papers from ICML abidlabs • about 21 hours ago • 13
Qwen-AgentWorld: Language World Models for General Agents Paper • 2606.24597 • Published Jun 23 • 158
Atemokoloporos Qwen3.5-0.8B retained checkpoints Collection Public evidence and retained failed or inconclusive LoRA checkpoints from the Atemokoloporos synthetic-fact study. • 9 items • Updated 5 days ago