SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning Paper • 2607.14777 • Published 6 days ago • 94
nvidia/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4 Text Generation • 45B • Updated 15 days ago • 54.3k • 122
SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe Paper • 2607.03451 • Published 19 days ago • 33
EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments Paper • 2607.05155 • Published 16 days ago • 18
Qwen-AgentWorld: Language World Models for General Agents Paper • 2606.24597 • Published 29 days ago • 150
yuxinlu1/gemma-4-12B-agentic-fable5-composer2.5-v2-3.5x-tau2-GGUF Text Generation • 12B • Updated Jun 19 • 515k • 1.26k
Towards Foundation Models for Learning on Tabular Data Paper • 2310.07338 • Published Oct 11, 2023 • 3
PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation Paper • 2606.18375 • Published Jun 16 • 12
Multi-LCB: Extending LiveCodeBench to Multiple Programming Languages Paper • 2606.20517 • Published Jun 18 • 60
Jackrong/Qwopus3.6-27B-Coder-Compat-MTP-GGUF Image-Text-to-Text • 0.5B • Updated 13 days ago • 559k • 125