HarnessEval-W: Agentifying the Evaluation of Visual Worlds Paper • 2608.16859 • Published 14 days ago • 339
EnvHarness: Awakening Static Worlds for Agent Learning Paper • 2608.19880 • Published 11 days ago • 271
OmniScientist: An Omni-Modal Omni-Discipline AI Scientist Paper • 2608.13558 • Published 18 days ago • 91
Demystifying Agent Skills: Why They Work-Until They Don't Paper • 2608.14036 • Published 17 days ago • 169
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill Paper • 2608.11924 • Published 19 days ago • 289
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 21 days ago • 342
Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs Paper • 2608.01755 • Published 28 days ago • 142
DistillAlign: Coordinating Mode Covering and Mode Seeking in Autoregressive Video Distillation Paper • 2607.26811 • Published Jul 29 • 93
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published Jul 23 • 154
NVIDIA-labs OO Agents: Native Python Object-Oriented Agents Paper • 2607.20709 • Published Jul 22 • 36
stefanocarrera/sqlautophagycode_D_test_Qwen3-8B_t1.25_g7_run0_metrics Viewer • Updated Jul 20 • 579 • 12 • 1
Navigating the Mirage: A Dual-Path Agentic Framework for Robust Misleading Chart Question Answering Paper • 2603.28583 • Published Jul 14 • 17
DuoMem: Towards Capable On-Device Memory Agents via Dual-Space Distillation Paper • 2606.29961 • Published Jun 29 • 10