Control-Data Flow Separation: Stable Prompt Optimization in Multi-Agent LLMs Paper • 2609.00621 • Published 2 days ago • 8
Agents in the Large: Perception-Centered Architecture for Persistent Agents Paper • 2608.30478 • Published 3 days ago • 9
Harness-of-Harness: Multi-Day Autonomous Software Development with Continual Improvement Paper • 2609.01481 • Published 2 days ago • 12
Evaluating Multimodal LLMs as Generalist Vision-Language-Action Agents for Drone Control: Commanding, Approaching, Tracking and Searching Paper • 2609.01404 • Published 2 days ago • 23
Hi-Q: Hierarchical Evidence-guided Query Refinement for Multi-Hop Question Answering Paper • 2608.30468 • Published 3 days ago • 30
H3-World: Turning Language Understanding into World Control Paper • 2609.01560 • Published 2 days ago • 44
SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers Paper • 2609.01343 • Published 2 days ago • 81
LightNav-0: Eliciting VLM Spatial Intelligence for Generalist Embodied Navigation Paper • 2608.30935 • Published 3 days ago • 27
Act with Intent: Distilling Behavior Intent for Vision-Language-Action Models Paper • 2608.23478 • Published 10 days ago • 30
LoopArena: Benchmarking Models as Runtime Controllers for Loop Engineering Paper • 2608.28281 • Published 6 days ago • 100
Training Agents to Evolve with Their Harness: TaoLive Digital Avatar Agent Technical Report Paper • 2608.15763 • Published 12 days ago • 50
VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning Paper • 2608.26105 • Published 8 days ago • 269
PILOT in the Loop: Live Self-Improvement for Long-Horizon Agents Paper • 2608.26530 • Published 7 days ago • 32
MA-VLA: Multi-Arm Vision-Language-Action Model for Collaboration and Compositional Generalization Paper • 2608.25864 • Published 8 days ago • 9