DataFlow-Harness: A Grounded Code-Agent Platform for Constructing Editable LLM Data Pipelines Paper • 2607.16617 • Published 8 days ago • 134
Scaling Laws for Hypernetwork-Based Knowledge Injection in Large Language Models Paper • 2607.19604 • Published 5 days ago • 14
SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD Paper • 2607.20145 • Published 4 days ago • 56
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 3 days ago • 124
RAGU: A Multi-Step GraphRAG Engine with a Compact Domain-Adapted LLM Paper • 2607.11683 • Published 13 days ago • 143
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources Paper • 2606.29538 • Published 10 days ago • 139
LLM-as-a-Coach: Experiential Learning for Non-Verifiable Tasks Paper • 2607.18110 • Published 6 days ago • 14
Rethinking the Evaluation of Harness Evolution for Agents Paper • 2607.12227 • Published 12 days ago • 9
From Human-Centric to Agentic Code Review: The Impact of Different Generations of Generative AI Technology on Review Quality Paper • 2607.13196 • Published 12 days ago • 29
AdvancedMathBench: A Benchmark Suite for Advanced Mathematical Proof Generation and Verification Paper • 2607.11849 • Published 13 days ago • 33
Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models Paper • 2607.12463 • Published 12 days ago • 108
Self-Improvements in Modern Agentic Systems: A Survey Paper • 2607.13104 • Published 12 days ago • 31
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning Paper • 2607.14777 • Published 10 days ago • 100
Ring-Zero: Scaling Zero RL to a Trillion Parameters for Emergent Reasoning Paper • 2607.12395 • Published 12 days ago • 98
Unified Audio Intelligence Without Regressing on Text Intelligence Paper • 2607.05196 • Published 20 days ago • 23
SWE-Review: Closing the Loop on Issue Resolution with Agentic Code Review Paper • 2607.06065 • Published 19 days ago • 8
SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe Paper • 2607.03451 • Published 23 days ago • 33