Procedural Graphs: Self-Evolving Execution Structures for LLM Agents Paper • 2609.09153 • Published 5 days ago • 36
view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • 10 days ago • 102
view article Article Training a coding model to paint watercolours with TRL and OpenEnv sergiopaniego • 10 days ago • 65
Kraken PP-OCRv6 text recognition models Collection Hub mirrors of Benjamin Kiessling's multilingual PP-OCRv6 line-recognition family for Kraken: tiny, small, and medium. • 3 items • Updated 9 days ago • 7
PILOT in the Loop: Live Self-Improvement for Long-Horizon Agents Paper • 2608.26530 • Published 17 days ago • 35
Agentic Transaction: Towards ACID-Compliant Agent Systems Paper • 2608.13900 • Published 30 days ago • 27
How Do Agents Fail on AutoResearch: End-to-End Diagnostic Evaluation on 100 Real-World Frontier Research Tasks Paper • 2608.14905 • Published 30 days ago • 31
Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization Paper • 2608.16072 • Published 27 days ago • 151
view article Article Multi-Vector (Late Interaction) Embedding Models with Sentence Transformers +1 tomaarsen, NohTow, raphaelsty • 26 days ago • 110
view article Article Meta is back with Muse Glimmer: local, agentic, multimodal, and open source +2 pcuenq, merve, burtenshaw, ariG23498 • Aug 10 • 111
LettuceDetect v2 Collection SOTA hallucination detection for agentic workflows, multilingual, long context • 6 items • Updated Aug 5 • 3
Explorative Modeling: Unlocking a Third Pretraining Axis and End-to-End Generation Paper • 2607.27372 • Published Jul 29 • 19
jina-reranker-v3.5: An Efficient Listwise Reranker with Hybrid Attention and Self-Distillation Paper • 2607.18152 • Published Jul 20 • 6
jina-reranker-v3: Last but Not Late Interaction for Document Reranking Paper • 2509.25085 • Published Sep 29, 2025 • 13