NEW Articles from Team or Enterprise organizations will get promoted to the main section. DistribAI, a new training framework
Enderchef
• • 1
How to Make a Small LLM a MiCA Specialist — What Went Right and What Went Wrong
svetlink
• One workstation, 6,341 papers, 18 days: the workflow is the product
ParetoOptimal
• Five machine raters, one definition, and answers ranging from 0 to 78
leventbulut
• On Measuring Progress Toward AGI
JosefAlbers
• What Is MiniMax H3 (Hailuo 3.0)? The Open-Weight Multimodal Video Model, Explained
ResterChed
• int-llm precision ladder: from a wide integer oracle to compact weights
nmicic
• DeepSeek V4 Flash Is Now Official: What Changed in the 0731 Build
Chapter 02 — AMNESIA BY DESIGN
SoulInPsyAbstract
• Exact E2M1 on Hopper
KissTheHabit
• Text-Only Models with mm-ctx Vision Toolkit vs. Native Vision Models
Accelerating Qwen3.6 on Intel® Core™ Ultra Series 3 with DFlash
mDenseOn with the mLateOn: Open Multilingual, Long-Context, and Code Retrieval Models
lightonai
• • 23
Can you train a model on Simon Willison's deeply unscientific pelican benchmark?
sergiopaniego
• • 1
What Do Memory Benchmarks Actually Measure? (Hint: Not Storage)
Geometric Memory FT4 — Distill Against a Consensus, Ship a Rotation
AbstractPhil
• I Built a RAG System, Then Tried to Break It
nazeerbashashaik
• • 1
🎲 Apprendre à un réseau à écrire avec seulement une récompense — il a atteint 99,9 % grammatical sans apprendre une seule règle 🇫🇷
RDTvlokip
• • 3
🎲 Teaching a network to write with reward only — it hit 99.9% grammatical without learning a single rule 🇫🇷
RDTvlokip
• • 1
24/24 Retrieval, Yet the Embeddings Changed: An MLX Q4–Q8 Sweep with CUDA Controls
TiGa-RCE
• • 2