Running 105 The ultimate guide to multi-harness RL 🔀 105 Train open models with RL inside real agent harnesses
Running on Zero Agents Featured 2.28k Qwen3-TTS Demo 🎙 2.28k Generate speech from text with voice design, cloning, or presets
Running Featured 94 Distilling 100B+ Models 40x Faster with TRL 📝 94 TRL distillation for 100B+ teachers, 40x faster
Runtime error MCP Featured 126 Mage-Flow 🎨 126 Efficient native-resolution image generation and editing
Running 68 Don't Train the Model, Evolve the Harness 🌿 68 Evolving an agent's harness, not its model, on Harvey's LAB
Running 252 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 252 Building and scaling RL environments for LLM training
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 21.7k • • 2.95k
Running on CPU Upgrade 281 The Synthetic Data Playbook: Generating Trillions of the Finest Tokens 📝 281 Explore synthetic data experiments as a visual bookshelf
mistralai/Voxtral-Mini-4B-Realtime-2602 Automatic Speech Recognition • 4B • Updated Mar 11 • 1.89M • 999
Running Featured 1.45k FineWeb: decanting the web for the finest text data at scale 🍷 1.45k Explore and download the FineWeb web‑scale text dataset
Running 4.06k The Ultra-Scale Playbook 🌌 4.06k The ultimate guide to training LLM on large GPU Clusters
Running on CPU Upgrade Featured 3.32k The Smol Training Playbook 📚 3.32k The secrets to building world-class LLMs