Running 220 The ultimate guide to multi-harness RL 🔀 220 Train open models with RL inside real agent harnesses
Running Featured 71 How to turn a game into an RL environment 🌍 71 From an idea to a trained 4B, with the dead ends left in
view article Article Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident +2 hlarcher, XciD, raphael-gl, chris-rannou • Jul 27 • 509
Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks Paper • 2606.29082 • Published Jun 27 • 44
Running 69 Don't Train the Model, Evolve the Harness 🌿 69 Evolving an agent's harness, not its model, on Harvey's LAB
view article Article Profiling in PyTorch (Part 1): A Beginner's Guide to torch.profiler +3 ariG23498, sayakpaul, sergiopaniego, ror, pcuenq • May 29 • 176
Running 259 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 259 Building and scaling RL environments for LLM training
Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe Paper • 2604.13016 • Published Apr 14 • 116
view article Article Training and Finetuning Multimodal Embedding & Reranker Models with Sentence Transformers tomaarsen • Apr 16 • 83
Running Featured 95 Distilling 100B+ Models 40x Faster with TRL 📝 95 TRL distillation for 100B+ teachers, 40x faster
view article Article TRL v1.0: Post-Training Library Built to Move with the Field +2 qgallouedec, stevhliu, pcuenq, sergiopaniego • Mar 31 • 60