Running 118 Unlocking On-Policy Distillation for Any Model Family 📝 118 Explore on-policy distillation visualization for any model
Paused Agents 7 Dataset Length Profiler 👁 7 Estimate optimal max_length for SFT training with token analysis
Running 3.95k The Ultra-Scale Playbook 🌌 3.95k The ultimate guide to training LLM on large GPU Clusters
Running Agents 88 Large Reasoning Models Leaderboard 🐳 88 A leaderboard to rank large reasoning models
Running 602 Scaling test-time compute 📈 602 Boost LLM answers with flexible test‑time search strategies
Running Agents 432 Reward Bench Leaderboard 📐 432 Explore and compare model scores on RewardBench benchmarks
Running on CPU Upgrade 14.1k Open LLM Leaderboard 🏆 14.1k Track, rank and evaluate open LLMs and chatbots