Running Agents 435 Reward Bench Leaderboard 📐 435 Explore and compare model scores on RewardBench benchmarks
Runtime error Featured 142 smolagents LLM leaderboard 🏆 142 A leaderboard for LLMs powering smolagents
microsoft/Phi-4-multimodal-instruct Automatic Speech Recognition • 6B • Updated Dec 10, 2025 • 566k • 1.61k
Running 3.97k The Ultra-Scale Playbook 🌌 3.97k The ultimate guide to training LLM on large GPU Clusters