Running Agents Featured 590 LLM-Perf Leaderboard 🏆 590 Compare LLM hardware performance and find the best model
Running Agents 1.52k Big Code Models Leaderboard 📈 1.52k Explore and compare code model performance on a leaderboard
Running on Zero Agents 18 Chat with Gemma-2-9B-Chinese-Chat 💬 18 Chat with a Chinese language assistant
Running Agents 436 Reward Bench Leaderboard 📐 436 Explore and compare model scores on RewardBench benchmarks