Running Agents 437 Reward Bench Leaderboard 📐 437 Explore and compare model scores on RewardBench benchmarks
sorry-bench/ft-mistral-7b-instruct-v0.2-sorry-bench-202406 Text Generation • 7B • Updated Jul 2, 2024 • 14.8k • 9