Running Agents 436 Reward Bench Leaderboard 📐 436 Explore and compare model scores on RewardBench benchmarks
nvidia/Llama-3.1-Nemotron-70B-Instruct-HF Text Generation • 71B • Updated Apr 13, 2025 • 14.4k • • 2.07k