Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
Pythonallenai/reward-bench

reward-bench

RewardBench: the first evaluation tool for reward models.

55.2/100
727Forks: 97
View on GitHubHomepage →
Loading report...

Similar Projects

LlamaFactory

92

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python73.5K

alignment-handbook

62

Robust recipes to align language models with human and AI preferences

Python5.6K

OpenClaw-RL

59

OpenClaw-RL: Train any agent simply by talking

Python5.6K

transformerlab-app

89

The open source research environment for AI researchers to seamlessly train, evaluate, and scale models from local hardware to GPU clusters.

Python5.2K
Back to List