| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Download Repo ZIP] [Original HTTPS Page] |
official implementation of ICLR'2025 paper: Rethinking Bradley-Terry Models in Preference-based Reward Modeling: Foundations, Theory, and Alternatives
LLM Post-Training, RLHF, PPO, DPO, etc
Add a description, image, and links to the llmalignment topic page so that developers can more easily learn about it.
To associate your repository with the llmalignment topic, visit your repo's landing page and select "manage topics."
| Back | FazBrowse Home | New Git URL |