Please reach out to naman_jain@berkeley.edu for questions or feedback on LiveCodeBench. We are also open to collaborations and suggestions for new scenarios to add to the benchmark. Finally, LiveCodeBench provides one axis of LLM coding evaluations and we recommend the following leaderboards for measuring code LM ability on various coding tasks, such as EvalPlus Leaderboard, CruxEval Leaderboard, Chatbot Arena Leaderboard, BigCode Models Leaderboard, InfiCoder-Eval, and TabbyML Leaderboard.
The source code from this website is borrowed from this template!