HumanEval pass@10 is an AI benchmark used to evaluate model capabilities. Review its overview, metrics, official resources, and model leaderboard results on DataLearnerAI.
Browse the latest scores, model modes, release dates, and parameter sizes for HumanEval pass@10.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
| Rank | Model | License | |||
|---|---|---|---|---|---|
![]() CodeLLaMA-Python-34B Standard Mode | 82.80 | 2023-08-24 | 34B | Free Commercial | |
![]() PanGu-Coder2 Standard Mode | 79.55 | 2023-07-27 | 15B | Closed | |
![]() CodeLLaMA-Python-13B Standard Mode | 77.40 | 2023-08-24 | 13B | Free Commercial | |
4 | ![]() CodeLLaMA-Instruct-34B Standard Mode | 77.20 | 2023-08-24 | 34B | Free Commercial |
5 | ![]() CodeLLaMA-34B Standard Mode | 76.80 | 2023-08-24 | 34B | Free Commercial |
6 | ![]() WizardCoder-15B-V1.0 Standard Mode | 73.32 | 2023-06-14 | 15B | Free Commercial |
7 | ![]() CodeLLaMA-Instruct-13B Standard Mode | 71.60 | 2023-08-24 | 13B | Free Commercial |
8 | ![]() CodeLLaMA-Python-7B Standard Mode | 70.30 | 2023-08-24 | 7B | Free Commercial |
9 | ![]() CodeLLaMA-13B Standard Mode | 69.40 | 2023-08-24 | 13B | Free Commercial |
10 | ![]() CodeLLaMA-Instruct-7B Standard Mode | 64.30 | 2023-08-24 | 7B | Free Commercial |
11 | ![]() CodeGeeX2-6B Standard Mode | 62.60 | 2023-07-25 | 6B | Commercial |
12 | ![]() CodeLLaMA-7B Standard Mode | 59.60 | 2023-08-24 | 7B | Free Commercial |
13 | ![]() LLaMA2 70B Standard Mode | 59.40 | 2023-07-18 | 70B | Free Commercial |
14 | ![]() LLaMA2 34B Standard Mode | 47.00 | 2023-07-18 | 34B | Free Commercial |
15 | ![]() Codex Standard Mode | 46.81 | 2021-08-10 | 175B | Closed |
16 | ![]() CodeGeeX Standard Mode | 39.57 | 2022-09-30 | 13B | Closed |
17 | ![]() LLaMA2 13B Standard Mode | 34.80 | 2023-07-18 | 13B | Free Commercial |
18 | ![]() LLaMA2 7B Standard Mode | 25.20 | 2023-07-18 | 7B | Free Commercial |