Llama3.1-8B-Instruct Benchmark Details
Llama3.1-8B-Instruct currently shows benchmark results led by MBPP (18 / 70, score 69.40), GSM8K (21 / 70, score 82.40), HumanEval (37 / 101, score 66.50).
Benchmark Results
Llama3.1-8B-Instruct
Benchmark Results
Coding and Software Engineer
2 evaluationsBenchmark / mode
Score
Rank/total
Writing and Creative Capabilities
1 evaluationsBenchmark / mode
Score
Rank/total