AI 모델 인텔리전스 플랫폼
← 벤치마크

코딩 벤치마크 최고 AI 모델

코딩 벤치마크 점수에 따른 AI 모델 순위.

순위모델공급자점수데이터셋버전날짜
1Claude 3.5 SonnetAnthropic92humanevalofficial-20242024-06-20
2Mistral Large 2Mistral92humanevalofficial-20242024-07-24
3GPT-4oOpenAI90.2humanevalofficial-20242024-05-13
4Llama 3.1 405BMeta89humanevalofficial-20242024-07-23
5Llama 3.3 70BMeta88.4humanevalofficial-20242024-12-06
6GPT-4o miniOpenAI87humanevalofficial-20242024-07-18
7Qwen2.5 72BAlibaba85.9humanevalofficial-20242024-09-19
8Claude 3 OpusAnthropic84.9humanevalofficial-20242024-03-04
9Gemini 1.5 ProGoogle84.1humanevaltechreport-20242024-05-14
10DeepSeek V3DeepSeek82.6humanevalpaper-20242024-12-26
11GLM-4Zhipu80.1humanevalofficial-20242024-01-16
12Claude 3 HaikuAnthropic75.9humanevalofficial-20242024-03-04
13Gemini 1.5 FlashGoogle71.7humanevaltechreport-20242024-05-14

Source

데이터셋
humaneval
버전
official-2024
날짜
2024-06-20

Benchmark

Related Resources

Use Cases