Docs
Search
Ctrl K
Models
Chat
Compare
Providers
Apps
Rankings
MMLU Chat Benchmark Leaderboard | Phaseo
MMLU Chat
MMLU Chat
AI benchmark results and model performance
Summary
▼
Summary
▼
Type: percentage
General
Recorded Results
1
Average Score
80.58%
Score Range
80.58% - 80.58%
Leading Model (lowest score)
80.58% - Llama 3.1 Nemotron 70B Instruct
Scores Over Time
Individual benchmark scores plotted by date.
Models Using This Benchmark
Organisation
Model
Reported
Top Score
Info
Self Reported
Source
Nvidia
Llama 3.1 Nemotron 70B Instruct
01 Oct 2024
80.58%
-
Yes
Source
Sign Up
Sign Up