WeirdML v3 is an AI benchmark used to evaluate model capabilities. Review its overview, metrics, official resources, and model leaderboard results on DataLearnerAI.
Browse the latest scores, model modes, release dates, and parameter sizes for WeirdML v3.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
| Rank | Model | License | |||
|---|---|---|---|---|---|
— | ![]() GPT-6 Astra Thinking Level · Extra HighTools | 0.4222 | 2026-09-03 | Unknown | Closed |
— | ![]() Claude Fable 5.1 Thinking Level · Extra HighTools | 0.2592 | 2026-09-01 | Unknown | Closed |
— | ![]() GPT-5.6 Sol Thinking Level · Extra HighTools | 0.1556 | 2026-06-26 | Unknown | Closed |
— | ![]() Gemini 3.8 Flash Thinking Level · HighTools | 0.0738 | 2026-09-02 | Unknown | Closed |
— | ![]() DeepSeek-V4.1-Flash Thinking Level · HighTools | 0.0605 | 2026-09-10 | 552B | Free Commercial |