Skip to content

LLM Benchmark Hub

Model evidence, daily

Models What's New Trends Benchmarks Skill Maps
How It Works Guesswork Metrics
LLM Benchmark Hub
Loading data...

DramaBench — leaderboard

Metric: Overall Score (%). Source: dramabench.pages.dev. 8 models tracked.

Top models

#ModelScore
1GPT-5.296.04
2GLM-4.693.04
3Qwen 3 Max91.73
4Kimi K2 (Thinking)91.65
5DeepSeek V3.290.99
6Claude Opus 4.586.7
7Gemini 3 Pro (Preview)83.9
8MiniMax-M280.94

Interactive version: aibenchmarks.dev/benchmark?slug=dramabench · How the rankings work · Data refreshed daily, snapshot 2026-07-20.

Built by Mikhail Doroshenko — AI researcher, co-author of Humanity’s Last Exam. Independent project; no affiliation with any AI lab.