LLM Benchmark Snapshot: August 30, 2026
Weekly frozen photo of six LLM leaderboards: Opus 5 max takes the text arena, Fable 5 holds LiveBench, gpt-5 keeps Aider coding, and the sources still disagree.
LLM Benchmark Snapshot: August 30, 2026
Every Sunday Wikiprompt freezes a photo of the major LLM leaderboards. This is the snapshot for the week ending August 30, 2026 - 14,164 entries across six sources, browsable in full at /benchmarks/snapshot/2026-08-30.
Leaders by source
What changed vs last week
The text-arena crown changed hands inside the same family: claude-opus-5-high (1504.2 last week) was edged out by claude-opus-5-max (1504.7). Statistically a coin flip, editorially a reminder that effort settings now matter as much as model choice at the top of the table.
The sources still disagree
As every week, no two leaderboards crown the same model: LMArena's community Elo favors Opus 5, LiveBench's contamination-free academic sets favor Fable 5, and Aider's real-world coding harness still belongs to gpt-5. Cross-reference before you choose - that is what the comparison view is for.
Browse the full dated table at /benchmarks/snapshot/2026-08-30, or the whole archive at /benchmarks/history.
Related Articles
- Instantâneo de Benchmark de LLM: 30 de agosto de 2026
Aug 31, 2026 · 4 min read
- # LLM基准快照:2026年8月30日
Aug 31, 2026 · 4 min read
- LLM-Benchmark-Momentaufnahme: 30. August 2026
Aug 31, 2026 · 4 min read
- Aperçu des benchmarks LLM : 30 août 2026
Aug 31, 2026 · 4 min read
- # Instantánea de Benchmarks de LLM: 30 de agosto de 2026
Aug 31, 2026 · 4 min read