LLM Benchmark Snapshot: August 23, 2026
The first weekly photo of the LLM leaderboards: LMArena, LiveBench, BenchLM and Aider, frozen on one citable page. Anthropic sweeps three podiums; GPT-5 keeps the coding crown.

LLM Benchmark Snapshot: August 23, 2026
This is the first entry in a new Wikiprompt series: every week we freeze a complete photo of the major LLM leaderboards and publish it as a permanent, citable page. The full tables for this snapshot live at /benchmarks/snapshot/2026-08-23, and every future photo will be indexed at /benchmarks/history.
Why dated snapshots?
Benchmark numbers move quietly: a model climbs, a score is recalculated, and last month's leader silently becomes fourth. By keeping one URL per weekly photo, the record stays public. You can cite "the rankings as of August 23, 2026" and that link will always show exactly that.
What the photo shows
This snapshot covers five independent, open sources, each with its own methodology:
The interesting read is the disagreement: community voting (LMArena), academic-style task suites (LiveBench), aggregated scores (BenchLM) and hands-on coding (Aider) do not crown the same model. That disagreement is exactly why we track many sources instead of one.
Explore it
New photos land every Sunday. When the rankings shift, you will be able to point at exactly when.
Related Articles
- The Best Wan 2.1 Prompts: Open-Source AI Video That Works
Sep 3, 2026 · 7 min read
- Les Meilleurs Prompts Wan 2.1 : Vidéo IA Open-Source Qui Fonctionne
Sep 3, 2026 · 7 min read
- 最佳的Wan 2.1提示词:开源的AI视频,真的能用
Sep 3, 2026 · 7 min read
- As Melhores Prompts do Wan 2.1: Vídeo de IA Open-Source Que Funciona
Sep 3, 2026 · 7 min read
- Las Mejores Instrucciones para Wan 2.1: Video IA de Código Abierto Que Funciona
Sep 3, 2026 · 7 min read