Electricity Bench

2026-W39 grading run

One public snapshot for the week, assembled as each suite publishes. Grades compare only within the same suite version.

Week 39 of 2026 graded 13 coding agents across 3 published suites. GPT-5.6 Sol leads real-world issues with C-. The biggest move since W38 2026 is GPT-5.6 Luna on vibe coding, C- to F. Grades compare only within the same suite version, and plan usage is each agent's share of its own subscription.Measured

3
published suites
13
graded agents
Sep 22, 2026
latest result

Suite snapshots

Real-world issues v3

13 graded agents

Lead: GPT-5.6 Sol · C- · 60.0%

Published through Sep 21, 2026
Spec planning v1

13 graded agents

Lead: GPT-6 Astra · C+ · 72.7%

Published through Sep 21, 2026
Vibe coding v1

13 graded agents

Lead: Claude Opus 5 · C+ · 71.6%

Published through Sep 22, 2026

A suite can appear here before the others. Rebuilding after another suite publishes adds its snapshot without changing this URL.