Electricity Bench

Grok 4.6, released today, ties Fable 5 for the vibe-coding lead

post

By the Electricity Bench editorial team · Drafted with AI, reviewed and approved by a human editor

xAI's Grok 4.6 shipped and graded the same day: it ties Claude Fable 5 for the vibe-coding lead, doubling Grok 4.5's task time and quota burn along the way.

xAI released Grok 4.6 today, and it went straight onto the bench: graded the same day, as a successor to Grok 4.5, on the Week 33 suites the rest of the field just ran. These are its day-one numbers, not the vendor's.

What the results show

The bottom line: Grok 4.6 is clearly the stronger model, but the strength comes at the cost of the one edge its predecessor had. Fast and good enough for everyday work was Grok 4.5's niche; twice the wait and twice the burn puts Grok 4.6 in a straight fight with the top tier - one it would be a surprise to see it win outside benchmarks.

The succession is stitched as one line on Trends, so the step shows in context. Grades are comparable within the same suite version - explore the complete Week 33 results.