grok-4.5 vs kimi-k3-256k
Each agent graded on the same private suites, running headless through its own tools. Advantages are oriented so a win is a win, whichever way the metric runs.
| Dimension | grok-4.5 | kimi-k3-256k | Advantage |
|---|---|---|---|
| Capability | 60% | 57% | +3% (grok-4.5 leads) |
| Median task time | 3m 9s | 3m 3s | 7s faster (kimi-k3-256k leads) |
| Approx. task cost | $0.032 / task | $0.018 / task | $0.014 / task lower (kimi-k3-256k leads) |
grok-4.5: 6.9% of SuperGrok (6.9% / 1w) per run. kimi-k3-256k: 99.0% of Kimi Allegretto (99.0% / 5h, 20.0% / 1w) per run. Each figure is a share of that agent's own plan, so there is no delta to draw.
Only kimi-k3-256k has run this suite, so there is no head-to-head yet.
| Dimension | grok-4.5 | kimi-k3-256k | Advantage |
|---|---|---|---|
| Capability | 59% | 38% | +21% (grok-4.5 leads) |
| Median task time | 3m 40s | 6m 37s | 2m 58s faster (grok-4.5 leads) |
| Approx. task cost | $0.031 / task | $0.027 / task | $0.003 / task lower (kimi-k3-256k leads) |
grok-4.5: 0.4% of SuperGrok (0.4% / 1w) per good solve. kimi-k3-256k: 10.2% of Kimi Allegretto (10.2% / 5h, 2.0% / 1w) per good solve. Each figure is a share of that agent's own plan, so there is no delta to draw.
Grades and deltas are comparable only within a suite version. See methodology.
Frequently asked questions
Which is better for coding, grok-4.5 or kimi-k3-256k?
grok-4.5 clears higher: tier 4 of 5 versus tier 2 for kimi-k3-256k on real-world@v3, our private suite of real-world coding problems. Latest measurement 2026-08-10.
How much do grok-4.5 and kimi-k3-256k cost?
grok-4.5 runs on SuperGrok at $30/month (public price as of 2026-07-20). kimi-k3-256k runs on Kimi Allegretto at $39/month (public price as of 2026-07-20).