Electricity Bench

antigravity-cli + gemini-3.1-pro

antigravity-cli harness · Google AI Pro

plan
Google AI Pro
price
no public price recorded
allowance
100 provider-percent per week
isolation
dedicated-account
usage method
reported-delta

Dedicated benchmark Google account ebbench.runner@gmail.com on eb-runner; no interactive or IDE use shares the AI Pro weekly Gemini window.

Solves up to tier 5 on real-world@v3 - the highest difficulty rung with a solved task
12345

Tier 1: localized single-file bug · tier 3: cause in a different module than the symptom · tier 5: problems that took a human hours. A cleared rung solved every task on it; a partial rung solved some.

overall pending Two distinct suite families are required for an overall grade.

Per-suite results

D
real-world@v3
47%
plan cost per run
16.8% of Google AI Pro
latency med / p95
150525 ms / 441288 ms
reliability
100%
error rate
0%

Run on antigravity-cli 1.1.10, model gemini-3.1-pro (--effort high), Aug 5, 2026.

One run of this suite consumes 16.8% of Google AI Pro - about 5.9 runs per week. Measured per suite run (reported-delta); an absolute cost on this plan, not comparable with another agent's plan. Measured on a benchmark-only account (Dedicated benchmark Google account ebbench.runner@gmail.com on eb-runner; no interactive or IDE use shares the AI Pro weekly Gemini window.).

Trajectory

Every scored run of this agent, oldest first. Grades compare only within a suite version, so each version gets its own line; while a suite is being retired both versions run, and the superseded one is dashed. The filled point is the run the scorecard shows today. All agents over time →

real-world · 2 runs · latest Aug 5, 2026
Capability
0%50%100%Aug 4Aug 5antigravity-cli + gemini-3.1-pro · v3 · Aug 4, 2026 · 100% · grade Aantigravity-cli + gemini-3.1-pro · v3 · Aug 5, 2026 · 47% · grade D0%50%100%Aug 4Aug 5antigravity-cli + gemini-3.1-pro · v3 · Aug 4, 2026 · 100% · grade Aantigravity-cli + gemini-3.1-pro · v3 · Aug 5, 2026 · 47% · grade D
Median latency
110157 ms285085 ms460012 msAug 4Aug 5antigravity-cli + gemini-3.1-pro · v3 · Aug 4, 2026 · 419644 ms · grade Aantigravity-cli + gemini-3.1-pro · v3 · Aug 5, 2026 · 150525 ms · grade D110157 ms285085 ms460012 msAug 4Aug 5antigravity-cli + gemini-3.1-pro · v3 · Aug 4, 2026 · 419644 ms · grade Aantigravity-cli + gemini-3.1-pro · v3 · Aug 5, 2026 · 150525 ms · grade D
Plan cost per run - a share of this agent's own plan, not a cross-agent number
0.0%9.6%19.3%Aug 4Aug 5antigravity-cli + gemini-3.1-pro · v3 · Aug 4, 2026 · 0.6% · grade Aantigravity-cli + gemini-3.1-pro · v3 · Aug 5, 2026 · 16.8% · grade D0.0%9.6%19.3%Aug 4Aug 5antigravity-cli + gemini-3.1-pro · v3 · Aug 4, 2026 · 0.6% · grade Aantigravity-cli + gemini-3.1-pro · v3 · Aug 5, 2026 · 16.8% · grade D

Aug 4, 2026: A on v3 · Aug 5, 2026: D on v3

Compare

Transcript excerpts

Outputs are shown only when they can be separated safely from the private suite. Scores and run metadata remain visible when an output is withheld.

real-world@v3
Task#TurnsLatencyOutput
b2r5h81-259070 mswithheldPrivate-suite content
c8g2m41-72034 mswithheldPrivate-suite content
d9b4x61-85099 mswithheldPrivate-suite content
h5n2v71-583001 mswithheldPrivate-suite content
h7w3j51-151555 mswithheldPrivate-suite content
k4f9t71-189205 mswithheldPrivate-suite content
m6t3q81-67454 mswithheldPrivate-suite content
p1x9k41-101157 mswithheldPrivate-suite content
q8k3j61-380554 mswithheldPrivate-suite content
r3v8m51-207724 mswithheldPrivate-suite content
s2j7f41-205391 mswithheldPrivate-suite content
t5s8n21-71062 mswithheldPrivate-suite content
w4j7q21-150525 mswithheldPrivate-suite content
y3p7k11-106506 mswithheldPrivate-suite content
z6q1v91-79984 mswithheldPrivate-suite content