Same standard of work. About 10.0× the energy.
GPT-5.4 scores higher than Gemini 2.5 Pro (Mar 2025) on Epoch's Capabilities Index — 156.4 against 145.59, across 58 benchmarks covering coding, agents, long context and writing — while using about 10.0× less energy per answer. In blind head-to-head voting, people rate them 1470 and 1458 respectively.
| GPT-5.4 | Gemini 2.5 Pro (Mar 2025) | |
|---|---|---|
| Energy per answer | 7.94 Wh | 79.3 Wh |
| Same carbon as driving | 15 m | 146 m |
| Energy rank | 91 of 134 | 126 of 134 |
| Capability (science questions) | 93.3 | 83.8 |
| Epoch Capabilities Index | 156.4 | 145.59 |
| Human vote rating | 1470 | 1458 |
| Votes cast | 60,614 | 122,589 |
| Confidence | Day-one estimate | Day-one estimate |
| Released | 2026-03-05 | 2025-03-31 |
Energy is calculated the same way for every model on this site, from public data and calibrated against laboratory measurement — accurate to roughly a factor of two to three, which is why we only ever state a difference of 5× or more. Capability is a science-question test; the Capabilities Index is a composite over 58 benchmarks; the vote rating comes from blind human comparisons. The full method is here.