Day-one estimate
Calculated from an estimated parameter count — the maker hasn't published one. Every model on this site is calculated the
same way, so the comparison holds even where the absolute value carries uncertainty.
Our figures are good to roughly a factor of two to three, which is why we only ever
state a difference of 5× or more.
The figure above is an understatement. GPT-4o generates reasoning text you never see, and that text costs energy to produce. Our calculation counts a fixed answer length, so the real figure is higher — we cannot say by how much, because the amount is not published.
Alibaba · capability 84.8. It is more capable than GPT-4o and uses about 21× less energy per answer: 0.351 Wh against 7.53 Wh.
What OpenAI publishes
Energy cannot be worked out without knowing how big a model is. Of the 31 models
from OpenAI that we track, 24 can be assessed at all — the rest publish
too little for anyone outside the company to calculate anything.