Sam Altman claims an “average” ChatGPT query uses ~0.34 watt-hours, or 1+ second of oven use, and ~0.000085 gallons of water, or “one fifteenth of a teaspoon”
Jay Peters / The Verge :
Context & Ripple Effects
OpenAI's compute costs were characterized as "eye-watering" when ChatGPT first scaled, making per-query resource use a consequential operating metric rather than a purely environmental one. More recent coverage argued that daily ChatGPT use remains a small share of an individual's electricity footprint, while water estimates have varied sharply by model and data-center location, including location-specific GPT-4 water estimates.
Altman's figure adds a company-supplied benchmark to a debate previously driven largely by external estimates. Its usefulness depends on what OpenAI includes in an "average" query and how that average maps to different model workloads.
First-order effects
- OpenAI now has a public, simple per-query energy and water claim that users, customers, and critics can use when discussing ChatGPT's footprint.
- The claim shifts immediate scrutiny toward methodology: query mix, model choice, data-center operations, and whether the reported average is comparable with prior estimates.
Second-order effects
- Rival AI providers face more pressure to publish similarly legible usage metrics; Google has separately reported a median Gemini text-prompt energy figure, though differently defined measurements are not automatically comparable.
- Enterprise buyers and sustainability teams gain a starting point for evaluating AI workloads, but may demand workload-specific disclosures rather than a single consumer-query average.
Third-order effects
- If major model providers standardize transparent, comparable inference reporting, environmental performance could become a competitive dimension alongside model quality, price, and latency.
- The pattern points toward AI resource reporting moving from broad company-level disclosures to product-level metrics, although common definitions and independent verification would determine whether those figures can support meaningful comparisons.
The trend: AI providers are increasingly translating inference infrastructure costs into per-prompt environmental metrics as assistants become a more routine computing interface.