Skip to main content
Back to Course 101
5/12
Unit 5 of 12
Unit 05 — Phase 01

This Stuff Isn't Free

How It Actually Works

Listen to this unit

Read-aloud narrates the unit in a natural voice, paragraph by paragraph. It needs an account.

In short

Training a frontier AI model costs tens to hundreds of millions of dollars. Running it costs fractions of a cent per query — but at scale, that adds up to billions.

Socratic Mode

The Hook

Every time you send a message to Claude or ChatGPT, a computer somewhere — probably in one of many filled with thousands of specialized chips running at full blast — does a massive amount of work to generate your answer. That building uses as much electricity as a small town. The chips inside it cost more than most houses. The engineers maintaining it get paid very well.

Then you get your response in three seconds, for free.

Something doesn't add up. If this costs so much, why isn't anyone charging you? And if someone is paying for it, who — and why?

Video — The Math That Doesn't Add UpWatch on YouTube

The Core Concept

Let's talk about money. Actual money.

Training GPT-4 — just the training, before anyone used it — cost roughly $78 million in alone. Google's Gemini Ultra cost an estimated $191 million. These numbers are growing fast: costs are roughly doubling every year. By 2027, a single training run could exceed $1 billion.

But training is a one-time cost. The ongoing cost — running the model every time someone asks a question — is called . Every message you send gets processed by thousands of specialized chips called , and each query costs a fraction of a cent. That sounds small, but multiply it by hundreds of millions of users sending multiple messages a day, and the numbers get enormous.

OpenAI is on track to lose roughly $14 billion in 2026, burning around $17 billion in cash. Anthropic, growing faster but spending heavily too, burned about $5.6 billion last year. These companies are spending far more money running AI than they're making from it. Sequoia Capital, one of Silicon Valley's most respected venture firms, published an analysis called "AI's $600B Question" that asked: the AI industry is spending enough on Nvidia chips alone to generate $600 billion in annual revenue — but actual AI revenue is a fraction of that. Where's the money coming from, and when does it start paying off?

Explained your way: the AI rewrites this idea around something you already know.

The answer, for now, is investors. Companies like Microsoft, Google, and Amazon are pouring billions into AI companies betting that the technology will eventually justify the cost. It's a gamble — a very large, very expensive gamble.

That's half the answer. The other half is you.

The idea is older than the internet. In 1973 the video piece Television Delivers People, by Richard Serra and Carlota Fay Schoolman, said it flatly: "You are the product of t.v." The phrasing you've probably heard — if you're not paying for it, you're the product being sold — was posted by Andrew Lewis to MetaFilter on 26 August 2010, and close variants were circulating on Usenet by 1999.

With AI assistants the trade is literal: your conversations can become training data for the next model. The providers genuinely differ, and the difference lives in the defaults.

On personal accounts the default leans toward sharing. Google says Keep Activity is on by default for anyone 18 or over. OpenAI ships "Improve the model for everyone" switched on for Free, Plus and Pro. Anthropic was opt-in for consumers until August 2025, then moved Free, Pro and Max to opt-out. All three run the opposite default for business, education and API customers, who are not trained on unless they ask to be — the account paying a bill gets the stricter setting. And opting out has a floor: Google keeps human-reviewed chats up to three years even after you delete your activity, and Anthropic still uses chats flagged for safety review.

These switches govern what happens next. They don't unsend what you already sent.

The Other Price: What You Pay in Data
Unit 05 · This Stuff Isn't Free

The calculator above prices your usage in their dollars. This one prices it in your data. Set the sliders to your own habits.

Messages you send a day20
Months you have been using it12 mo
How personal the material isWork plus daily life
Chat history saved

New chats are kept on the provider's servers until you delete them — and deleting is a request to a company, not an undo on your own machine.

Chats used to improve the model

What you send from here on can shape future models. Once material is trained in, there is no thread back to your individual chat to pull it out.

The record you have already handed over
7,200 messages
20 a day × 30 days × 12 months — your own numbers, multiplied out

Both switches above govern what happens next. Neither one reaches these.

What that record could show
Rounded profile

Your work, your routine, who you deal with, what you are trying to fix this month.

How long it stays around

As long as the provider's policy allows, which they can change without asking you.

Training is the part with no delete button: a model that has learned from your material cannot unlearn one person's share of it.

This is an illustrative model, not a measurement. The message count is your own arithmetic and nothing more. The two bands are a teaching device — the band steps up once when the record passes 10,000 messages, and that rule was written to make the point visible, not fitted to any study. No provider is being described here; read the retention and training terms of the assistant you actually use.

Then the second currency: attention. OpenAI put ChatGPT at more than 900 million weekly users in February 2026. A habit that size is an asset whether or not it's billed today.

None of this is hidden. It's just unread.

In the copyright case that the New York Times and other news plaintiffs brought against OpenAI, the fight was over consumer ChatGPT conversation logs. The plaintiffs first sought 120 million of them. OpenAI countered that a sample of 20 million was more than enough, and the plaintiffs took the deal. Then, in October 2025, OpenAI changed course and offered to hand over only the conversations a keyword search tied to the plaintiffs' articles.

In January 2026, US District Judge Sidney Stein refused, affirming Magistrate Judge Ona Wang's order to produce the whole 20-million-chat sample, and unfiltered, as "neither clearly erroneous nor contrary to law." The court took the privacy interest seriously and called it sincere, but held it adequately protected by the reduced sample size, OpenAI's de-identification, and the protective order already in the case. It also noted that users had voluntarily submitted those conversations to OpenAI in the first place.

Nobody in that sample chose to be in it. That's the shape of the trade: data you hand over becomes a thing that exists, and it is then subject to other people's decisions — a policy change, a court order, a change of ownership.

Knowledge check
Your personal chats with a major AI assistant may be used to train the next model. Based on this section, the most accurate statement is:
Knowledge checks save to your account.

In June 2024, Sequoia Capital partner David Cahn published a simple calculation that shook the AI industry. Nvidia was on track to sell $150 billion in AI chips per year. The companies buying those chips needed to earn roughly 4x their hardware cost to break even (accounting for energy, buildings, staff, and profit). That meant the AI industry needed to generate $600 billion in annual revenue to justify its spending.

At the time, the entire AI industry's revenue — including all of OpenAI, Google's AI products, and everyone else — was roughly $100 billion. The gap between spending and revenue was half a trillion dollars.

Cahn's conclusion wasn't that AI was a bubble. It was that either AI revenue needs to grow 6x, or a lot of companies are going to lose a lot of money. That question hasn't been answered yet.

Here's why this matters to you: the economics explain almost everything about the AI industry that otherwise seems confusing.

Why are there different models? Because bigger models cost more to run. Claude Opus is smarter than Claude Sonnet, but it costs roughly 5-10x more per query. Companies offer cheaper, smaller models for simple tasks and expensive, powerful models for hard ones.

Why do free tiers have limits? Because every message costs money. When you hit a usage cap on the free version of ChatGPT or Claude, it's because the company literally can't afford to give away unlimited access to their most powerful model.

Why is AI improving so fast? Partly because the massive investment is funding enormous research teams and massive compute. The race is fueled by a fear that whoever builds the best AI first will dominate the next era of technology.

Why is AI concentrated in a few companies? Because training requires billions of dollars in hardware and energy. This isn't a garage startup situation. You need data centers, chip allocations from Nvidia, and billions in funding. As of 2025, the top AI labs — OpenAI, Anthropic, Google DeepMind, Meta AI — are all backed by trillion-dollar companies.

Data centers globally consumed roughly 415 terawatt-hours of electricity in 2024 — about 1.5% of all electricity generated on Earth, on the International Energy Agency's numbers. That is every data center, not just the AI ones: AI is a fast-growing part of the total rather than the whole of it. The IEA projects the total will more than double by 2030, reaching the equivalent of Japan's entire electricity consumption.

A single text question is small: Google measured an average Gemini prompt at about 0.24 watt-hours — roughly what a television uses in nine seconds — and independent analysts put ChatGPT in the same range. You may still see the claim that one AI query costs ten times a web search; that figure compared a worst-case AI estimate against a 2009 number for search, and it no longer holds.

The weight is in training and in scale. Epoch AI estimates that training xAI's Grok 4 in 2025 took about 310 gigawatt-hours — roughly what 30,000 American homes use in a year — and the power behind these training runs has been roughly doubling annually. Water follows the same shape: an average Gemini text prompt uses about 0.26 millilitres, five drops, but Microsoft's total water use jumped 34% in 2022 as AI training ramped up, to about 6.4 million cubic metres.

AI isn't just an abstract technology. It has a physical footprint that shows up in power grids, water bills, and carbon emissions.

Understanding the economics doesn't make AI less useful. But it does make you a more informed user.

AI Cost Calculator
Unit 05 · This Stuff Isn't Free
Pricing (Claude Sonnet)
Input: $3.00 / 1M tokensOutput: $15.00 / 1M tokens

Adjust the sliders to see how compute costs scale. Every message you send costs real money.

Queries per user per day50
Input tokens per query (your prompt)500
Output tokens per query (AI response)800
Users1,000
$675.00
Per day
$20.3K
Per month
$243.0K
Per year
$20.25
Per user / month

When a tool is "free," you're not the customer — you're the product, or you're the bet. And when you're paying, you should understand what you're paying for.

Knowledge check
AI companies like OpenAI and Anthropic are currently losing billions of dollars per year. The primary reason is:
Knowledge checks save to your account.

Live Demo

Step 1: Go to openai.com/pricing and anthropic.com/pricing. Look at the per-token costs for different models.

Step 2: Open any AI tool. Write a medium-length prompt (about 200 words) and get a response. Estimate how many tokens that exchange used (rough rule: 1 token ≈ ¾ of a word, so a 200-word prompt + 500-word response ≈ 930 tokens).

Step 3: Calculate what that conversation cost the company. Using GPT-4o pricing as an example: input tokens might cost $2.50 per million, output tokens $10 per million. Your conversation might cost a fraction of a cent — but multiply by a million users doing the same thing.

Step 4: Calculate how much it would cost to run 10,000 queries per day on the most expensive model. Then on the cheapest model. The difference in cost directly shows why companies offer different model tiers.

Step 5: Look at the Visual Capitalist infographic:

Interactive tool — Training Costs of AI Models Over Time — Visual Capitalist

Training Costs of AI Models Over Time — Visual Capitalist

This tool opens in a new tab where you can interact with it directly.

Launch Tool

The original Transformer model in 2017 cost $930 to train. GPT-4 cost $78 million. That's an 84,000x increase in seven years.

Take a conversation you've had with AI today. Count the rough word count of your inputs and the AI's outputs. Convert to tokens (÷ 0.75). Look up the per-token price for the model you used. Multiply. Now multiply that by the number of messages you send per week. That's what your usage costs someone — and it's why "free" tiers can't last forever.

Why This Matters

The economics of AI determine who gets access and who doesn't. If the most powerful AI costs $200/month, that creates a divide between people who can afford it and people who can't. If AI training requires billions of dollars, only a handful of companies will build frontier models — and those companies will decide what the models can and can't do.

This isn't theoretical. It's happening now. The gap between free-tier AI and paid AI is already significant, and it's growing. Understanding the economics helps you make smarter decisions about which tools to use, why they work the way they do, and what's really going on when something is offered "for free."

It also sets up the next unit perfectly: if every token costs money and every query burns compute, then the way you structure your input to AI isn't just about getting better answers — it's about efficiency. The people who communicate well with AI aren't just getting better output. They're getting better output per dollar.

Knowledge check
The fact that training a frontier AI model costs hundreds of millions of dollars has which direct consequence?
Knowledge checks save to your account.

The Challenge

AI Budget Planner

30 minutesHands-on

Imagine you're starting a small business and want to use AI for three tasks: writing marketing copy, answering customer questions, and summarizing daily reports.

  1. Research pricing — Go to openai.com/pricing and anthropic.com/pricing. Pick a model for each task.
  2. Estimate token usage — How many tokens would each task use daily? (Remember: 1 token ≈ ¾ of a word.)
  3. Calculate monthly cost — Multiply daily token usage × per-token price × 30 days for each task.
  4. Optimize — Redo the calculation using the cheapest model that could still do each job. How much did you save?
  5. Write a recommendation — Which model for which task, and why? When is the expensive model worth it?
Success criteria: You can explain why different tasks justify different models, estimate real costs, and make a reasonable cost-optimization argument.
Submitting your work needs an account.

Key Takeaways

  1. 1Training a frontier AI model costs tens to hundreds of millions of dollars. Running it costs fractions of a cent per query — but at scale, that adds up to billions.
  2. 2The economics explain why models come in tiers, why free versions have limits, and why the industry is concentrated in a few companies.
  3. 3Data centers use about 1.5% of global electricity (IEA, 2024), projected to double by 2030. AI is one growing part of that, and it has a real physical footprint.
  4. 4When AI is free, someone else is paying. Understanding who — and why — makes you a more informed user.

The Rabbit Hole

Type: Article Title: AI's $600B Question — Sequoia Capital URL: https://sequoiacap.com/article/ais-600b-question/ Description: Short, sharp analysis asking whether the AI industry's spending can ever be justified by revenue. Written by one of the most respected venture capital firms in the world.

Explore Further

TypeTitleURLDescription
ArticleSequoia Capital, "AI's $600B Question" (2024)sequoiacap.com/article/ais-600b-qu…The analysis that questioned whether AI spending can be justified
ReportIEA, "Energy and AI" (2025)iea.org/reports/energy-and-…International Energy Agency's assessment of AI's energy footprint
ArticleEpoch AI, "How Much Does It Cost to Train Frontier AI Models?"epoch.ai/blog/how-much-does-…Detailed breakdown of training costs across major models
VisualVisual Capitalist, "Training Costs of AI Models Over Time"visualcapitalist.com/training-costs-of-a…From $930 (2017 Transformer) to $78M (GPT-4) — the cost explosion visualized
AudioNPR Planet Money, "What AI Data Centers Are Doing to Your Electric Bill" (2025)npr.org/2025/12/19/nx-s1-56…How AI infrastructure affects local energy costs
ToolOpenAI Pricingopenai.com/pricingCurrent per-token costs for all OpenAI models
ToolAnthropic Pricinganthropic.com/pricingCurrent per-token costs for all Claude models
ReportStanford HAI AI Index Report 2025hai.stanford.edu/news/ai-index-2025-…Comprehensive annual snapshot of the AI industry including economics
BookAgrawal, Gans & Goldfarb, Power and Prediction (2022)How AI economics will reshape industries and decision-making
ArticleQuote Investigator, "Quote Origin: You're Not the Customer; You're the Product"quoteinvestigator.com/2017/07/16/productTraces the line from Serra and Schoolman's 1973 video piece through Usenet to Andrew Lewis's 2010 post
DocsGoogle, "Gemini Apps Privacy Hub" (updated August 2026)support.google.com/gemini/answer/13594…Keep Activity's default, the 18-month auto-delete, human review, and the three-year reviewed-chat copy
DocsOpenAI Help Center, "How your data is used to improve model performance"help.openai.com/en/articles/5722486…The "Improve the model for everyone" control, and why personal and business accounts default differently
DocsAnthropic, "Updates to Consumer Terms and Privacy Policy" (2025)anthropic.com/news/updates-to-our…The August 2025 move to opt-out, five-year retention, and which plans are excluded
DocsAnthropic Privacy Center, "How do you use personal data in model training?"privacy.claude.com/en/articles/1002354…The Model Improvement setting, Incognito chats, and five-year versus 30-day retention
ArticleABA Journal, "ChatGPT creator must turn over 20M chat logs in copyright litigation" (2026)abajournal.com/news/article/chatgp…The January 2026 ruling ordering production of de-identified consumer chat logs

Last updated: May 31, 2026. Economics and creativity statistics were updated to current figures.

Track your progress

Marking a unit complete, the Prove It check and your place in the course all need an account. The reading stays free.