The Debrief

Free AI Still Has a Bill

13 min read

The Short Version

The AI market is having a very normal software moment:

Everyone is giving things away.

How generous.

Google is now explaining Gemini access in terms of usage limits and compute resources, not just a clean number of prompts. Wired notes that the same subscription can behave differently depending on whether you use text, image generation, video, research, or heavier reasoning features.

OpenAI, meanwhile, just published a guide on managing AI investments in the agentic era. That is the grown-up version of the same story: once AI moves from chat into workflows, the question stops being "can we try it?" and becomes "what does this cost when it becomes part of the company?"

And on the startup side, the market is full of credits, programs, and discounted access. OpenAI has a startup program with API credits. Anthropic has Claude for Startups. Google Cloud advertises startup credits that can be used to build on its infrastructure and AI stack.

None of that is bad.

Free credits are useful. Subsidized experiments are useful. Usage limits are more honest than pretending every AI request costs the same.

But the useful takeaway is simple:

Free AI is not free.

It is a customer acquisition strategy, a product-learning loop, and sometimes a way to make a team build around one vendor before it has developed the discipline to measure the bill.

The dangerous part is not that AI companies charge money.

Of course they do.

The dangerous part is that the real bill usually arrives after the workflow has already changed.

Update: the bill is becoming the product

The bill is no longer theoretical.

On August 11, OpenAI said ChatGPT Ads had launched in the United Kingdom, Mexico, Brazil, Japan, and South Korea, after earlier pilots. Ads are for Free and Go users, not Plus, Pro, Business, Enterprise, or Education. OpenAI says ads do not influence ChatGPT's answers, are labeled and visually separated, and advertisers do not get access to chats, memories, history, or personal details.

Good.

Also: this is the bill.

A week earlier, OpenAI said it was making GPT-5.6 Luna the default for Free and Go users, adding unlimited text chats, and giving users a Think button for harder questions, while keeping limits on files, images, and other tools. The free product gets more capable. The expensive stuff still has meters. The ad layer expands. The paid tiers stay cleaner.

That is a very familiar bargain.

The free version gets broader access.

The paid version gets fewer interruptions, more capacity, and cleaner control.

The platform gets usage, data about product behavior, ad inventory, upgrade pressure, and a better story about democratizing access.

Very generous. The subsidy has discovered a media plan.

Ads make trust part of pricing

OpenAI is trying to draw a bright line:

The answer is the answer.

The ad is the ad.

The advertiser does not see the conversation.

That line matters.

But in a conversational product, the line is harder to maintain than it is in a search results page or a social feed. People use ChatGPT while deciding things: what to buy, where to travel, how to study, what tool to use, which vendor to compare, which recipe to cook, which health question to ask a real doctor about.

An ad beside that workflow is not just a banner.

It is commercial context inside a decision assistant.

OpenAI says ads are not eligible near sensitive or regulated topics like health, mental health, or politics, and that it will build protections against scams and narrow targeting. Necessary. The hard part is that categories blur in real use. A meal plan can become a health question. A budgeting chat can become financial vulnerability. A school assignment can reveal a minor. A travel search can reveal family situation, income, location, or religion.

The ad product will not be judged by the principle page.

It will be judged by edge cases.

The free tier is becoming an economic router

The same week, OpenAI also introduced Premium seats for ChatGPT Business: 5x more usage than Standard, no five-hour usage limit, and a higher monthly price. That is the enterprise version of the same meter.

Free users see the ad-funded bargain.

Power users see the capacity bargain.

Businesses see the seat-tier bargain.

Developers see the token and tool-call bargain.

All of this points to the same product reality: AI pricing is becoming a router. The system routes people by workload, tolerance for ads, need for privacy, need for capacity, need for admin control, and willingness to pay when the task becomes important.

That is not bad.

It is the only way the economics can work.

But it means builders should stop saying "we use ChatGPT" as if that describes one stable object. Which plan? Which model? Which tools? Which quotas? Which data terms? Which ads? Which admin controls? Which country? Which user age? Which sensitive-topic rules?

The AI product is no longer just the model.

It is the commercial policy around the model.

Update: the agent price war is about workflow margin

The meter moved again.

On August 21, OpenAI updated its GPT-5.6 launch post to say it had dropped API and credit pricing for GPT-5.6 Sol by more than 20% for the next three months. That follows its July 30 cuts for Luna and Terra.

Google is using the same play from the other side. Gemini 3.7 Flash, introduced August 13, is pitched as Google's most intelligent workhorse model yet for coding and agents, with an introductory price through December 31 of $0.75 per million input tokens and $3.75 per million output tokens. Google says that is half the original 3.6 Flash price, and explicitly ties the price/performance story to scaling production-ready agents. It also says Spark, its 24/7 personal agent in the Gemini app, is moving onto 3.7 Flash for Workspace tool use.

This is not a coupon story.

It is agent strategy.

A chatbot answer can be subsidized as a user-acquisition cost. An agent workflow is different. It plans, reads, calls tools, retries, asks for clarification, summarizes, and sometimes hands the result back to a human reviewer. One user request can become a small bundle of model calls and tool calls.

So the real unit is no longer "one prompt."

It is a completed workflow.

That is why the price cuts matter. The winning model is not just the smartest model in a benchmark table. It is the model that is good enough, cheap enough, and predictable enough to leave inside a real process without turning every successful automation into a margin problem.

For builders, the trap is obvious: do not build your business model around promotional economics. Google's introductory Gemini 3.7 Flash pricing has an end date. OpenAI's Sol cut is explicitly temporary. Treat those prices as a distribution signal, not a durable cost basis.

The practical checklist changes:

Measure cost per completed workflow, not cost per call.

Track retries and clarification loops.

Route easy steps to cheaper models when quality allows it.

Include human review time in the bill.

Write down what happens when the promotional price expires.

If the answer is "we will figure it out later," congratulations, you have found the real pricing page.

What buyers should take from this

For normal users, the practical lesson is simple:

If a free AI product gets much better, ask what pays for it.

Maybe the answer is ads.

Maybe it is upgrade conversion.

Maybe it is a bundled cloud deal.

Maybe it is enterprise cross-subsidy.

Maybe it is investor money that wants usage now and margin later.

For businesses, the lesson is stricter:

Do not let employees build important workflows on a consumer tier whose economics, ads, limits, and data controls are not designed for your risk model. And do not assume that a paid plan removes every meter. It may just move the meter into seats, workspace credits, tool limits, model routing, or premium capacity.

Free AI can still be useful.

It just stopped pretending the bill was imaginary.

Credits are distribution, not charity

In software, "free" often means "we are still deciding who captures the value."

AI makes that bargain more complicated because the marginal unit is not just a seat. It can be tokens, images, video generations, tool calls, retrieval, memory, long-running agents, evaluations, latency guarantees, compliance features, and human review time.

That is why credits matter.

Credits let a startup test a model without worrying about every experiment. They also make one provider the default during the most formative part of the product. The first architecture is often the stickiest one. The first eval set becomes the scorecard. The first model's quirks become product behavior. The first successful workflow becomes the thing customers expect.

By the time procurement shows up, the "free" part may be over.

But the dependency remains.

This is not a conspiracy. It is distribution.

Cloud companies did this. SaaS companies did this. Developer platforms did this. AI labs are doing it with a more expensive raw material and a more intimate role inside the workflow.

The difference is that AI does not just host the product.

It can shape the product.

The meter is getting more honest

Google's Gemini usage page is a small but useful signal because it makes the abstraction visible.

A prompt is not a prompt.

Asking a model to summarize an email is not the same economic object as generating video, running deep research, using a reasoning-heavy model, or letting an agent take many tool-backed steps.

That sounds obvious until a team tries to budget for it.

The old consumer mental model was:

I pay for the AI subscription, then I use AI.

The new model is:

I pay for access to a changing basket of capabilities, and each capability consumes a different amount of scarce compute.

That is a much better description of reality.

It is also less comfortable.

Because it means AI pricing will keep moving toward meters, quotas, priority lanes, model tiers, and "fair use" language that depends on what the model is actually doing.

For consumers, that means the product may feel less magical when the meter appears.

For companies, it means the finance and engineering teams need to understand the same dashboard.

Good luck, everyone.

The data question is not one question

There is another bill hiding behind the credit conversation:

data terms.

Teams often ask one vague question: "Do they train on our data?"

That is necessary.

It is not sufficient.

OpenAI says on its enterprise privacy page that it does not train on business data by default. Anthropic says in its privacy center that it does not use commercial Claude inputs or outputs to train its models by default, unless the customer opts in or provides feedback in a way covered by its policy.

Those controls matter.

But buyers still need to ask more precise questions:

  • What counts as customer content?
  • What counts as metadata?
  • What is retained?
  • Can logs be disabled or shortened?
  • Are prompts, outputs, files, tool traces, evals, or feedback treated differently?
  • Can the company export usage history?
  • Can the team prove which model handled which request?
  • What changes when a product moves from a free program to a paid enterprise plan?

The answer may be completely reasonable.

But "reasonable" is not the same as "understood."

Most teams do not get hurt because a vendor secretly does something cartoonishly evil.

They get hurt because nobody wrote down the boring operational details before the workflow became important.

ROI arrives after the pilot

The awkward truth is that free credits can make AI look better than it is.

Not because the model is fake.

Because the pilot is fake.

A pilot usually ignores half the cost:

  • integration work
  • prompt and eval maintenance
  • security review
  • human approval loops
  • failed runs
  • context management
  • vendor switching costs
  • internal support
  • model upgrades that change behavior
  • governance for who can use what

That is fine when the goal is exploration.

It is dangerous when the pilot becomes the business case.

Agentic AI makes this worse because agents turn one user request into many model calls. A workflow that feels like one action may include planning, file search, web access, tool calls, retries, code execution, summarization, and review.

The user sees one button.

The invoice sees a small expedition.

That is why OpenAI's investment-management framing is more interesting than the usual launch blog. The market is moving from "try AI" to "operate AI." That means usage controls, internal chargebacks, approval gates, evals, data policies, and model routing are no longer enterprise theater.

They are the product.

The buyer checklist

If you are a startup, free credits are great.

Take them.

Just do not let them do the thinking for you.

Before a team builds a real workflow around subsidized AI, it should know:

  • What happens when the credits expire?
  • Which features are included in the free or startup tier?
  • Are data protections the same across free, startup, team, and enterprise plans?
  • Can the workflow run on another model without rewriting the product?
  • Are prompts, evals, traces, and usage logs portable?
  • What is the cost per successful task, not per prompt?
  • Which tasks need the expensive model?
  • Which tasks can use a cheaper model or local system?
  • Who can approve higher-cost runs?
  • What user-facing behavior changes when the quota is hit?

That last one is underrated.

Users do not care that your quota model is economically sophisticated.

They care that the feature worked yesterday and now it does not.

Free is a phase, not a strategy

The AI market is still in land-grab mode.

Model companies want usage. Cloud companies want workloads. Startups want runway. Enterprises want productivity without admitting they do not yet know how to measure it.

So the market will keep producing discounts, bundles, pilots, credits, free tiers, and confusingly generous offers.

Use them.

But treat them as a temporary pricing environment, not a permanent law of physics.

The real AI bill is not only the invoice.

It is the meter, the data policy, the workflow dependency, the switching cost, the eval burden, the human review loop, and the moment a subsidized experiment becomes infrastructure.

Free AI can be a very good deal.

Just make sure you know what you are buying before the price appears.