Business Pill 54 · Paying for work, not for seats
Usage-based pricing means you pay for how much you use, not for how many people have access. The video shows why AI agents push vendors toward it.
A short explainer video, under a minute. The company and its seat counts are an illustration.
At a glance
The Arithmetic of the Squeeze
This publication's own arithmetic: ten seats at thirty dollars is three hundred dollars, and five seats at thirty dollars is a hundred and fifty.
What Usage-Based Pricing Is
So vendors change the meter. It is called usage-based pricing: you pay for how much you use, not for how many people have access.
Why It Matters
The video turns this into one question, and it is about the busiest month: what is the bill then?
The Short Answer
A company buys software for ten employees: ten seats at thirty dollars each. Then it brings in one AI agent that does the work of five of them.
Next year, it needs five seats. The vendor’s software is doing more work than ever, and earning half as much.
The Arithmetic of the Squeeze
This publication’s own arithmetic: ten seats at thirty dollars is three hundred dollars, and five seats at thirty dollars is a hundred and fifty. The video gives no billing period for the thirty dollars.
The seat count fell by half, so the bill fell by half, even though the software now does more work.

What Usage-Based Pricing Is
So vendors change the meter. It is called usage-based pricing: you pay for how much you use, not for how many people have access.
Per Seat Versus Per Use
Per seat, the bill is easy to predict, but it follows headcount.
Per use, the bill follows the work, but it is harder to forecast.
Why It Matters
The video turns this into one question, and it is about the busiest month: what is the bill then?
Seats count people. Usage counts work. And agents work without a seat.
The One Question to Ask
- What does this bill come to in our busiest month?
Related Frameworks
More Business Pills
- The Memory Wall: Why a Faster AI Chip Is Not Faster AI
- Tokens per Watt: What an AI Data Centre Actually Produces
- Why AI Can’t Be Both Instant and Cheap: Latency vs Throughput
- The Model and the Harness: Why Same-Model Products Differ
- Context, Not Capability: Why a Smart AI Model Gives Poor Answers
- Discardable Software: When Code Is Cheap Enough to Throw Away
- Human in the Loop vs Human on the Loop: Supervising AI Agents
- The AI Audit Problem: When Making Work Is Cheaper Than Checking It
- AI Evaluation as Acceptance Test: How to Know It’s Good Enough
- Extensibility Is Control: Who Holds the Power in an AI Product
- RLHF Explained: How an AI Model Learns What People Prefer
- Pretraining Explained: How an AI Model Learns Before Anyone Teaches It
- Fine-Tuning Explained: How to Adapt a General AI Model to One Job
- Tokens Explained: The Unit AI Reads, Writes and Bills In
- Embeddings Explained: How a Machine Compares Meaning
- RAG Explained: How a Model Answers From Your Documents
- AI Hallucination Explained: Why Models State False Things
- Distillation Explained: How a Small Model Learns From a Large One
- Reasoning Models Explained: What Changes When AI Thinks First
- Tool Use Explained: How an AI Model Goes From Text to Action
- Capex and Depreciation Explained: Why Chip Lifetime Drives AI Profits
- Run-Rate Revenue Explained: What an AI Company’s Number Means
- Backlog Explained: Revenue That Is Signed but Not Yet Earned
- Switching Costs Explained: Why It Is Hard to Leave an AI Supplier
- Temperature Explained: The Dial That Sets How Predictable an AI Model Is
- Guardrails Explained: Rules Enforced Around an AI Model, Not Inside It
- Benchmarks Explained: Why a Test Score Is Not Your Own Result
- Agent Memory Explained: How an AI Assistant Remembers You Between Chats
- MCP Explained: The Shared Plug That Connects AI Models to Tools
- Quantization Explained: How a Large AI Model Is Shrunk to Fit
- Parameters Explained: Where an AI Model’s Knowledge Lives
- Scaling Laws Explained: Why AI Labs Keep Building Bigger
- Synthetic Data Explained: When a Model Writes Its Own Training Examples
- Transformers and Attention Explained: How AI Reads Every Word at Once









