Investors debate agent pricing and unit economics in plain language.
What Is an AI Agent? (pricing and economics discussion)
a16z May 2025
Listen on Spotify open.spotify.com →Token prices for equivalent quality fall roughly 10x a year, yet agent bills keep rising because agents burn 50 to 500 times the tokens of a chat: every step re-sends the whole history. A simple task can cost a fraction of a cent while a long agent session runs several dollars. The levers that matter: prompt caching (up to 90% off repeated context), routing easy steps to cheap models, trimming history, and measuring cost per completed task rather than per token.
18 resources.
Investors debate agent pricing and unit economics in plain language.
a16z May 2025
Listen on Spotify open.spotify.com →The single chart that explains why what is unaffordable today is cheap next year.
a16z (Guido Appenzeller) Nov 2024
Open a16z.com →The headline numbers in thread form for a two-minute read.
a16z Nov 2024
Open x.com →Names the paradox every founder hits: cheaper tokens, bigger bills.
NavyaAI 2026
Open navyaai.com →Explains the quadratic context growth that quietly multiplies agent costs.
LeanOps 2026
Open leanopstech.com →How large buyers model agent economics, useful when selling to them.
EY 2025
Open ey.com →Concrete per-task cost ranges from cheap chatbot turns to $5+ agent runs.
Cowork 2026
Open cowork.ink →Walks through the exact math of an agent whose unit economics collapse.
Klaus Hofenbitzer 2025
Open medium.com →The feature that cuts up to 90% off repeated context, table stakes for agents.
Anthropic Aug 2024
Open anthropic.com →One page comparing how each provider's caching discount actually works.
PromptHub 2024
Open prompthub.us →Why KV-cache hit rate is the one production metric that decides agent margins.
Yichao 'Peak' Ji (Manus) Jul 2025
Open manus.im →Shows 72% of production cost sits outside the model invoice.
Optimum Partners 2026
Open optimumpartners.com →The full stack of costs behind a token price, for founders selling AI products.
Introl 2025
Open introl.com →A hands-on implementation guide with before-and-after cost numbers.
DigitalOcean 2025
Open digitalocean.com →Saves you from the premature self-host-a-model cost trap.
ML6 2025
Open ml6.eu →Why the price curve should keep falling, and what that means for your roadmap.
Weighty Thoughts 2025
Open weightythoughts.com →Estimate a task's cost across models before you commit to one.
TokenCalculator 2025
Open tokencalculator.ai →Live per-model pricing so your cost model never goes stale.
UsagePricing 2026
Open usagepricing.com →The same ground, over in Build the product, our Starting Up track.