The Calcaas Blog

Cost-first thinking for AI pricing.

Pricing frameworks, LLM economics, product updates, and founder playbooks from the team building the AI cost calculator.

128 articles
Moonshot AI's 300 Billion Daily Tokens: What It Signals About the Model Market
LLM EconomicsLatest

Moonshot AI's 300 Billion Daily Tokens: What It Signals About the Model Market

A Chinese lab targeting $2B in revenue is notable, but the more useful number for founders is the sheer token throughput reshaping provider economics.

Sep 15, 2026 · 3 min read
What Nvidia's Vera Rubin Platform Means for Your Token Costs
LLM Economics
Sep 15, 20264 min read

What Nvidia's Vera Rubin Platform Means for Your Token Costs

Better performance-per-watt on the hardware side eventually shows up as lower token prices, but the timeline and who captures the savings first is worth understanding.

Why Hasn't SaaS Pricing Raced to the Bottom? The Uncomfortable Answer for Founders
Pricing Strategy
Sep 15, 20263 min read

Why Hasn't SaaS Pricing Raced to the Bottom? The Uncomfortable Answer for Founders

It's not because competition is weak, it's because price is rarely the thing customers are actually optimizing for, and founders who compete on price alone usually lose anyway.

The Real Total Cost of an LLM Gateway: Why the License Fee Is Never the Whole Story
LLM Economics
Sep 15, 20263 min read

The Real Total Cost of an LLM Gateway: Why the License Fee Is Never the Whole Story

A vendor comparison of LiteLLM's enterprise pricing against a managed alternative shows why a $250/month license number hides most of the actual 3-year cost.

What Prompt Caching Actually Saves: The Token Math Behind LLM Gateways
LLM Economics
Sep 15, 20264 min read

What Prompt Caching Actually Saves: The Token Math Behind LLM Gateways

Once caching is on, output tokens dominate your bill, so the input-token discounts everyone advertises matter less than you'd think.

Is Subscription Pricing Still Right for Your SaaS, or Should You Switch to Usage-Based?
Pricing Strategy
Sep 15, 20264 min read

Is Subscription Pricing Still Right for Your SaaS, or Should You Switch to Usage-Based?

Neither wins by default: the right model depends on whether your costs scale with usage, and for AI features, they almost always do.

Does AI Pricing Really Need to Fall 90%? What the Token-Cost Math Actually Shows
LLM Economics
Sep 15, 20265 min read

Does AI Pricing Really Need to Fall 90%? What the Token-Cost Math Actually Shows

No, prices don't need to crash 90% overnight, but the trend underneath the headline is real, and it changes how you should price any AI feature you ship this year.

If AI Coding Is Not Winner-Take-All, Gross Margin Decides Who Survives
LLM Economics
Sep 9, 20265 min read

If AI Coding Is Not Winner-Take-All, Gross Margin Decides Who Survives

Cognition raised at a $48B valuation, which TechCrunch reads as investors betting that AI coding is not a winner-take-all market, and in a market with many survivors the winners are usually decided by unit economics rather than by being first.

How to Cut Agent Token Spend by 80% Without Changing Models
Founder Guides
Sep 9, 20266 min read

How to Cut Agent Token Spend by 80% Without Changing Models

A practitioner reports cutting token spend on dynamic agent workflows by roughly 80% through prompt and workflow restructuring alone, which points at an uncomfortable truth: most agent spend is not on the task, it is on context you resend by default.

Token Volume Is Up 25x and Mid-Tier Models Do 90% of the Job. Your Pricing Has to Move.
Pricing Strategy
Sep 9, 20266 min read

Token Volume Is Up 25x and Mid-Tier Models Do 90% of the Job. Your Pricing Has to Move.

Tom's Hardware reports token volume exploding 25-fold while mid-tier models deliver roughly 90% of flagship capability at one-sixth the cost, which means the default choice of flagship-everything is now a margin decision rather than a quality one.

A 33x Price Gap for the Same Model Is Not a Market. It Is a Tax on Not Checking.
LLM Economics
Sep 9, 20265 min read

A 33x Price Gap for the Same Model Is Not a Market. It Is a Tax on Not Checking.

A developer who tracks LLM API list prices daily reports a 33x cost gap for the identical model depending on which provider you call, which means most of what teams call their AI cost is really a procurement decision they never made.

The Same Model, 14 Providers: Why Your LLM Bill Depends on Where You Run It
LLM Economics
Sep 9, 20266 min read

The Same Model, 14 Providers: Why Your LLM Bill Depends on Where You Run It

Choosing a model sets your capability ceiling, but choosing the provider that serves it sets your actual bill, because price, throughput and prompt caching all differ between endpoints running the identical weights.

AI Agent Cost Per Hour: Why the $6 Headline Is a Floor, Not a Bill
LLM Economics
Sep 6, 20265 min read

AI Agent Cost Per Hour: Why the $6 Headline Is a Floor, Not a Bill

An AI agent quoted at under $6 an hour is priced on output tokens at a single stream, so your real bill depends on input volume, how many agents you run in parallel, and how often work has to be redone.

The Margin Memo

Pricing math, in your inbox.

One short note a week on AI pricing, token economics, and margin. No spam, unsubscribe anytime.