All articles
The Calcaas Blog

All articles

Page 3 of 11

OpenAI Is Gaining Ground on Anthropic with Business Users: What That Says About Vendor Lock-In
LLM Economics
Aug 25, 20264 min read

OpenAI Is Gaining Ground on Anthropic with Business Users: What That Says About Vendor Lock-In

New data shows businesses swinging between OpenAI and Anthropic as each ships new models, a signal that enterprise AI spend is far less sticky than most vendor contracts assume.

Stripe Paid $7.5B for OpenRouter. It's Not About the Singularity, It's About Owning Your AI Spend
Company News
Aug 21, 20264 min read

Stripe Paid $7.5B for OpenRouter. It's Not About the Singularity, It's About Owning Your AI Spend

Stripe's real reason for buying OpenRouter for a reported $7.5 billion isn't a philosophical bet on AI merging with humanity, it's a bet on owning the layer where developers see, route, and pay for every token they use.

AI Compute Just Got a Price Index. Here's What It Means for Your LLM Margins
LLM Economics
Aug 21, 20264 min read

AI Compute Just Got a Price Index. Here's What It Means for Your LLM Margins

Silicon Data just raised a $30 million Series A to become the reference price for GPU rental, but a standardized compute market won't automatically make the token prices you pay OpenAI, Anthropic, or Google any more predictable.

Why a 500% Spike in Memory Prices Could Quietly Raise Your AI Costs
LLM Economics
Aug 20, 20264 min read

Why a 500% Spike in Memory Prices Could Quietly Raise Your AI Costs

DRAM and HBM prices have climbed roughly 500% in 12 months, a hardware shock that sits underneath every GPU and inference bill and could slow or reverse the token-price deflation AI builders have gotten used to.

Model Routing: Why Cost, Not Capability, Is Driving Enterprise AI Strategy in 2026
LLM Economics
Aug 20, 20265 min read

Model Routing: Why Cost, Not Capability, Is Driving Enterprise AI Strategy in 2026

Model routing, sending each AI request to the cheapest model that can handle it, has become a core enterprise cost lever because frontier model prices are rising faster than most teams' AI budgets.

What OpenRouter's $7B Sale Teaches Builders About Margin at the Routing Layer
LLM Economics
Aug 19, 20264 min read

What OpenRouter's $7B Sale Teaches Builders About Margin at the Routing Layer

Stripe's reported $7 billion purchase of OpenRouter values a token-routing business at roughly 50x revenue, and the real story is how a company with no GPUs of its own posted a 70% gross margin.

Anthropic's $65B Revenue Run Rate: What It Means for Your Token Pricing
LLM Economics
Aug 19, 20264 min read

Anthropic's $65B Revenue Run Rate: What It Means for Your Token Pricing

Anthropic's annualized revenue run rate hit $65 billion in July, a sevenfold jump in seven months, and the growth curve says more about enterprise usage-based pricing than about model quality.

Stripe's Reported $7B OpenRouter Deal: What It Says About LLM Cost Routing
LLM Economics
Aug 18, 20266 min read

Stripe's Reported $7B OpenRouter Deal: What It Says About LLM Cost Routing

Stripe has reportedly agreed to buy OpenRouter for more than $7 billion, roughly five times the $1.3 billion valuation the model-routing startup raised at in May 2026, which prices model choice as core financial infrastructure rather than a developer convenience.

Airtable Sold at 2.7x ARR. Here Is What Seat-Based Pricing Is Worth Now
Pricing Strategy
Aug 18, 20266 min read

Airtable Sold at 2.7x ARR. Here Is What Seat-Based Pricing Is Worth Now

Airtable sold its core platform to Bending Spoons for $1.285 billion, roughly 2.7 times its $480 million in ARR, while still growing more than 20% a year, which hands the market a public clearing price for seat-based software in the agent era.

LLM VRAM Sizing: Your Context Limit Is a Pricing Decision
Founder Guides
Aug 17, 20266 min read

LLM VRAM Sizing: Your Context Limit Is a Pricing Decision

Llama 4 Scout needs about 231 GB of VRAM at a 1,024-token context and about 3,671 GB at its advertised 10M-token window, so the context length you configure sets your hardware bill more than the model size does.

vLLM vs TensorRT-LLM: The 12% Cost Gap and When It Is Worth Paying For
LLM Economics
Aug 17, 20265 min read

vLLM vs TensorRT-LLM: The 12% Cost Gap and When It Is Worth Paying For

On the same H100, TensorRT-LLM serves a million output tokens for about $0.66 against vLLM's $0.75, and whether that 12% is worth having depends almost entirely on how often you change models.

AI Agent Hosting Cost: Price Per Completed Task, Not Per GPU-Hour
LLM Economics
Aug 17, 20266 min read

AI Agent Hosting Cost: Price Per Completed Task, Not Per GPU-Hour

On the same GPU at the same hourly rate, a ten-step agent loop costs roughly 14x a single chat call per unit of finished work, and most of that gap is context you re-send rather than work you do.

The Margin Memo

Pricing math, in your inbox.

One short note a week on AI pricing, token economics, and margin. No spam, unsubscribe anytime.