AI Model Pricing & Token Economics
Articles on token pricing tiers, cost structures, and economic models of AI inference for GPT-4o, Claude, Mistral, and other frontier models.
Reassessing RL Scaling's Economic Impact: Cost Burdens Are More Transient Than They Appear
Toby Ord argues RL scaling primarily increases inference costs, creating a persistent economic burden. The counter-analysis finds that while RL scaling does increase inference costs, the cost to reach a given capability level falls rapidly over time due to algorithmic improvements, model distillation, hardware advances, and more efficient reasoning. The RL scaling data is thin, and the 10,000x compute estimate is uncertain. Inference cost reductions of 5-10x per year make the burden more transie
LLM Cost Breakdown: How Sapient Trained a Foundation Model for $1,500
Training a foundation LLM from scratch typically costs millions and requires internet-scale data — which is why most enterprises don't bother. Publicis Sapient says it has found a cheaper path, potentially opening the door for organizations to build custom foundation models without the massive infrastructure investment that has historically limited them to off-the-shelf API access.
The AI Bill You Didn’t Know You Were Running Up
Autonomous AI agents are quietly turning corporate inference costs into a runaway line item—and here’s how the smartest teams are getting control.