Essay26

Essay · Published Aug 12, 2026

AI Pricing Is Becoming UX

GPT-5.6 turns capability into a price ladder, while Codex Spark turns speed into a separate button. AI pricing is becoming part of the interface.

A field-guide drawing of a pine-lined mountain trail

Codex is starting to feel less like one tool and more like a control panel: four model buttons, two usage buckets, and three token prices. The comparison matters more than the branding.

The model matrix

ModelContextImagesAPI input / output per 1MUsage on my $200 Pro 20× planI would use it forMain downside
GPT-5.6 Sol1,050,000Yes$5 / $30Shared bucketLong migrations, ambiguous debuggingHighest price
GPT-5.6 Terra1,050,000Yes$2.50 / $15Shared bucketLarge repositories, everyday workMiddle choice still requires guessing
GPT-5.6 Luna1,050,000Yes$1 / $6Shared bucketCheaper multimodal tasksUses the same shared allowance
GPT-5.3 Codex Spark128,000NoNot finalizedSeparate preview bucketSmall edits, refactors, quick CSS workOlder model, 8.2× less context, no screenshots

Spark’s separate bucket is the interesting part. During the research preview, it does not consume my shared 20× allowance, so I can preserve GPT-5.6 usage for larger jobs. Pricing and availability are still provisional.

The separate allowance is also a budget-design decision. A pooled cap encourages substitution; a separate bucket encourages experimentation. I saw the organizational version when my company combined several AI budgets into one pool.

What I would choose

TaskFirst choiceWhy it fitsUsage consequenceWhat would make me switch
Rename, focused refactor, quick CSS editCodex SparkFast feedback for a small text-only changePreserves the shared 20× bucketThe change spreads across the repository
Fix a UI from a screenshot or mockupGPT-5.6 LunaImage input at the lowest GPT-5.6 API priceUses the shared bucketThe debugging requires more reasoning
Explore an unfamiliar large repositoryGPT-5.6 Terra1.05M context with middle-tier pricingUses the shared bucketThe task is either trivial or unusually difficult
Plan a migration across many filesGPT-5.6 SolMaximum reasoning headroom and full contextUses the shared bucket fastestA smaller model proves sufficient
Continue after the shared limit gets tightCodex SparkSeparate research-preview allowanceLeaves the shared bucket aloneI need images or more than 128k context

The trade is not subtle: Spark is fast and quota-friendly, but text-only with roughly one-eighth of GPT-5.6’s context. A screenshot or a large repository can send me straight back to another button.

AI pricing is becoming UX because choosing a model now means choosing which limitation I want to hit.

Quota shape matters as much as token price. My Grok Build trial made that concrete: a weekly ceiling changed whether I was willing to begin a task at all.