#API pricing
5 articles tagged API pricing. At twelve articles this topic starts lifting the whole site in search.
Timeline
-
Anthropic's Claude Fable 5.1 — same sticker price, cache reads cut 75%
-
What prompt caching is — from the second call, your input costs one-fortieth
-
What an AI token is — how much text $10 per million tokens actually buys
-
GPT-5.6 Sol API price cut (August 21) — output falls from $30 to $20 per million tokens
All articles
-
Tech · 4 min readAnthropic's Claude Fable 5.1 — same sticker price, cache reads cut 75%
Price — input $10 and output $50 unchanged; cache reads $1.00 → $0.25 (−75%)
-
Tech · 4 min readWhat prompt caching is — from the second call, your input costs one-fortieth
Mechanism — an identical prefix yields identical intermediate values, which can be stored
-
Tech · 4 min readWhat an AI token is — how much text $10 per million tokens actually buys
Unit — a token is a word fragment, not a word, and counts differ sharply by language
-
Tech · 3 min readGPT-5.6 Sol API price cut (August 21) — output falls from $30 to $20 per million tokens
Price — input $5→$4, output $30→$20, cached input $0.50→$0.40 per million tokens. The 33% cut to output is the largest
-
Tech · 5 min readWhat AI token pricing is — why output costs about 5× input
Billing is per token, with separate input and output rates — output typically costs four to five times input