- AI pricing is based on tokens (roughly 3/4 of a word), with separate charges for input and output; output is typically 3-5x more expensive than input.
- Prompt caching can reduce repeated input costs by up to 90%, and batch processing offers ~50% discounts, which many cost estimates ignore.
- A live table shows pricing for 232+ flagship models, including GPT-5.x, Claude 5 variants, DeepSeek V4, Gemini 3.x, Grok 4.x, and GLM 5.2, with costs ranging from $0.09 to $10.00 per million input tokens.
- Calculators are provided to convert token prices into real monthly costs for chatbots, agents, or API workloads.