Show HN: CostPerPrompt – Live AI API pricing and real-workload cost calculators

ahmed_hassan71 pts0 comments

CostPerPrompt — Live AI API Pricing & LLM Cost Calculators Skip to content

What will AI actually cost you?

Live pricing for 232+ models, refreshed automatically — plus calculators that<br>turn token prices into real answers: what your chatbot, agent, or API workload will cost per<br>month.

API cost calculator Chatbot cost calculator All 232 model prices GPU rental prices

Every cost question, answered<br>$_ API Cost Calculator Any workload, compared across all 232 models — with caching & batch discounts. 💬 Chatbot Cost Simulates real conversations — growing history, resent context, cache hits. 🤖 Agent Cost Multi-step loops, tool schemas, retries — see why agents cost 10–30× more than you think. 📚 RAG Cost Indexing, retrieval and generation priced separately — spoiler: embeddings are pennies. 🎙️ Voice AI Cost STT + LLM + TTS per minute, per call, per month — the full three-part bill. 🖥️ GPU Rental Pricing H100, A100, RTX 4090 across 10 providers — the same chip at a 5× price spread. 🖼️ Image API Pricing $0.003 to $0.17 per image — three market tiers and a batch calculator. 🔢 Token Counter Paste text, get tokens and cost on any model — 100% in your browser. ⚡ Cost-Cutting Guides Prompt caching (−90% input) and the Batch API (−50% everything) explained.

Flagship model pricing<br>per 1M tokens · updated 2026-08-02<br>Model Vendor Input / 1M Output / 1M Context GPT-5.6 Sol OpenAI $5.00 $30.00 1.1M GPT-5.5 OpenAI $5.00 $30.00 1.1M GPT-5.4 OpenAI $2.50 $15.00 1.1M Claude Fable 5 Anthropic $10.00 $50.00 1M Claude Opus 5 Anthropic $5.00 $25.00 1M Claude Sonnet 5 Anthropic $2.00 $10.00 1M Kimi K3 Moonshot (Kimi) $3.00 $15.00 1M DeepSeek V4 Pro DeepSeek $0.435 $0.87 1M DeepSeek V4 Flash 0731 DeepSeek $0.09 $0.18 1M Gemini 3.1 Pro Preview Google $2.00 $12.00 1M Gemini 3.6 Flash Google $1.50 $7.50 1M Grok 4.5 xAI $2.00 $6.00 500K Grok 4.3 xAI $1.25 $2.50 1M GLM 5.2 Z.ai (GLM) $0.4186 $1.32 1M<br>See the full table of 232 models →

advertisement

How AI API pricing works — the 60-second version

Every major AI provider bills the same way: you pay per token (roughly ¾ of<br>a word), with separate rates for input (what you send) and<br>output (what the model writes back). Output is usually 3–5× more expensive<br>than input. A model listed at $5 / $25 per million tokens costs $5 for every million tokens<br>you send and $25 for every million it generates.

Two discounts change the math dramatically: prompt caching cuts repeated<br>input costs by up to 90% (critical for chatbots that resend conversation history), and<br>batch processing takes ~50% off when you can wait for results. Our<br>calculators account for both — most "how much will this cost" articles don't, which is why<br>their estimates run 2–3× too high or too low.

advertisement

© 2026 CostPerPrompt. Some outbound links may be affiliate links that earn us a commission<br>at no extra cost to you.

cost pricing model input calculators calculator

Related Articles