AI Cost Calculator
What AI actually costs to run in production — not the sticker price.
Model list prices are the tip of the iceberg. Retries and context growth routinely push real production cost to 5–20× the per-token rate on the pricing page. RAGRetrieval-Augmented Generation (RAG) retrieves relevant passages at query time and feeds them into an LLM's context. overhead and eval runs add more on top. Model your own usage and see where the multiplier comes from.
A production cost estimate you can defend to finance, with the assumptions shown.
Model your cost →