Running on your device · 0 bytes uploaded LLM API cost estimator Tokenize your prompt with the model's own encoding, add the reply you expect, and see the exact dollar figure at dated vendor prices. exact BigInt arithmetic · prices effective 2026-08-26 — re-check the vendor page before you budgetEstimateClearModel $ 1.25 in / $ 10 out per 1M · $0.125 cached in Output tokens per requestRequests You cannot know output tokens in advance — enter the reply length you are budgeting for. Set the inputs, then press Estimate. - Input tokens are counted exactly, with the selected model's encoding. Output tokens are your number — the tool cannot predict a reply's length. - Prices are a dated snapshot of the vendor's published per-1M rates (standard tier; documented long-context tiers applied above their boundary). Costs are exact BigInt arithmetic on those prices — no float rounding anywhere in a dollar figure. - Not modeled, on purpose: cached-input discounts on your billed line (shown as a separate reference), batch rates, tool-call framing, and per-request platform fees where a provider charges them. A fake "efficiency" factor would just be lying with decimals. - Prices change — the effective date ships with every result. Re-check the vendor page before committing a budget. - gpt-5.6-sol$ 4 / $ 20 - gpt-5.6-terra$ 2 / $ 12 - gpt-5.6-luna$ 0.2 / $ 1.2 - gpt-5.6-cyber$ 12.5 / $ 75 - gpt-5.5$ 5 / $ 30 - gpt-5.5-pro$ 30 / $ 180 - gpt-5.4$ 2.5 / $ 15 - gpt-5.4-mini$ 0.75 / $ 4.5 - gpt-5.4-nano$ 0.2 / $ 1.25 - gpt-5.4-pro$ 30 / $ 180 - gpt-5.3-codex$ 1.75 / $ 14 - gpt-5.2$ 1.75 / $ 14 - gpt-5.2-pro$ 21 / $ 168 - gpt-5.1$ 1.25 / $ 10 - gpt-5$ 1.25 / $ 10 - gpt-5-mini$ 0.25 / $ 2 - gpt-5-nano$ 0.05 / $ 0.4 - gpt-5-pro$ 15 / $ 120 - gpt-4.1$ 2 / $ 8 - gpt-4.1-mini$ 0.4 / $ 1.6 - gpt-4.1-nano$ 0.1 / $ 0.4 - gpt-4o$ 2.5 / $ 10 - gpt-4o-mini$ 0.15 / $ 0.6 - o3$ 2 / $ 8 - o3-mini$ 1.1 / $ 4.4 - o3-pro$ 20 / $ 80 - o4-mini$ 1.1 / $ 4.4 - o1$ 15 / $ 60 - chat-latest$ 5 / $ 30 - gpt-4-turbo$ 10 / $ 30 - gpt-3.5-turbo$ 0.5 / $ 1.5 Nothing in yet Paste, drop, or type to begin. Everything stays on this device. nothing estimates until you press Estimate Next door in AI - Context fit (/ai/context-fit) - Tokenizer playground (/ai/tokenizer) All AI tools (/ai) Canonical HTML: https://nutter.tools/ai/cost-estimator Markdown version: https://nutter.tools/ai/cost-estimator/index.md Plain-text version: https://nutter.tools/ai/cost-estimator/index.txt Agent index: https://nutter.tools/llms.txt