100% free

Select your AI cost issue?

Select AI cost issues to filter the AI Cost Calculators you need

Selected: none

🗂️All 96 AI Cost Calculators, organized by what they solve.

Popular tools first, then grouped by problem, then alphabetical for power users.

I’m a:

🔥Popular this week

What other people are running through aicost.ai right now.

Popular
🧮

AI Cost Calculator

The foundational "what does this workload cost?" tool.

Try the calculator →
📖 Guide
Popular
🔤

Token Estimator

Paste text → instant cost across every model.

Try the calculator →
📖 Guide
Popular
🔍

AI Model Finder

Compare every model side-by-side. Filter by context, price, modality.

Try the calculator →
📖 Guide
Popular
✂️

Token Reduction Stack

Stack 6 techniques. See exact savings on YOUR workload.

Try the calculator →
📖 Guide
Popular
🧾

Plan Overage & Tier-Break Calculator

Pay overage or upgrade? Break-even across GitHub, Claude, Cursor, Copilot Studio.

Try the calculator →
📖 Guide
Popular
🪙

Credit Decoder

What is an AI credit worth? Decode credits → dollars → tokens & messages.

Try the calculator →
📖 Guide

🚨My bill is spiking (10)

Audit a runaway bill, find the biggest line items, model what one bad day costs.

🔦

MSP Client AI Exposure Audit

Scores a whole client book for undocumented model routing: allowlist, data class, and who would answer an auditor.

Try the calculator →
📖 Guide
Popular
🔤

Token Estimator

Paste text → instant cost across every model.

Try the calculator →
📖 Guide
💰

Cheapest Model Finder

Lowest-cost model that meets your quality bar.

Try the calculator →
📖 Guide
Popular
✂️

Token Reduction Stack

Stack 6 techniques. See exact savings on YOUR workload.

Try the calculator →
📖 Guide
💹

Margin Calculator

Is your AI feature profitable? Gross margin + grade.

Try the calculator →
📖 Guide
📆

Quarterly Spend Forecaster

The agentic shift. Your CFO is about to be surprised.

Try the calculator →
📖 Guide
⚠️

Overage Forecaster

True monthly cost on hybrid (seat + overage) plans.

Try the calculator →
📖 Guide
Popular
🧾

Plan Overage & Tier-Break Calculator

Pay overage or upgrade? Break-even across GitHub, Claude, Cursor, Copilot Studio.

Try the calculator →
📖 Guide
Popular
📈

Pricing History Explorer

2.5+ years of vendor pricing. 62K+ historical price points.

Try the calculator →
📖 Guide
🏢

Business AI Bill Diagnose

Break a multi-vendor team AI bill into seats, API, and agentic spend — surface leakage, concentration risk, and the savings plays.

Try the calculator →
📖 Guide

📉Cut what I'm spending (36)

Caching, batching, model routing, distillation — stack the savings on what you already run.

🗜️

Context Compaction

Vendors quote the gross reduction. This prices what you keep after the compaction cost and the re-derivation the agent pays back.

Try the calculator →
📖 Guide
⚖️

Model Family Comparison

One workload across OpenAI, Anthropic and Gemini. Headline prices converged; cache economics and the long-context cliff decide it.

Try the calculator →
📖 Guide
🔗

Stacked Savings

Routing 45% plus caching 30% is not 75%. Each lever only works on what the last one left behind.

Try the calculator →
📖 Guide
🚦

Gateway Stage Advisor

Which of four layers a client belongs at, and what the layer above would have to save before it pays for itself.

Try the calculator →
📖 Guide
🛣️

Router and Gateway Comparison

Eliminates on hard requirements, prices the survivors at this client's spend, and shows what each one gives up.

Try the calculator →
📖 Guide
🔦

MSP Client AI Exposure Audit

Scores a whole client book for undocumented model routing: allowlist, data class, and who would answer an auditor.

Try the calculator →
📖 Guide
🔓

Open-Weight Model Suitability

Can this client self-host it, and may they. Four gates in order: data class, licence, hardware, and who supports it.

Try the calculator →
📖 Guide
Popular
🔤

Token Estimator

Paste text → instant cost across every model.

Try the calculator →
📖 Guide
📏

Context Window Cost

At what doc size does your cost double?

Try the calculator →
📖 Guide
💰

Cheapest Model Finder

Lowest-cost model that meets your quality bar.

Try the calculator →
📖 Guide
🧭

Open-Source Model Finder

Browse open-weight & Chinese LLMs (GLM, Qwen, Kimi, DeepSeek, Nemotron) by cost, intelligence & risk.

Try the calculator →
📖 Guide
🌏

Chinese vs Western AI Cost

Compare open-weight / Chinese models against Western frontier models at your workload.

Try the calculator →
📖 Guide
💸

GPU Spot Savings

Savings from spot / preemptible GPUs vs on-demand for self-hosting.

Try the calculator →
📖 Guide
⚖️

Jurisdiction Risk Scorer

Score data-sensitivity x jurisdiction x compliance for open-weight / Chinese models.

Try the calculator →
📖 Guide
🏠

Local vs Cloud Inference

Serverless host vs dedicated GPU rental vs self-host — the 3-way serving cost for open-weight models.

Try the calculator →
📖 Guide
🏭

Open-Weight Hosting Compare

Serverless host comparison plus self-host GPU crossover for open-weight models.

Try the calculator →
📖 Guide
Popular
✂️

Token Reduction Stack

Stack 6 techniques. See exact savings on YOUR workload.

Try the calculator →
📖 Guide
💾

Prompt Cache ROI

Will caching save money or cost money? Break-even hit-rate by vendor.

Try the calculator →
📖 Guide
🔀

Multi-Model Router

Route easy queries to cheap, hard ones to flagship. Tier savings.

Try the calculator →
📖 Guide

Batch vs Realtime

~50% off async. What fraction of your workload tolerates 24h SLA?

Try the calculator →
📖 Guide
⚖️

Buy vs Build

Outcome-priced vendor or build it yourself? Engineering cost included.

Try the calculator →
📖 Guide
🧠

RAG vs Fine-Tuning

Head-to-head cost across both approaches at YOUR query volume.

Try the calculator →
📖 Guide
🧩

Chunking Optimizer

Chunk size vs retrieval quality vs token cost. Find the sweet spot.

Try the calculator →
📖 Guide
🔎

Hybrid Search Cost

Vector + BM25. Twice the cost, often 30%+ better recall.

Try the calculator →
📖 Guide
🎓

Fine-Tuning Cost

Training + inference + break-even vs base. 4 providers.

Try the calculator →
📖 Guide
🖥️

Self-Host Breakeven

Llama / Qwen on rented GPUs vs API. Real ops costs included.

Try the calculator →
📖 Guide
Popular
🧠

Reasoning Token Cost

The hidden thinking-token surcharge — what reasoning tokens add to your bill.

Try the calculator →
📖 Guide
Popular
🔁

Retry Cost

What failed-and-retried API calls quietly cost you each month.

Try the calculator →
📖 Guide
🖥️

Inference Serving Cost

Self-hosted serving economics — the true $/1M tokens on your own GPUs.

Try the calculator →
📖 Guide
🔒

Guardrail / Security Cost

Runtime safety, prompt-injection and PII screening cost, and the percent it adds over your token bill.

Try the calculator →
📖 Guide
🗓️

Annual vs Monthly

Lock-in cost vs commit discount. Quantified.

Try the calculator →
📖 Guide
🛠️

Dev Stack Calculator

Find overlap in your AI dev tools. Stop paying for redundancy.

Try the calculator →
📖 Guide
⚖️

API vs Pro+SDK Breakeven

Direct API or a Claude Pro/Max/Team plan? Models the June-15 Agent SDK carve-out.

Try the calculator →
📖 Guide
💺

AI Seat ROI

Is your coding-assistant seat spend worth it? ROI multiple + break-even hours, live seat prices.

Try the calculator →
📖 Guide
🎧

Support Deflection ROI

AI cost per resolved conversation vs the human cost it avoids — with the human-touch share priced in.

Try the calculator →
📖 Guide
🏗️

Subscription Picker for Builders

Best AI subscription for your team by persona — founder, CTO, FP&A, consultant.

Try the calculator →
📖 Guide

🧮Plan a new workload (59)

Model the workload before you build it. Forecast, scale, and break-even analysis.

⚖️

Model Family Comparison

One workload across OpenAI, Anthropic and Gemini. Headline prices converged; cache economics and the long-context cliff decide it.

Try the calculator →
📖 Guide
🔗

Stacked Savings

Routing 45% plus caching 30% is not 75%. Each lever only works on what the last one left behind.

Try the calculator →
📖 Guide
🚦

Gateway Stage Advisor

Which of four layers a client belongs at, and what the layer above would have to save before it pays for itself.

Try the calculator →
📖 Guide
🛣️

Router and Gateway Comparison

Eliminates on hard requirements, prices the survivors at this client's spend, and shows what each one gives up.

Try the calculator →
📖 Guide
🔎

AI Cost Diligence for VC/PE

Stress-test a startup's claimed AI unit economics against vendor-exact pricing. Verdict + deal-memo.

Try the calculator →
📖 Guide
🔥

Burn Multiple Stress-Test (VC/PE)

Reported vs true burn multiple. Reclassifies subsidised inference back into COGS to expose real gross margin.

Try the calculator →
📖 Guide
🖥

GPU Cost-per-Inference (VC/PE)

API vs self-hosted cost per inference, break-even volume, and whether the claimed GPU fleet can serve the traffic.

Try the calculator →
📖 Guide
📈

Portfolio EBITDA-Lift Estimator (VC/PE)

Turns achievable AI and cloud savings into bps of EBITDA expansion, EV lift at your exit multiple, and payback.

Try the calculator →
📖 Guide
🤖

Humanoid Labor ROI

Honest humanoid payback: realized FTE, effective $/hr vs labor, and the AI-ops layer.

Try the calculator →
📖 Guide
🤖

Robot Training-Data / Teleoperation Cost

Price the demonstration corpus to train a robot policy: total program cost + cost per usable demo.

Try the calculator →
📖 Guide
🤖

Robot Buy vs RaaS Breakeven

CAPEX vs Robotics-as-a-Service: the break-even month, hidden fees, and the AI layer both paths carry.

Try the calculator →
📖 Guide
🤖

AI Robot Fleet Cost

Price the AI layer of an AI-driven robot fleet: cost per successful task.

Try the calculator →
📖 Guide
Popular
🧮

AI Cost Calculator

The foundational "what does this workload cost?" tool.

Try the calculator →
📖 Guide
📏

Context Window Cost

At what doc size does your cost double?

Try the calculator →
📖 Guide
Popular
🔍

AI Model Finder

Compare every model side-by-side. Filter by context, price, modality.

Try the calculator →
📖 Guide
🧭

Open-Source Model Finder

Browse open-weight & Chinese LLMs (GLM, Qwen, Kimi, DeepSeek, Nemotron) by cost, intelligence & risk.

Try the calculator →
📖 Guide
🌏

Chinese vs Western AI Cost

Compare open-weight / Chinese models against Western frontier models at your workload.

Try the calculator →
📖 Guide
🎯

ROI Quick Check

Hours saved vs cost — a fast ROI sanity check on an AI initiative.

Try the calculator →
📖 Guide
🏦

TCO Quick Estimate

Quick total-cost-of-ownership estimate for an AI initiative.

Try the calculator →
📖 Guide
🏛️

TCO Complete

Full total-cost-of-ownership build across the whole AI stack.

Try the calculator →
📖 Guide
📚

Corpus Onboarding Cost

Cost to embed and index a document corpus for RAG.

Try the calculator →
📖 Guide
🧪

Eval & Benchmark Cost

What a model evaluation / benchmark run costs across vendors.

Try the calculator →
📖 Guide
📄

OCR & Parsing Cost

Document OCR and parsing cost across vendors at your page volume.

Try the calculator →
📖 Guide
🏋️

Training Run Cost

Estimate the GPU cost of a fine-tuning or training run.

Try the calculator →
📖 Guide
⚖️

Jurisdiction Risk Scorer

Score data-sensitivity x jurisdiction x compliance for open-weight / Chinese models.

Try the calculator →
📖 Guide
🏠

Local vs Cloud Inference

Serverless host vs dedicated GPU rental vs self-host — the 3-way serving cost for open-weight models.

Try the calculator →
📖 Guide
🏭

Open-Weight Hosting Compare

Serverless host comparison plus self-host GPU crossover for open-weight models.

Try the calculator →
📖 Guide
⚖️

Buy vs Build

Outcome-priced vendor or build it yourself? Engineering cost included.

Try the calculator →
📖 Guide
💹

Margin Calculator

Is your AI feature profitable? Gross margin + grade.

Try the calculator →
📖 Guide
💼

Budget Planner

Plan AI spend across every use case. Where the budget goes.

Try the calculator →
📖 Guide
📅

Annual Cost Forecaster

Project monthly AI spend for 12 months with growth assumptions.

Try the calculator →
📖 Guide
📆

Quarterly Spend Forecaster

The agentic shift. Your CFO is about to be surprised.

Try the calculator →
📖 Guide
Popular
🧾

Plan Overage & Tier-Break Calculator

Pay overage or upgrade? Break-even across GitHub, Claude, Cursor, Copilot Studio.

Try the calculator →
📖 Guide
Popular
🪙

Credit Decoder

What is an AI credit worth? Decode credits → dollars → tokens & messages.

Try the calculator →
📖 Guide
⚔️

Tier Showdown

The $100 & $200 power-user tiers, head to head: ChatGPT, Claude, Google, Cursor.

Try the calculator →
📖 Guide
🏢

Enterprise Unbundling

Claude Team vs seat-based Enterprise ($20/seat + metered): the crossover.

Try the calculator →
📖 Guide
📊

Scale Projection

Show CFOs what scaling 10× actually costs. Tier-transition warnings.

Try the calculator →
📖 Guide
⚠️

Vendor Concentration Risk

What if your #1 AI vendor goes down? HHI score + diversification plan.

Try the calculator →
📖 Guide
🌍

Region Cost Map

US vs EU vs APAC pricing. Data-residency premiums up to 25%.

Try the calculator →
📖 Guide
🏢

Business AI Bill Diagnose

Break a multi-vendor team AI bill into seats, API, and agentic spend — surface leakage, concentration risk, and the savings plays.

Try the calculator →
📖 Guide
Popular
🧠

Reasoning Token Cost

The hidden thinking-token surcharge — what reasoning tokens add to your bill.

Try the calculator →
📖 Guide
Popular
🔁

Retry Cost

What failed-and-retried API calls quietly cost you each month.

Try the calculator →
📖 Guide
🖼️

Image Generation Cost

$/image and monthly spend across the major image models.

Try the calculator →
📖 Guide
🎬

Video Generation Cost

$/second and monthly spend for AI video generation.

Try the calculator →
📖 Guide
🖥️

Inference Serving Cost

Self-hosted serving economics — the true $/1M tokens on your own GPUs.

Try the calculator →
📖 Guide
👥

Human-in-the-Loop Review Cost

Labor cost of human review on agent actions, and whether catching incidents pays for it.

Try the calculator →
📖 Guide
🔬

Eval / Anti-Hallucination Cost

Continuous eval cost: a golden set scored every release by an LLM-as-judge to catch silent regressions.

Try the calculator →
📖 Guide
📋

AI Compliance / Governance Cost

Annual compliance overhead by cost center, driven by EU AI Act risk tier.

Try the calculator →
📖 Guide
🔌

AI Agent Integration Cost

One-time build + maintenance to connect an agent to CRM, ERP, ticketing, and identity systems.

Try the calculator →
📖 Guide
🔄

Data Freshness / Re-Embedding Cost

The recurring data treadmill: re-embed, reindex, pipeline, and drift monitoring. The data-drift bill.

Try the calculator →
📖 Guide
🔧

Model Maintenance Cost

The model-drift fix bill: forced migrations, prompt re-baselining, re-tuning, and prompt-ops.

Try the calculator →
📖 Guide
🔭

LLM Observability / Tracing Cost

The tracing bill by span: compare per-span, per-trace, and self-host for agentic workloads.

Try the calculator →
📖 Guide
🗼

Agentic TCO + ROI Control Tower

The orchestrator: all twelve cost layers, risk-weighted ROI, and a board-ready GO / NO-GO verdict.

Try the calculator →
📖 Guide
🧮

Agentic TCO + ROI Builder

Raw drivers per layer run live engines behind the scenes, then roll up to a risk-weighted GO / NO-GO verdict.

Try the calculator →
📖 Guide
⚖️

API vs Pro+SDK Breakeven

Direct API or a Claude Pro/Max/Team plan? Models the June-15 Agent SDK carve-out.

Try the calculator →
📖 Guide
💺

AI Seat ROI

Is your coding-assistant seat spend worth it? ROI multiple + break-even hours, live seat prices.

Try the calculator →
📖 Guide
🎧

Support Deflection ROI

AI cost per resolved conversation vs the human cost it avoids — with the human-touch share priced in.

Try the calculator →
📖 Guide
🛟

Agentic Variance Reserve

Re-forecast the plan year with the agentic share compounding: breach month + the 20-40% reserve.

Try the calculator →
📖 Guide
🏗️

Subscription Picker for Builders

Best AI subscription for your team by persona — founder, CTO, FP&A, consultant.

Try the calculator →
📖 Guide

🤖Cost an agent / RAG / voice stack (32)

Multi-step agents, retrieval pipelines, voice bots — full integrated stack costs.

🔎

AI Cost Diligence for VC/PE

Stress-test a startup's claimed AI unit economics against vendor-exact pricing. Verdict + deal-memo.

Try the calculator →
📖 Guide
🔥

Burn Multiple Stress-Test (VC/PE)

Reported vs true burn multiple. Reclassifies subsidised inference back into COGS to expose real gross margin.

Try the calculator →
📖 Guide
🖥

GPU Cost-per-Inference (VC/PE)

API vs self-hosted cost per inference, break-even volume, and whether the claimed GPU fleet can serve the traffic.

Try the calculator →
📖 Guide
📈

Portfolio EBITDA-Lift Estimator (VC/PE)

Turns achievable AI and cloud savings into bps of EBITDA expansion, EV lift at your exit multiple, and payback.

Try the calculator →
📖 Guide
🧠

RAG vs Fine-Tuning

Head-to-head cost across both approaches at YOUR query volume.

Try the calculator →
📖 Guide
🧩

Chunking Optimizer

Chunk size vs retrieval quality vs token cost. Find the sweet spot.

Try the calculator →
📖 Guide
🔎

Hybrid Search Cost

Vector + BM25. Twice the cost, often 30%+ better recall.

Try the calculator →
📖 Guide
🧬

Embedding Cost

RAG indexing + re-indexing + query cost across 9 embedding models.

Try the calculator →
📖 Guide
🗄️

Vector DB Cost

Pinecone vs Qdrant vs Weaviate vs pgvector. 8 options compared.

Try the calculator →
📖 Guide
📚

RAG Pipeline Cost

Full RAG stack on one screen — ingest, storage, query, generation.

Try the calculator →
📖 Guide
📰

Multimodal RAG Stack

Doc + vision + audio retrieval. Composed cost across all stages.

Try the calculator →
📖 Guide
🔁

Agent Loop Cost

Multi-turn agent cost with context growth + runaway warnings.

Try the calculator →
📖 Guide
Popular
🤖

Agentic AI Stack

Planner + executor loop + verifier. Full agent cost on one screen.

Try the calculator →
📖 Guide
Popular
⚙️

Agentic Workflow Cost

Claude Code · Cursor · Copilot. Monthly burn across 4 vendors.

Try the calculator →
📖 Guide
🎙️

Voice Agent Stack

STT + LLM + TTS with interrupt-waste modeling.

Try the calculator →
📖 Guide
🎧

Audio Cost

Speech-to-text + TTS pricing. Voice agent per-call breakdown.

Try the calculator →
📖 Guide
🖼️

Vision Cost

Image analysis pricing. Per-image, per-tile, with detail levels.

Try the calculator →
📖 Guide
🎓

Fine-Tuning Cost

Training + inference + break-even vs base. 4 providers.

Try the calculator →
📖 Guide
🖥️

Self-Host Breakeven

Llama / Qwen on rented GPUs vs API. Real ops costs included.

Try the calculator →
📖 Guide
👥

Human-in-the-Loop Review Cost

Labor cost of human review on agent actions, and whether catching incidents pays for it.

Try the calculator →
📖 Guide
🔒

Guardrail / Security Cost

Runtime safety, prompt-injection and PII screening cost, and the percent it adds over your token bill.

Try the calculator →
📖 Guide
🔬

Eval / Anti-Hallucination Cost

Continuous eval cost: a golden set scored every release by an LLM-as-judge to catch silent regressions.

Try the calculator →
📖 Guide
📋

AI Compliance / Governance Cost

Annual compliance overhead by cost center, driven by EU AI Act risk tier.

Try the calculator →
📖 Guide
🔌

AI Agent Integration Cost

One-time build + maintenance to connect an agent to CRM, ERP, ticketing, and identity systems.

Try the calculator →
📖 Guide
🔄

Data Freshness / Re-Embedding Cost

The recurring data treadmill: re-embed, reindex, pipeline, and drift monitoring. The data-drift bill.

Try the calculator →
📖 Guide
🔧

Model Maintenance Cost

The model-drift fix bill: forced migrations, prompt re-baselining, re-tuning, and prompt-ops.

Try the calculator →
📖 Guide
🔭

LLM Observability / Tracing Cost

The tracing bill by span: compare per-span, per-trace, and self-host for agentic workloads.

Try the calculator →
📖 Guide
🗼

Agentic TCO + ROI Control Tower

The orchestrator: all twelve cost layers, risk-weighted ROI, and a board-ready GO / NO-GO verdict.

Try the calculator →
📖 Guide
🧮

Agentic TCO + ROI Builder

Raw drivers per layer run live engines behind the scenes, then roll up to a risk-weighted GO / NO-GO verdict.

Try the calculator →
📖 Guide
🛟

Agentic Variance Reserve

Re-forecast the plan year with the agentic share compounding: breach month + the 20-40% reserve.

Try the calculator →
📖 Guide
Playbook
📖

Agentic AI Playbook

Step-by-step guided journey. Learn the cost levers in ~10 min.

Try the calculator →
📖 Guide
Playbook
📖

Multimodal RAG Playbook

Build a RAG system. Understand cost levers step by step.

Try the calculator →
📖 Guide

🤖Physical AI & robotics (4)

Robot fleets, humanoids, teleoperation data, buy-vs-RaaS — the AI-layer economics of embodied AI.

🤖

Humanoid Labor ROI

Honest humanoid payback: realized FTE, effective $/hr vs labor, and the AI-ops layer.

Try the calculator →
📖 Guide
🤖

Robot Training-Data / Teleoperation Cost

Price the demonstration corpus to train a robot policy: total program cost + cost per usable demo.

Try the calculator →
📖 Guide
🤖

Robot Buy vs RaaS Breakeven

CAPEX vs Robotics-as-a-Service: the break-even month, hidden fees, and the AI layer both paths carry.

Try the calculator →
📖 Guide
🤖

AI Robot Fleet Cost

Price the AI layer of an AI-driven robot fleet: cost per successful task.

Try the calculator →
📖 Guide

📊Track price trends (6)

2.5 years of vendor pricing history. Free-tier checker. Currency conversion.

Popular
🔍

AI Model Finder

Compare every model side-by-side. Filter by context, price, modality.

Try the calculator →
📖 Guide
Popular
🪙

Credit Decoder

What is an AI credit worth? Decode credits → dollars → tokens & messages.

Try the calculator →
📖 Guide
⚔️

Tier Showdown

The $100 & $200 power-user tiers, head to head: ChatGPT, Claude, Google, Cursor.

Try the calculator →
📖 Guide
💱

Currency Converter

AI spend in 12 currencies — EUR, GBP, INR, JPY, more.

Try the calculator →
📖 Guide
Popular
📈

Pricing History Explorer

2.5+ years of vendor pricing. 62K+ historical price points.

Try the calculator →
📖 Guide
🆓

Free Tier Checker

Do you actually need to pay? Free-tier limits across major vendors.

Try the calculator →
📖 Guide

👤Personal / consumer AI (8)

Subscription decisions for individuals, families, creators, and developers.

🆓

Free Tier Checker

Do you actually need to pay? Free-tier limits across major vendors.

Try the calculator →
📖 Guide
Popular

Subscription Picker

Which AI subscription should I pay for? 6 quick questions.

Try the calculator →
📖 Guide
🖼️

Image Generation Cost

$/image and monthly spend across the major image models.

Try the calculator →
📖 Guide
🎬

Video Generation Cost

$/second and monthly spend for AI video generation.

Try the calculator →
📖 Guide
🗓️

Annual vs Monthly

Lock-in cost vs commit discount. Quantified.

Try the calculator →
📖 Guide
🎬

Creator Bundle

YouTuber, podcaster, blogger? Optimal AI stack + redundancy report.

Try the calculator →
📖 Guide
👨‍👩‍👧

Family Plan Comparator

Cheapest AI plan for a household. Per-person cost ranked.

Try the calculator →
📖 Guide
🛠️

Dev Stack Calculator

Find overlap in your AI dev tools. Stop paying for redundancy.

Try the calculator →
📖 Guide

🎯Calculators for your selected issues

Showing every calculator that addresses the issues you selected. Pick more pills above to expand the list, or clear all to browse everything.

I’m a:
📚 All 96 calculators · alphabetical For power users & search engines
🔁Agent Loop Cost 🤖Agentic AI Stack 🧮Agentic TCO + ROI Builder 🗼Agentic TCO + ROI Control Tower 🛟Agentic Variance Reserve ⚙️Agentic Workflow Cost 🔌AI Agent Integration Cost 📋AI Compliance / Governance Cost 🧮AI Cost Calculator 🔎AI Cost Diligence for VC/PE 🔍AI Model Finder 🤖AI Robot Fleet Cost 💺AI Seat ROI 📅Annual Cost Forecaster 🗓️Annual vs Monthly ⚖️API vs Pro+SDK Breakeven 🎧Audio Cost Batch vs Realtime 💼Budget Planner 🔥Burn Multiple Stress-Test (VC/PE) 🏢Business AI Bill Diagnose ⚖️Buy vs Build 💰Cheapest Model Finder 🌏Chinese vs Western AI Cost 🧩Chunking Optimizer 🧾Consumer AI Bill Diagnose 🗜️Context Compaction 📏Context Window Cost 📚Corpus Onboarding Cost 🎬Creator Bundle 🪙Credit Decoder 💱Currency Converter 🔄Data Freshness / Re-Embedding Cost 🛠️Dev Stack Calculator 🧬Embedding Cost 🏢Enterprise Unbundling 🔬Eval / Anti-Hallucination Cost 🧪Eval & Benchmark Cost 👨‍👩‍👧Family Plan Comparator 🎓Fine-Tuning Cost 🆓Free Tier Checker 🚦Gateway Stage Advisor 🖥GPU Cost-per-Inference (VC/PE) 💸GPU Spot Savings 🔒Guardrail / Security Cost 👥Human-in-the-Loop Review Cost 🤖Humanoid Labor ROI 🔎Hybrid Search Cost 🖼️Image Generation Cost 🖥️Inference Serving Cost ⚖️Jurisdiction Risk Scorer 🔭LLM Observability / Tracing Cost 🏠Local vs Cloud Inference 💹Margin Calculator ⚖️Model Family Comparison 🔧Model Maintenance Cost 🔦MSP Client AI Exposure Audit 🔀Multi-Model Router 📰Multimodal RAG Stack 📄OCR & Parsing Cost 🧭Open-Source Model Finder 🏭Open-Weight Hosting Compare 🔓Open-Weight Model Suitability ⚠️Overage Forecaster 🧾Plan Overage & Tier-Break Calculator 📈Portfolio EBITDA-Lift Estimator (VC/PE) 💾Prompt Cache ROI 📆Quarterly Spend Forecaster 📚RAG Pipeline Cost 🧠RAG vs Fine-Tuning 🧠Reasoning Token Cost 🌍Region Cost Map 🔁Retry Cost 🤖Robot Buy vs RaaS Breakeven 🤖Robot Training-Data / Teleoperation Cost 🎯ROI Quick Check 🛣️Router and Gateway Comparison 📊Scale Projection 🖥️Self-Host Breakeven 🔗Stacked Savings Subscription Picker 🏗️Subscription Picker for Builders 🎧Support Deflection ROI 🏛️TCO Complete 🏦TCO Quick Estimate ⚔️Tier Showdown 🔤Token Estimator ✂️Token Reduction Stack 🏋️Training Run Cost 🗄️Vector DB Cost ⚠️Vendor Concentration Risk 🎬Video Generation Cost 🖼️Vision Cost 🎙️Voice Agent Stack
🚀 New · Public Beta For AI agents · MCP server

Ask AI cost questions inside Claude, ChatGPT, Cursor & Perplexity — natively.

aicost shipped a Model Context Protocol (MCP) server. Plug it into your favorite AI assistant and it can call our 48+ calculators mid-conversation. No more switching tabs to look up pricing — your AI just answers with verified numbers and cites the source.

✓ Working today
  • · Claude.ai Pro/Max/Enterprise
  • · ChatGPT Plus/Pro/Team (OAuth)
  • · Perplexity Pro/Max (OAuth)
  • · Cursor · Continue · Zed · Cody · Goose
🔮 Coming next
  • · Hybrid pricing (subscription vs API)
  • · TCO + ROI playbooks for enterprise
  • · Domain calcs (healthcare, finance, dev)
  • · Self-serve API keys at aicost.ai/account
📨
Want a beta invite?
Email [email protected] with the AI client you'd use (Claude / ChatGPT / Cursor / Perplexity / etc.) and we'll send setup instructions. Free during beta.

Open standard · Model Context Protocol · live at https://mcp.aicost.ai

Go deeper

Our playbooks on cutting this number.

📚
All Cost Topics
12 playbooks for cost problems
📋
Vendor Pricing Guides
15 vendors · weekly refresh
🎯
Book a Quickscan
$1.5K · 5-day cost review

Need help using this calculator for your workloads?

AICost.ai has 50+ calculators and playbooks. Schedule an AvatarVA meeting and we'll work through your real cost scenarios across AI & Cloud: visibility, cost reduction, optimization, forecasting and capacity planning, without sacrificing accuracy or performance.

📅 Schedule an AvatarVA meeting →
📖 Data sources & methodology 163 text models · 9 embeddings · 37 vision · 55 audio · 8 vector DBs across 10 vendor pages · last verified 2026-07-28

Methodology

  • All prices are USD per 1 million tokens, current as of 2026-07-28.
  • Vendor-published values have no mark. Inferred/extrapolated values are marked with * and listed below.
  • Batch API discounts are 50% off standard rates across providers that offer Batch mode.
  • Prompt caching discounts vary by provider (typically 80-90% off cached input tokens).
  • Regional data-residency surcharges (Anthropic 1.1x, OpenAI 1.1x, Google regional tiers) are NOT included in base rates.
  • Long-context pricing tiers apply when input exceeds model threshold.
  • Embedding prices are input-only (no output tokens generated).

Primary sources

Last-verified date is the most recent successful daily snapshot (aicost_pricing_snapshots) or, when no snapshot exists yet, the latest successful crawler run (aicost_crawler_runs). 10 of 10 vendors are currently verified. Aggregator services (TokenCost, AI Pricing Guru, etc.) are not listed.

Anthropic
2026-07-28
https://www.anthropic.com/pricing
Daily snapshot since Sep 2023 · 631 days captured
Anthropic Docs
2026-07-28
https://platform.claude.com/docs/en/about-claude/pricing
Daily snapshot since Sep 2023 · 631 days captured
OpenAI
2026-07-28
https://openai.com/api/pricing/
Daily snapshot since Sep 2023 · 632 days captured
Google AI
2026-07-28
https://ai.google.dev/gemini-api/docs/pricing
Daily snapshot since Dec 2023 · 607 days captured
Google Vertex
2026-07-28
https://cloud.google.com/vertex-ai/generative-ai/pricing
Daily snapshot since Dec 2023 · 607 days captured
DeepSeek
2026-07-28
https://api-docs.deepseek.com/quick_start/pricing
Daily snapshot since May 2024 · 546 days captured
xAI
2026-07-28
https://x.ai/api
Daily snapshot since Nov 2024 · 464 days captured
Mistral
2026-07-28
https://mistral.ai/pricing
Daily snapshot since Dec 2023 · 605 days captured
Cohere
2026-07-28
https://cohere.com/pricing
Daily snapshot since Sep 2023 · 631 days captured

Inferred values (marked with * in calculator tables)

Derived from industry conventions, not directly published by the vendor. Typical conventions: cached input = 10% of base (90% off), Batch API = 50% of base (50% off).

Vendor / Model Field Why it’s inferred
Anthropic — Claude Sonnet 4.6 cachedInput Derived at 10% of input rate — Anthropic publishes 90% cache-hit discount on this tier.
Anthropic — Claude Sonnet 4.5 cachedInput Derived at 10% of input rate; same 90% cache-hit convention as Sonnet 4.6.
Anthropic — Claude Sonnet 4.5 batchInput Derived at 50% of standard input — Anthropic documents uniform 50% Batch discount.
Anthropic — Claude Sonnet 4.5 batchOutput Derived at 50% of standard output — Anthropic documents uniform 50% Batch discount.
Anthropic — Claude Haiku 4.5 cachedInput Derived at 10% of input rate — Anthropic 90% cache-hit discount convention.
OpenAI — GPT-5.4 Mini cachedInput Derived at 10% of input — OpenAI documents automatic 90% discount on cache hits across GPT-5.x tier.
OpenAI — GPT-5.4 Nano cachedInput Derived at 10% of input — OpenAI 90% cache-hit convention.
OpenAI — GPT-5.4 Nano batchInput Derived at 50% of input — OpenAI Batch API uniform 50% discount.
OpenAI — GPT-5.4 Nano batchOutput Derived at 50% of output — OpenAI Batch API uniform 50% discount.
OpenAI — GPT-5.4 Pro cachedInput Derived at 10% of input — OpenAI 90% cache-hit convention.
OpenAI — GPT-5.4 Pro batchInput Derived at 50% of input — OpenAI Batch API uniform 50% discount.
OpenAI — GPT-5.4 Pro batchOutput Derived at 50% of output — OpenAI Batch API uniform 50% discount.
OpenAI — GPT-5.2 cachedInput Derived at 10% of input; no residency uplift.
OpenAI — GPT-5.2 batchInput Derived at 50% of input.
OpenAI — GPT-5.2 batchOutput Derived at 50% of output.
OpenAI — GPT-5 cachedInput Derived at 10% of input.
OpenAI — GPT-5 batchInput Derived at 50% of input.
OpenAI — GPT-5 batchOutput Derived at 50% of output.
OpenAI — GPT-5.5 Pro cachedInput Derived at 10% of input — OpenAI does not publish a cached rate for *-pro models; using the family convention.
OpenAI — GPT-5.5 Pro batchInput Derived at 50% of input.
OpenAI — GPT-5.5 Pro batchOutput Derived at 50% of output.
OpenAI — GPT-5.2 Pro cachedInput Derived at 10% of input — pro-tier convention.
OpenAI — GPT-5.2 Pro batchInput Derived at 50% of input.
OpenAI — GPT-5.2 Pro batchOutput Derived at 50% of output.
OpenAI — GPT-5.1 batchInput Derived at 50% of input.
OpenAI — GPT-5.1 batchOutput Derived at 50% of output.
OpenAI — GPT-5 Pro batchInput Derived at 50% of input.
OpenAI — GPT-5 Pro batchOutput Derived at 50% of output.
OpenAI — GPT-5 Nano cachedInput Derived at 10% of input.
OpenAI — GPT-5 Nano batchInput Derived at 50% of input.
OpenAI — GPT-5 Nano batchOutput Derived at 50% of output.
Google — Gemini 3 Flash cachedInput Derived at 10% of input — Google caching discount convention ~90%.
Google — Gemini 3.1 Flash-Lite cachedInput Derived at 10% of input — Google caching convention.
Google — Gemini 3.1 Flash-Lite batchInput Derived at 50% of input — Google Batch API uniform 50% discount.
Google — Gemini 3.1 Flash-Lite batchOutput Derived at 50% of output — Google Batch API uniform 50% discount.
Google — Gemini 2.5 Pro cachedInput Derived at 10% of input.
Google — Gemini 2.5 Flash cachedInput Derived at 10% of input.
Google — Gemini 2.5 Flash-Lite cachedInput Derived at 10% of input — Google caching convention.
Google — Gemini 2.5 Flash-Lite batchInput Derived at 50% of input — Google Batch API uniform 50% discount.
Google — Gemini 2.5 Flash-Lite batchOutput Derived at 50% of output — Google Batch API uniform 50% discount.
Google — Gemini 2.0 Flash cachedInput Derived at 25% of input per Google 2.0 family caching rates.
Google — Gemini 2.0 Flash batchInput Derived at 50% of input — Google Batch API uniform 50% discount.
Google — Gemini 2.0 Flash batchOutput Derived at 50% of output — Google Batch API uniform 50% discount.
Google — Gemini 2.0 Flash-Lite cachedInput Derived at 10% of input — Google caching convention.
Google — Gemini 2.0 Flash-Lite batchInput Derived at 50% of input — Google Batch API uniform 50% discount.
Google — Gemini 2.0 Flash-Lite batchOutput Derived at 50% of output — Google Batch API uniform 50% discount.
xAI — Grok 4 (legacy) cachedInput Extrapolated at 25% of base.

Pricing is cross-verified against the LiteLLM community registry when available. Daily snapshots are kept in aicost_pricing_snapshots; every change is logged to aicost_price_changelog with old & new values for full audit trail. Read the full methodology →