Free tool

AI Chatbot Cost Calculator

Estimate what an AI chatbot costs to run each month from your conversation volume, token usage and model prices, and see how caching and smaller models change the bill.

Usage

1000

Chats started on your website, WhatsApp or other channels.

6

Booking and FAQ chats are often 4 to 8 replies.

2000

Instructions + retrieved knowledge + history. Often 1,000 to 4,000.

150

A short reply is roughly 50 to 200 tokens.

Prices (example rates, use your provider’s)

$3

Small models cost far less than large ones.

$15

Output usually costs several times more than input.

0%

Prompt caching can cut the cost of repeated instructions.

$50

Chat widget, hosting, vector database, monitoring.

Estimated monthly cost

$0

$0 per conversation

Model usage

$0

Tokens per month

0

Want a chatbot that pays for itself?

We design prompts and retrieval to keep token costs low without hurting answer quality.

Book a free strategy call

Estimates only. Cost = conversations × replies × (input tokens × input price × (1 − cache share × 0.9) + output tokens × output price) ÷ 1,000,000 + platform costs. Cached input is assumed to cost 10% of the normal rate; check your provider’s caching prices. Default prices are examples, not quotes.

Lowering the cost

Three levers for cheaper chatbots

Most chatbot bills are driven by input tokens, not replies.

Right-size the model

Routine FAQ and booking chats often work well on smaller, cheaper models. Reserve larger models for complex questions.

Trim the context

Retrieve only the most relevant knowledge, summarize long histories and keep the system prompt focused.

Cache what repeats

Instructions and policies that don’t change between messages are ideal for prompt caching where your provider supports it.

FAQ

AI Chatbot Cost Calculator FAQ

Straight answers to what buyers ask us most. Don't see yours? Ask on a free call.

Ask us directly

Language models charge per token, roughly a short word or part of a word. Each message sends input tokens (instructions, knowledge and conversation so far) and receives output tokens (the reply). Monthly cost = conversations × messages × tokens × price per token, plus any platform fees.

Every message usually resends the system prompt, retrieved knowledge and conversation history, so input tokens per message are often ten times the output. Shorter prompts, focused retrieval and prompt caching reduce them.

Use the current per-million-token prices from your model provider’s pricing page. Prices differ widely between small and large models and change over time.

No. The calculator runs in your browser and nothing is sent or saved.

Taking projects for next month

Let's start something new together

Book a free 30-minute strategy call. We'll map the agents, automations and integrations worth building first, and tell you honestly what isn't.

No obligation Fixed-scope proposal You own the code