—per conversation
—per year
—vs human agents
Cheapest models for this chatbot
Same traffic, every model ranked by monthly cost.
| Model | Cost / month | Per conversation |
|---|
⚠️ Estimate using reference prices (July 2026) and list rates. Real bills vary with caching, tools, retrieval, batching, region and tiers. Long chats cost more than people expect because history is resent every turn — that's modelled here. ·
Report outdated price →
Why chatbot cost grows faster than message count
The trap with chatbots is resent history. To stay coherent, most bots send the whole conversation so far with every new message. So the 1st reply pays for one message of context, the 10th pays for nine. Total input tokens grow roughly with the square of the conversation length — which is why a few very long chats can dominate your bill. Trimming or summarising history, capping conversation length, and using a smaller model for routine turns are the biggest levers. Try the "Summarised history" option above to see the difference.
Comparing to humans? A short chat on a small model often costs a fraction of a cent — far below a human ticket. But a long chat on a frontier model with full history can cross into dollars and beat a human's cost. The break-even above tells you where you stand. Building something bigger than support? Use the AI app cost estimator or compare models head-to-head on the AI API cost calculator.
Cheapest LLM APITranslation API CostContext Window CalculatorAPI Pricing / Monetization CalculatorLLM Evaluation Cost Calculator
How this calculator works
Chatbot Cost Calculator estimates the monthly bill for running an AI chatbot on a per-token pricing model. You enter your conversations per month, the number of bot replies per conversation, and the average token size of user messages, bot replies, and your system prompt. The calculator multiplies these across your chosen model to project total input and output tokens, then applies that model's rates. The biggest driver is conversation memory: when a bot resends prior history with each new turn, earlier messages get re-billed as input on every reply, so a long chat can cost several times more than the raw message count suggests. An optional human agent cost per conversation lets you compare automation against staffing.
The key trade-off to watch is reply length versus resent context. Because history is re-sent each turn, the token cost of a conversation grows roughly with the square of its length, not linearly — so trimming the number of replies or capping how much memory you carry forward often saves more than shrinking individual messages. Test a realistic replies-per-conversation figure and try lowering the memory setting to see how much a shorter context window reduces the total. A large but cheaper model can also beat a small expensive one at scale, so compare model choice against your actual conversation shape rather than headline per-token rates.
Frequently asked questions
Why do long chatbot conversations cost so much more?
Most chatbots resend the whole conversation history with every new message so the model remembers context. That means the input tokens grow with each turn, and the total cost of a conversation rises roughly with the square of the number of messages. A 20-message chat can cost far more than four 5-message chats, even though it has the same number of messages.
Is an AI chatbot cheaper than a human support agent?
Usually yes per conversation, but it depends on length and model. A short chat on a small model can cost a fraction of a cent, while a long chat on a frontier model with full history can run into dollars. This calculator compares your AI cost per conversation against what you pay a human to handle the same ticket, so you can see the real break-even.