AI strategy8 min read
How to Choose an LLM for Your Business
How to choose an LLM for your business: define the task, build a test set, compare quality, cost, speed and data terms, then decide with a scorecard.
Enter your token volumes and your provider's prices to see what a chatbot, assistant or automation will cost before you build it.
| Scenario name | Small model (example) | Mid-size model (example) | Large model (example) |
|---|---|---|---|
| Per request | $0.00062 | $0.00310 | $0.02 |
| Per 1,000 requests | $0.62 | $3.10 | $15.50 |
| Per day | $0.12 | $0.62 | $3.10 |
| Per month | $3.72 | $18.60 | $93.00 |
| Per year | $44.64 | $223.20 | $1,116.00 |
Paste a typical prompt or answer. The estimate uses common rules of thumb (about 4 characters or 0.75 English words per token); real counts depend on the model's tokenizer and the language.
Runs entirely in your browser. No data is sent to a server.
No. They are round example values to show how the calculation works. Prices change often and differ by provider and model, so always enter the current figures from your provider's pricing page.
Input tokens times the input price plus output tokens times the output price, each divided by one million. Cached input tokens are charged at the cached price instead of the normal input price.
As a rough rule, one token is about four characters or three quarters of an English word. German, French and Spanish usually need more tokens for the same content. The exact number depends on the model's tokenizer.
Generating text is more compute-intensive than reading it, so most providers charge a higher rate for output. Long answers can therefore dominate your bill even when prompts are long.
No. It estimates usage-based API charges only. Add VAT, platform fees, minimum commitments or tool subscriptions separately.
AI strategy8 min read
How to choose an LLM for your business: define the task, build a test set, compare quality, cost, speed and data terms, then decide with a scorecard.
Costs & ROI8 min read
Twelve practical ways to reduce LLM API costs without hurting quality: model routing, prompt caching, shorter context, output limits, batching and monitoring.
Costs & ROI8 min read
How LLM API pricing works: input and output tokens, context, caching, batch discounts and hidden multipliers, with a worked monthly cost example.
Costs & ROI7 min read
What tokens in LLMs are, how text is split into them, why languages and formats differ, and how tokens drive context limits, speed and API cost, with examples.
Calculate the ROI of automating a task with AI: hours saved, net monthly benefit, payback period and first-year return, including review time and running costs.
Build clear, structured prompts for ChatGPT, Claude, Gemini and other AI assistants from role, task, context, format and rules. Free, with ready-made templates.
Answer a few questions to see which EU AI Act risk category your AI system likely falls into and which obligations apply. Free orientation, not legal advice.