All comparisons · Prices checked 2026-10-01
Gemini 3.8 Flash vs Claude Haiku 4.5: API cost comparison
For a typical chatbot reply, Gemini 3.8 Flash costs $0.0025 and Claude Haiku 4.5 costs $0.0037 per request: Gemini 3.8 Flash is 32% cheaper. Gemini 3.8 Flash is cheaper in every workload below.
API prices
| Gemini 3.8 Flash | Claude Haiku 4.5 | |
|---|---|---|
| Maker | Anthropic | |
| Input / 1M tokens | $0.75 | $1 |
| Output / 1M tokens | $3.75 | $5 |
| Output vs input price | 5.0× | 5.0× |
| Context window | 1.05M tokens | 200K tokens |
| Tokens for the same English text | ~95% of GPT's count (estimate) | ~105% of GPT's count (estimate) |
Cost by workload
| Workload (per 1,000 requests) | Gemini 3.8 Flash | Claude Haiku 4.5 | Cheaper |
|---|---|---|---|
| Chatbot reply short question plus some chat history · 1,500 in / 400 out | $2.49 | $3.67 | Gemini 3.8 Flash 32% cheaper |
| Document Q&A (RAG) retrieved passages plus a question · 8,000 in / 500 out | $7.48 | $11.03 | Gemini 3.8 Flash 32% cheaper |
| Coding agent step large working context, short edit · 25,000 in / 800 out | $20.66 | $30.45 | Gemini 3.8 Flash 32% cheaper |
| Summarize a long report long input, short summary · 20,000 in / 300 out | $15.32 | $22.58 | Gemini 3.8 Flash 32% cheaper |
| Write an article short brief, long draft · 800 in / 2,000 out | $7.69 | $11.34 | Gemini 3.8 Flash 32% cheaper |
Gemini 3.8 Flash is cheaper on both input and output, so it costs less for every kind of request. The gap is 1.5× on input and 1.5× on output once the tokenizer difference is included.
Same text, different token counts. Each company uses its own tokenizer. By our estimate Claude Haiku 4.5 counts about 11% more tokens than Gemini 3.8 Flash for the same English text, and the costs above already include that. Paste your own prompt into the token counter to see both counts.
Monthly cost
| Monthly volume | Gemini 3.8 Flash | Claude Haiku 4.5 |
|---|---|---|
| 10,000 chatbot replies / month | $24.94 | $36.75 |
| 100,000 chatbot replies / month | $249 | $368 |
| 1,000,000 chatbot replies / month | $2,494 | $3,675 |
Calculate your own
Enter your typical request size and volume.
Token counts are for English text measured with GPT's tokenizer; each model's own tokenizer difference is applied automatically.
Questions
Which is cheaper, Gemini 3.8 Flash or Claude Haiku 4.5?
For a typical chatbot reply (1,500 input and 400 output tokens), Gemini 3.8 Flash costs $0.0025 and Claude Haiku 4.5 costs $0.0037 per request, so Gemini 3.8 Flash is 32% cheaper. Gemini 3.8 Flash is cheaper in every workload below.
How much do Gemini 3.8 Flash and Claude Haiku 4.5 cost per million tokens?
Gemini 3.8 Flash costs $0.75 per 1M input tokens and $3.75 per 1M output tokens. Claude Haiku 4.5 costs $1 input and $5 output. Prices checked 2026-10-01.
Do both models count the same text as the same number of tokens?
No. Each model family has its own tokenizer, so the same text can be a different number of tokens. The costs on this page already include that difference (estimated from GPT's tokenizer).
Related comparisons
Paste a real prompt to see its exact tokens and cost on Gemini 3.8 Flash, Claude Haiku 4.5 and 30+ other models.
Open the token counter →Standard API list prices, short-context tier, no caching or batch discounts. Prices change; check each provider's pricing page before large jobs.