AI Model Cost Comparator - Compare GPT, Claude, Gemini and Llama Prices

Paste a prompt and compare per-call, per-day and per-month cost across 12 AI models

๐Ÿ”’ 100% ๋ธŒ๋ผ์šฐ์ € ๋‚ด๐Ÿ†“ ํ‰์ƒ ๋ฌด๋ฃŒโšก ๊ฐ€์ž… ์—†์Œ
์ฐธ๊ณ : ์•„๋ž˜ ํˆด ์ธํ„ฐํŽ˜์ด์Šค๋Š” ํ˜„์žฌ ์˜์–ด๋กœ๋งŒ ์ œ๊ณต๋ฉ๋‹ˆ๋‹ค. UI๋Š” ํ•˜๋‚˜์”ฉ ๋ฒˆ์—ญ ์ค‘์ž…๋‹ˆ๋‹ค. ์ด ํˆด์„ ์˜์–ด๋กœ ๋ณด๊ธฐ โ†’
Estimated ~20 input tokens (rough: 4 chars/token).
Cheapest
GPT-4o mini ยท $0.000303/call
Most expensive
Claude Opus 4 ยท $0.0378/call
Spread
124.8x
Model$/1M in$/1M outPer callPer dayPer monthFits?
GPT-4o mini
OpenAI ยท 128k ctx
$0.15$0.6$0.000303$0.3030$9.09โœ“
Llama 3.3 70B
Meta (via Groq) ยท 128k ctx
$0.59$0.79$0.000407$0.4068$12.204โœ“
DeepSeek V3
DeepSeek ยท 64k ctx
$0.27$1.1$0.000555$0.5554$16.662โœ“
Gemini 2.5 Flash
Google ยท 1000k ctx
$0.3$2.5$0.001256$1.256$37.68โœ“
Claude Haiku 4.5
Anthropic ยท 200k ctx
$1$5$0.002520$2.52$75.60โœ“
Mistral Large 2
Mistral ยท 128k ctx
$2$6$0.003040$3.04$91.20โœ“
Gemini 2.5 Pro
Google ยท 1000k ctx
$1.25$10$0.005025$5.025$150.75โœ“
GPT-4o
OpenAI ยท 128k ctx
$2.5$10$0.005050$5.05$151.50โœ“
o1-mini
OpenAI ยท 128k ctx
$3$12$0.006060$6.06$181.80โœ“
Claude Sonnet 4.5
Anthropic ยท 200k ctx
$3$15$0.007560$7.56$226.80โœ“
o1 (reasoning)
OpenAI ยท 200k ctx
$15$60$0.0303$30.30$909.00โœ“
Claude Opus 4
Anthropic ยท 200k ctx
$15$75$0.0378$37.80$1,134.00โœ“
Prices are public list pricing per 1M tokens and change often. Token counts are estimated (~4 chars/token in English) - use exact input tokens for precise cost. Cached rate applies to Anthropic/OpenAI/DeepSeek prompt caching where supported.
์ถ”์ •์น˜์ผ ๋ฟ์ž…๋‹ˆ๋‹ค. ๊ฒฐ๊ณผ๋Š” ์—ฌ๋Ÿฌ๋ถ„์ด ์ž…๋ ฅํ•œ ๊ฐ’์—์„œ ์ผ๋ฐ˜ ๊ณต์‹์œผ๋กœ ์ƒ์„ฑ๋˜๋ฉฐ, ์ •๋ณด ์ œ๊ณต ๋ชฉ์ ์œผ๋กœ๋งŒ ์ œ๊ณต๋ฉ๋‹ˆ๋‹ค. ๊ธˆ์œต, ์„ธ๋ฌด, ๋ฒ•๋ฅ , ์˜๋ฃŒ, ๋˜๋Š” ์ „๋ฌธ๊ฐ€ ์กฐ์–ธ์ด ์•„๋‹ˆ๋ฉฐ, ์—ฌ๋Ÿฌ๋ถ„์˜ ๊ฐœ์ธ์  ์ƒํ™ฉ, ๊ด€ํ• , ์„ธ๊ธˆ, ์ˆ˜์ˆ˜๋ฃŒ, ๊ฑด๊ฐ• ์ƒํƒœ ๋“ฑ์„ ๊ณ ๋ คํ•˜์ง€ ์•Š์Šต๋‹ˆ๋‹ค. ์ค‘์š”ํ•œ ์ˆ˜์น˜๋Š” ๋ฐ˜๋“œ์‹œ ์ž๊ฒฉ์„ ๊ฐ–์ถ˜ ์ „๋ฌธ๊ฐ€์™€ ํ™•์ธํ•˜์„ธ์š”. ์ด ํˆด์„ ์‚ฌ์šฉํ•จ์œผ๋กœ์จ ๋‹ค์Œ์— ๋™์˜ํ•œ ๊ฒƒ์œผ๋กœ ๊ฐ„์ฃผ๋ฉ๋‹ˆ๋‹ค: ์ด์šฉ์•ฝ๊ด€ & ๋ฉด์ฑ…์กฐํ•ญ.

์‚ฌ์šฉ๋ฒ•

  1. 1
    Paste your prompt

    The tool estimates input tokens (~4 characters per token in English). Enter exact token counts on the right for precise numbers.

  2. 2
    Set expected output tokens and call volume

    Roughly how long the responses will be, and how many calls you make per day.

  3. 3
    Adjust the cache hit rate

    If you use Anthropic, OpenAI or DeepSeek prompt caching, drag the slider up to see the discounted cost.

  4. 4
    Read the sorted table

    Models are sorted cheapest first, with per-call, per-day and per-month totals and whether the context window fits your prompt.

AI Model Cost Comparator - Compare GPT, Claude, Gemini and Llama Prices ์†Œ๊ฐœ

Free AI model cost comparator. Paste your prompt or enter exact input tokens, set expected output length and calls per day, and see per-call and monthly cost across Claude Opus, Sonnet and Haiku, GPT-4o, o1, Gemini 2.5, Llama 3.3, Mistral Large and DeepSeek V3. Factor in prompt caching. Helps you pick the right model for your budget without spinning up 12 accounts. 712 Tools์˜ AI Model Cost Comparator - Compare GPT, Claude, Gemini and Llama Prices์€ ์ตœ์‹  JavaScript API๋ฅผ ์‚ฌ์šฉํ•ด ์—ฌ๋Ÿฌ๋ถ„์˜ browser์—์„œ ์™„์ „ํžˆ ๋™์ž‘ํ•ฉ๋‹ˆ๋‹ค - server๊ฐ€ ์—ฌ๋Ÿฌ๋ถ„์˜ ๋ฐ์ดํ„ฐ๋ฅผ ๋ณผ ์ผ์ด ์ ˆ๋Œ€ ์—†์Šต๋‹ˆ๋‹ค. ์ฆ‰, ์ฆ‰๊ฐ์ ์ธ ๊ฒฐ๊ณผ, ์™„์ „ํ•œ ํ”„๋ผ์ด๋ฒ„์‹œ, ๊ทธ๋ฆฌ๊ณ  ์—…๋กœ๋“œ ์ œํ•œ์ด ์—†์Šต๋‹ˆ๋‹ค.

debugging ์„ธ์…˜, ๋น ๋ฅธ sanity check, ๋˜๋Š” ํ”„๋กœ๋•์…˜ ์ธ์‹œ๋˜ํŠธ - ์–ด๋–ค ์ƒํ™ฉ์ด๋“  ์ด ํˆด์€ ์›ํ•˜๋Š” ๋งŒํผ ๋ฌด๋ฃŒ๋กœ ์‚ฌ์šฉํ•  ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค. ์›Œํ„ฐ๋งˆํฌ ์—†์Œ, ๊ฐ€์ž… ์—†์Œ, ํˆด ๋‚ด ๊ด‘๊ณ  ์—†์Œ.

์ž์ฃผ ๋ฌป๋Š” ์งˆ๋ฌธ

How accurate are the token estimates?

The character-based estimate is within 10 to 15% for English text. Different models tokenize slightly differently, so for a precise budget, use the exact input token field or run a small sample through the real API.

Are these prices current?

Prices are public list prices at the time of the last update. Providers change pricing periodically; check the provider's pricing page before committing large volumes.

What is prompt caching?

Anthropic, OpenAI and DeepSeek offer a discount (roughly 90%) on input tokens that are repeated across calls, such as a long system prompt. If most of your prompt is a stable prefix, the cache hit rate is high and cost drops sharply.

Is my prompt uploaded anywhere?

No. All calculation runs in your browser. Your prompt text is not sent to any server.

๋” ์ฝ๊ธฐ