LLM Token Counter & Cost Estimator

Estimate tokens and API cost for GPT-4o, Claude, Gemini, and Llama

๐Ÿ”’ 100% ๋ธŒ๋ผ์šฐ์ € ๋‚ด๐Ÿ†“ ํ‰์ƒ ๋ฌด๋ฃŒโšก ๊ฐ€์ž… ์—†์Œ
์ฐธ๊ณ : ์•„๋ž˜ ํˆด ์ธํ„ฐํŽ˜์ด์Šค๋Š” ํ˜„์žฌ ์˜์–ด๋กœ๋งŒ ์ œ๊ณต๋ฉ๋‹ˆ๋‹ค. UI๋Š” ํ•˜๋‚˜์”ฉ ๋ฒˆ์—ญ ์ค‘์ž…๋‹ˆ๋‹ค. ์ด ํˆด์„ ์˜์–ด๋กœ ๋ณด๊ธฐ โ†’
Characters
0
Words
0
Tokens (est.)
0
Context used
0.00%
Estimated cost per call
Input (0 ร— $2.5/1M)$0.000000
Output (500 ร— $10/1M)$0.005000
Total per call$0.005000
ร— 1,000 calls$5.00
ร— 1,000,000 calls$5000.00

Token counts are estimates based on average characters-per-token (~4 for GPT-4o). Actual counts from the model's tokenizer may differ by 5โ€“15%. Prices reflect published rates and may change - check the provider's pricing page for authoritative numbers.

์‚ฌ์šฉ๋ฒ•

  1. 1
    Paste your prompt

    Drop the full prompt - system message, user message, or a long RAG context.

  2. 2
    Pick a model

    Choose GPT-4o, Claude, Gemini, or Llama. Each model has its own price and tokenizer ratio.

  3. 3
    Set expected output tokens

    Add the response length you expect. The tool shows per-call cost and 1K / 1M-call totals.

LLM Token Counter & Cost Estimator ์†Œ๊ฐœ

Free online LLM token counter and cost calculator. Paste any prompt to estimate the token count for GPT-4o, GPT-4o mini, Claude Opus/Sonnet/Haiku, Gemini 2.0, and Llama 3, plus the per-call and per-1M-call API cost. Runs entirely in your browser. 712 Tools์˜ LLM Token Counter & Cost Estimator์€ ์ตœ์‹  JavaScript API๋ฅผ ์‚ฌ์šฉํ•ด ์—ฌ๋Ÿฌ๋ถ„์˜ browser์—์„œ ์™„์ „ํžˆ ๋™์ž‘ํ•ฉ๋‹ˆ๋‹ค - server๊ฐ€ ์—ฌ๋Ÿฌ๋ถ„์˜ ๋ฐ์ดํ„ฐ๋ฅผ ๋ณผ ์ผ์ด ์ ˆ๋Œ€ ์—†์Šต๋‹ˆ๋‹ค. ์ฆ‰, ์ฆ‰๊ฐ์ ์ธ ๊ฒฐ๊ณผ, ์™„์ „ํ•œ ํ”„๋ผ์ด๋ฒ„์‹œ, ๊ทธ๋ฆฌ๊ณ  ์—…๋กœ๋“œ ์ œํ•œ์ด ์—†์Šต๋‹ˆ๋‹ค.

debugging ์„ธ์…˜, ๋น ๋ฅธ sanity check, ๋˜๋Š” ํ”„๋กœ๋•์…˜ ์ธ์‹œ๋˜ํŠธ - ์–ด๋–ค ์ƒํ™ฉ์ด๋“  ์ด ํˆด์€ ์›ํ•˜๋Š” ๋งŒํผ ๋ฌด๋ฃŒ๋กœ ์‚ฌ์šฉํ•  ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค. ์›Œํ„ฐ๋งˆํฌ ์—†์Œ, ๊ฐ€์ž… ์—†์Œ, ํˆด ๋‚ด ๊ด‘๊ณ  ์—†์Œ.

์ž์ฃผ ๋ฌป๋Š” ์งˆ๋ฌธ

How accurate are the token counts?

They're estimates based on average characters-per-token (~4 for GPT, ~3.6 for Claude). Actual counts from the model's real tokenizer usually differ by 5โ€“15%.

Where do the prices come from?

Published rates from OpenAI, Anthropic, Google, and hosted-Llama providers. Rates change - always verify with the provider before building financial projections.

Does it upload my prompt?

No. Counting and cost math run entirely in your browser.

Why does context-used % matter?

If a prompt approaches the model's context window, the model may truncate or refuse. The gauge helps you stay well under the limit.

๋” ์ฝ๊ธฐ