LLM Token Counter & Cost Estimator

Estimate tokens and API cost for GPT-4o, Claude, Gemini, and Llama

🔒 100 % im Browser🆓 Für immer kostenlos⚡ Kein Signup
Hinweis: Das Tool-Interface unten ist aktuell nur auf Englisch verfügbar. Wir übersetzen die UIs Stück für Stück. Dieses Tool auf Englisch ansehen →
Characters
0
Words
0
Tokens (est.)
0
Context used
0.00%
Estimated cost per call
Input (0 × $2.5/1M)$0.000000
Output (500 × $10/1M)$0.005000
Total per call$0.005000
× 1,000 calls$5.00
× 1,000,000 calls$5000.00

Token counts are estimates based on average characters-per-token (~4 for GPT-4o). Actual counts from the model's tokenizer may differ by 5–15%. Prices reflect published rates and may change - check the provider's pricing page for authoritative numbers.

So funktioniert's

  1. 1
    Paste your prompt

    Drop the full prompt - system message, user message, or a long RAG context.

  2. 2
    Pick a model

    Choose GPT-4o, Claude, Gemini, or Llama. Each model has its own price and tokenizer ratio.

  3. 3
    Set expected output tokens

    Add the response length you expect. The tool shows per-call cost and 1K / 1M-call totals.

Über LLM Token Counter & Cost Estimator

Free online LLM token counter and cost calculator. Paste any prompt to estimate the token count for GPT-4o, GPT-4o mini, Claude Opus/Sonnet/Haiku, Gemini 2.0, and Llama 3, plus the per-call and per-1M-call API cost. Runs entirely in your browser. LLM Token Counter & Cost Estimator auf 712 Tools läuft komplett in deinem Browser über native JavaScript APIs - kein Server sieht je deine Daten. Das heißt: sofortige Ergebnisse, volle Privatsphäre und keine Upload-Limits.

Ob für eine Debugging-Session, einen schnellen Sanity Check oder einen Produktions-Incident - dieses Tool ist kostenlos so oft nutzbar, wie du willst. Keine Wasserzeichen, kein Signup und keine Ads im Tool selbst.

Häufige Fragen

How accurate are the token counts?

They're estimates based on average characters-per-token (~4 for GPT, ~3.6 for Claude). Actual counts from the model's real tokenizer usually differ by 5–15%.

Where do the prices come from?

Published rates from OpenAI, Anthropic, Google, and hosted-Llama providers. Rates change - always verify with the provider before building financial projections.

Does it upload my prompt?

No. Counting and cost math run entirely in your browser.

Why does context-used % matter?

If a prompt approaches the model's context window, the model may truncate or refuse. The gauge helps you stay well under the limit.

Weiterlesen