Knowledge Base Token Estimator

Drop your knowledge-base documents to estimate total tokens, chunks and embedding cost.

Runs in your browser — nothing is uploaded

How to use Knowledge Base Token Estimator

  1. Drop files or choose them (text-based formats, up to 50 MB in total).
  2. See tokens per file and in total for each model family.
  3. Check whether they fit a model’s context window and what embedding them costs.

Formula & assumptions

  • Counts are estimates: the tokenizer approximation was calibrated against OpenAI’s o200k_base (typically within a few percent for English prose and code; up to about 20% off for some other languages). Nothing you paste leaves your browser.