Knowledge Base Token Estimator
Drop your knowledge-base documents to estimate total tokens, chunks and embedding cost.
Runs in your browser — nothing is uploadedFiles
| File | Characters | Words | Tokens (GPT) |
|---|
How to use Knowledge Base Token Estimator
- Drop files or choose them (text-based formats, up to 50 MB in total).
- See tokens per file and in total for each model family.
- Check whether they fit a model’s context window and what embedding them costs.
Formula & assumptions
- Counts are estimates: the tokenizer approximation was calibrated against OpenAI’s o200k_base (typically within a few percent for English prose and code; up to about 20% off for some other languages). Nothing you paste leaves your browser.