Extract PDFs as optimized text. Count tokens in real-time before pasting into Claude, ChatGPT, or any LLM. See compression and reduction. 100% browser-based.
Drop any PDF. Instantly extract text and detect structure. Identify tables, sections, content hierarchy. No data stored, 100% browser-based processing.
Remove boilerplate (copyright, document IDs, report headers), strip decorative chars (©®™→), eliminate chart artifacts (axis numbers), remove headers/footers, collapse spacing, join short lines. Preserve all meaningful content. Savings vary 0-80% depending on PDF structure.
Estimated token counts. See before/after and reduction percentage. Copy to clipboard or download as .txt. Works with Claude, ChatGPT, and any LLM.
PDF Extraction: Mozilla's PDF.js — the trusted tool used by Firefox and major browsers to render PDFs accurately.
Token Counting: gpt-tokenizer — the official JavaScript tokenizer for all OpenAI models (GPT-5, GPT-4o, o1, o3, GPT-4, GPT-3.5). Fastest, smallest footprint tokenizer available. Same results as OpenAI's official tokenizer.
Processing: Everything happens locally in your browser. No servers, no uploads, no internet connection needed (after initial page load). Your files never leave your computer.
Your Privacy: We don't store, track, or see your PDFs. No logins, no accounts, no data collection. 100% private. Files up to 50MB work best.