Skip to content

Extract From PDF

PDF Word Counter

A real word count for a PDF, with the counting rule stated up front.

Processed locally in your browser. Your file is not uploaded to any server. How this works

  1. 01Add your file
  2. 02PDF Word Counter
  3. 03Download

Runs in your browser · nothing uploaded

About PDF word counter

Word counts only mean something if you know what counts as a word. This tool extracts the text layer and then splits on whitespace, treating each resulting token as one word after trimming surrounding punctuation. So "state-of-the-art" is one word, "don't" is one word, an em dash standing alone is not a word, and "10" is. That is the same convention most editors and publishers use, which is what makes the number comparable to a brief that says 2,000 words. The report also fixes the two problems a single number hides: it shows the count per page, so you can see where a chapter runs long, and it separates body text from the short fragments typical of headers and footers, since a running footer repeated across forty pages otherwise inflates a total by hundreds of words nobody wrote.

How to PDF word counter

  1. 01

    Add the PDF

    Drop in the document you need counted. Parsing happens locally.

  2. 02

    Read the total

    The document total appears first, with the per-page table beneath it.

  3. 03

    Scan the per-page column

    Use it to find pages that are unexpectedly dense or nearly empty before you edit.

  4. 04

    Export if needed

    Download the report when you have to show a count to a client or editor.

What this tool does

  • Explicit counting rule: whitespace-delimited tokens, punctuation trimmed, hyphenated compounds counted once
  • Per-page breakdown alongside the document total
  • Repeated header and footer fragments identified so they can be excluded
  • Numerals counted as words, matching standard editorial practice
  • Whole calculation runs locally on a document you never upload

Limitations worth knowing

Every PDF tool has constraints. Stating them plainly is more useful than discovering them halfway through your work.

  • Scanned PDFs have no text layer, so the count is zero until the file is OCR'd.
  • Text inside images, charts and logos is invisible to the count.
  • Word counts from Word or Google Docs may differ slightly because each product draws the token boundary in its own way.

How your file is handled

This tool runs entirely inside this browser tab. When you choose a file, your browser reads it from your own disk and hands the bytes to JavaScript running on this page — no network request carries your document anywhere. You can confirm that yourself: open your browser’s developer tools, switch to the Network panel, and run the tool. You will see no upload.

Nothing is stored after the fact. Closing or reloading this tab discards the file, the result and everything derived from them, because none of it ever left your machine. Read how local processing works.

Questions about PDF word counter

Why does my word processor report a different number?

Because the rules differ. Some products count a hyphenated compound as two words, some count footnote markers, some skip text boxes. Our rule is stated above so you can reconcile the gap.

Are headers, footers and page numbers included?

They are detected as repeated short fragments and reported separately, so you can quote a body-text count rather than one padded by forty repetitions of a running title.

Can I count words on just part of the document?

The per-page table gives you that directly — add the rows for the pages you care about rather than re-running on a split file.

Do you keep a copy of my document to count it?

No. The text layer is read in memory in this tab and discarded when you close it. No bytes are transmitted.

Read more about this