All PDF Tools
Navigation
Home
PDF Tools
🔒 Files never leave your device

PDF Word Counter

Count words, characters, sentences, and paragraphs in any PDF — plus reading time, readability score, keyword frequency, and page-by-page analytics. No signup, no upload.

Drop your PDF to count words

Click to browse or drag & drop

No signup 100% private SSL encrypted
Pair This With These Free Tools

TL;DR — Count words in a PDF in seconds

  • What it does: extracts text from a PDF and computes word/character/sentence/paragraph counts, reading time, a readability score, keyword frequency, and per-page statistics.
  • How to use it: upload your PDF and the full dashboard appears automatically — search keywords, sort pages by word count, or export the results as CSV, JSON, or TXT.
  • Privacy: runs 100% in your browser using PDF.js — your file and its text are never uploaded to a server.
  • Cost & limits: completely free, no signup, no daily limit.
  • Limitation: only works on PDFs with a real text layer — scanned PDFs need OCR PDF first, since there's no text to count otherwise.
Step-by-Step Guide

How to Count Words in a PDF Free Online

Four steps. Under a minute. Nothing ever leaves your browser.

Upload Your PDF

Click Browse or drag and drop your PDF. Text extraction with PDF.js begins immediately, entirely inside your browser.

Drag & Drop 100% Private

Review the Automatic Analysis

A full dashboard appears the instant extraction finishes — words, characters, sentences, paragraphs, reading time, and a readability score, with no clicks required.

11 Stats Instant

Explore Pages & Search Text

Sort the page-by-page table by word count, search the extracted text with live highlighting, or look up a keyword to see occurrences and context snippets.

Sortable Table Keyword Search

Export the Statistics

Download the full analysis as CSV, JSON, or TXT — ready for a report, an audit, or another tool.

3 Formats No Signup
💡 Pro Tip: If the dashboard shows 0 words on a PDF you know has text, the file is almost certainly a scanned image with no text layer — run OCR PDF first, then come back and count words on the result.

Why PDFcrest

More Than a Word Counter — a PDF Text Analytics Platform

Eleven primary statistics, keyword search with context snippets, and per-page breakdowns — most free tools stop at a single total.

11 Primary Statistics, Not Just One

Words, characters (with and without spaces), sentences, paragraphs, pages, three averages, plus reading and speaking time — all computed the moment extraction finishes.

Real Readability Scoring

A genuine Flesch Reading Ease score and Flesch-Kincaid grade level — the same standard formulas used in publishing and government plain-language guidelines, not a made-up "easy/hard" label.

Full-Text Keyword Search

Search any word or phrase and get occurrence counts, a list of matching pages, and context snippets — not just a yes/no "found it."

Sortable Page-by-Page Table

See exactly which pages are longest or shortest, sorted in one click — useful for spotting where a document's content is concentrated.

Export CSV, JSON, or TXT

Take the full statistics — including the top keywords and every page's numbers — into a spreadsheet, a script, or a report.

Private by Design — No Upload

Your PDF is read directly in your browser's memory with PDF.js. It's never transmitted anywhere — after the page itself loads, the tool makes zero network requests.

Is This Tool Right For You?

Who Should Use PDF Word Counter

A quick decision guide — this tool reads and measures text, it doesn't edit or restructure it.

Who It's For

  • Writers, students, and researchers checking length against a word-count limit
  • Editors and publishers reviewing readability before print
  • Translators and localization teams quoting jobs by word count
  • Legal, business, and SEO professionals auditing document content

When To Use It

  • A submission, contract, or assignment has a strict word or page limit
  • You need a defensible reading-level score before publishing
  • You're estimating translation cost or narration/presentation time
  • You want to find every occurrence of a specific term across a long document

When NOT To Use It

  • You need to edit the text itself — convert to PDF to Word first
  • Your PDF is a scan with no text layer — run OCR PDF first
  • You need document properties, not content stats — try Edit Metadata
  • Your PDF requires a password to open — unlock it first with Unlock PDF

Comparison

PDFcrest vs Other Ways to Count Words in a PDF

Adobe Acrobat and most free online counters show a single word total and nothing else.

FeaturePDFcrestAdobe AcrobatTypical Free Online Tools
PDF Word Count — Free✓ Always free✗ Full Acrobat subscription~ Often free, some cap daily use
No Signup Required✓ Zero signup✗ Adobe account required~ Varies by tool
File Ever Uploaded to a Server✓ Never — runs in your browser✗ Uploaded to Adobe's servers✗ Usually uploaded to a server
Character, Sentence & Paragraph Counts✓ All included✗ Not shown~ Rarely more than word count
Reading & Speaking Time Estimates✓ 3 reading speeds + speaking✗ Not offered✗ Rarely offered
Readability Score (Flesch)✓ Built in✗ Not offered✗ Rarely offered
Keyword Frequency Table✓ Built in, stopword filter✗ Not offered✗ Not offered
Page-by-Page Word Count Table✓ Sortable✗ Not offered✗ Not offered
Export Statistics (CSV/JSON/TXT)✓ All three formats✗ Not offered~ Some offer basic export
Mobile Support✓ Any modern mobile browser~ App or desktop only for full features~ Varies

Comparison reflects each tool's general publicly documented free-tier behavior at time of writing and may change — always check the provider's current terms.

Use Cases

Who Counts Words in PDFs — and Why

Anyone who needs to measure, verify, or search the text inside a PDF relies on a PDF word counter.

✍️

Writers & Authors

Check a manuscript or article against a publisher's word count requirement before submission, without re-opening the original document.

🎓

Students & Researchers

Verify an essay, thesis chapter, or paper meets a professor's or journal's length requirement, and search for how often a key term appears.

📖

Editors & Publishers

Check the Flesch reading level before print, compare chapter lengths in the page table, and confirm estimated reading time for back-cover copy.

⚖️

Legal Professionals

Search a contract or filing for every occurrence of a defined term or clause reference, with the exact page and surrounding context for each match.

🌐

Translators & Localization Teams

Get an accurate source word count to quote a translation job, without manually copying text out of the PDF into a word processor first.

📈

SEO Professionals & Content Marketers

Audit a PDF resource or whitepaper's content depth and top keywords before repurposing it into a web page or blog post.

🏢

Businesses & Government Offices

Confirm a report or public notice meets a required length or plain-language readability target before it's published or filed.

Workflow Recipes

Common PDF Word Counter Workflows

Counting words is rarely the only step. Here are the tool combinations PDFcrest users chain together most often.

Scanned Document → OCR → Count Words

A scanned PDF has no text layer to count. Run OCR first to make it searchable, then come back here for accurate statistics.

OCR PDFWord Counter (this tool)

Count Words → Edit to Fit a Limit

Found out you're 500 words over the limit? Convert to an editable Word file and trim the content down.

Word Counter (this tool)PDF to Word

Split by Chapter → Count Each Section

Extract each chapter or section into its own file first, then run word counts on each part separately for a per-section breakdown.

Split PDFWord Counter (this tool)

Count Words → Align Metadata Keywords

Use the keyword frequency table to see what a document is actually about, then make sure the PDF's title and keywords metadata match.

Word Counter (this tool)Edit Metadata

Full Document Audit: Size + Content

Analyze what's inflating a PDF's file size, then analyze its actual text content — a complete before-you-publish check.

PDF Size AnalyzerWord Counter (this tool)

Need the text itself, not just statistics about it? Use PDF to Word to get fully editable content.

People Also Search For

pdf word countercount words in pdfcount words in pdf documentcount words on pdfcount words from pdfcount words in pdf onlineonline pdf word countpdf word count toolpdf word count checkerpdf character count onlinecount characters in pdfpdf page counterpdf text statisticsfree pdf word countpdf to word counter

Related Tools

ocr pdfpdf to wordedit pdf metadatapdf size analyzersplit pdf

Technical Detail

How PDF Word Counting Actually Works

How Text Extraction Works

A PDF page doesn't store "paragraphs" or "lines" — it stores a flat list of positioned text runs, each with its own x/y coordinates. This tool uses PDF.js, the same open-source rendering engine behind Firefox's built-in PDF viewer, to read that list via its getTextContent() API. Runs that share roughly the same y-coordinate are joined into one line; a vertical gap larger than about 1.7x the page's typical line spacing is treated as a paragraph break rather than a simple line wrap. That reconstruction is what turns a flat list of text fragments back into readable lines and paragraphs.

How Words Are Actually Counted

A "word" here is any run of Unicode letters or digits that may contain an internal apostrophe or hyphen — this correctly counts don't and well-known as one word each, and works on non-English scripts, not just ASCII. It does not treat punctuation as part of a word, so an email address or URL splits into its letter/number segments the same way most word processors count them, and a value like $3.50 counts as two number tokens rather than one currency value. These are deliberate, documented tradeoffs of a simple, fast tokenizer — not silent inaccuracy.

How Sentences and Paragraphs Are Detected

Sentence counting is rule-based: it protects common abbreviations (Mr., Dr., e.g., U.S., and about two dozen others), decimal numbers, and ellipses from being misread as sentence endings, then splits on a period, question mark, or exclamation point followed by whitespace and a capital letter, digit, or quotation mark. Paragraph counts come directly from the line-reconstruction step above — each block separated by a detected paragraph break is one paragraph.

How the Readability Score Is Calculated

The reading-level score is the standard Flesch Reading Ease formula, alongside a Flesch-Kincaid Grade Level — the same formulas referenced in U.S. federal plain-language guidance for public-facing writing. Both need an estimated syllable count per word, computed with a vowel-group heuristic (the same general approach standard readability libraries use). This heuristic is tuned for English; documents in non-Latin scripts fall back to one syllable per word, so the score is presented as an English-text estimate, not a universal metric.

Why Your Word Count Might Differ From Microsoft Word's

Small differences between word counters are normal and expected — Microsoft Word, Google Docs, and PDFcrest all use slightly different rules for what counts as a "word" (hyphenated compounds, numbers with decimals, and stray formatting characters are the most common sources of disagreement). A difference of a few words per thousand is a tokenization-rule difference, not an error in either tool.

Best Practices

  • If the count reads zero, check for a text layer first. A 0-word result on a PDF you know has content almost always means it's a scanned image — run OCR PDF, then recount.
  • Use the page table to spot outliers — a page with an unusually low word count relative to its neighbors is worth a manual look; it may be an image-heavy page or a formatting artifact.
  • Toggle stop words on when you want raw frequency (e.g. for a linguistic count), and leave them off for a more meaningful "what is this document about" keyword list.
  • Export JSON if you're piping results into another tool or script — it includes the full page-level breakdown and top 100 keywords, not just the headline numbers.

Common Mistakes

  • Assuming a scanned PDF will "just work." Without OCR, there's no text to count — the tool will say so rather than silently returning zero.
  • Comparing counts across tools without checking their rules. A different tokenizer, not a bug, is the usual explanation for small discrepancies.
  • Reading the Flesch score as a universal judgment of quality. It measures sentence and word length, not accuracy, tone, or argument quality — and it's calibrated for English.
  • Forgetting the reading-time estimate is a speed, not a guarantee. Change the dropdown to match your actual reading style for a more useful number.

Glossary

Key Terms, Explained Simply

PDF (Portable Document Format): a file format that preserves a document's exact layout, fonts, and images regardless of the device or software used to open it.

Word Count: the number of distinct words in a text, typically counted as whitespace- and punctuation-separated runs of letters and digits.

Character Count: the total number of characters in a text, usually reported both including and excluding spaces.

Sentence Count: the number of sentences, detected by punctuation that ends a sentence (., !, ?) while excluding abbreviations and decimal numbers.

Paragraph Count: the number of distinct text blocks, typically separated by a blank line or a significant vertical gap on the page.

Reading Time: an estimate of how long a document takes to read silently, based on an average words-per-minute reading speed.

Speaking Time: an estimate of how long a document takes to read aloud, based on an average spoken words-per-minute pace.

OCR (Optical Character Recognition): the process of detecting text within a scanned image and converting it into selectable, searchable characters.

Searchable PDF: a PDF containing an actual text layer that can be selected, copied, and searched — as opposed to a scanned image with no underlying text.

Unicode: the international standard for representing text in virtually every writing system, allowing accurate word and character counting in any language.

Text Extraction: the process of reading the underlying text content out of a PDF's internal structure, rather than treating the page as a flat image.

Keyword Frequency: a count of how many times each distinct word appears in a document, typically ranked from most to least common.

Lexical Diversity: the ratio of unique words (vocabulary size) to total words — a higher ratio means less word repetition.

Document Analytics: the broader set of statistics and insights — length, structure, readability, keyword patterns — that describe a document beyond its raw word count.

Under the Hood

The exact counting logic this tool runs

No black box — this mirrors the real tokenizer, sentence splitter, and readability formulas that run in your browser the moment your PDF finishes extracting.

Word tokenizer — real regex, Unicode-aware: /[\p{L}\p{N}][\p{L}\p{N}'’\-]*[\p{L}\p{N}]|[\p{L}\p{N}]/gu Flesch Reading Ease: 206.835 − 1.015 × (words / sentences) − 84.6 × (syllables / words) Flesch-Kincaid Grade Level: 0.39 × (words / sentences) + 11.8 × (syllables / words) − 15.59 Reading speeds — words per minute: Silent Reading = 238 wpm Presentation (aloud) = 130 wpm Speed Reading = 450 wpm Speaking Time = 150 wpm
ComponentWhat it actually does
Word regexMatches a Unicode letter/digit run that may contain an internal apostrophe or hyphen — works across non-English scripts, not just ASCII.
Abbreviation list~28 common abbreviations (Mr, Dr, e.g., i.e., U.S., Inc., etc.) are protected from being misread as sentence endings before the sentence split runs.
Paragraph gap thresholdA vertical gap greater than 1.7× a page's median line-to-line spacing is treated as a paragraph break, not just a line wrap.
Syllable counterA vowel-group heuristic feeds both Flesch formulas — tuned for English; other scripts default to one syllable per word.
Stopword filter~120 common English words (the, and, of…) are excluded from the keyword frequency table unless explicitly re-enabled.
Scanned-PDF detectionIf total extracted non-whitespace characters fall under 10 across the whole document, the tool shows the OCR notice instead of a misleading all-zero dashboard.

FAQ

People Also Ask About PDF Word Counting

Answers to the most common questions about the PDF Word Counter tool

A PDF word counter is a tool that extracts the text inside a PDF and counts its words, characters, sentences, and paragraphs. PDFcrest's version also adds reading time, a readability score, and keyword frequency — full document analytics, not just a single total.
Upload your PDF to PDFcrest's Word Counter and the word count, along with a full statistics dashboard, appears automatically within seconds — no clicking a separate "count" button required, since extraction and analysis run the moment the file is read.
Yes. PDFcrest counts words entirely in your browser using the open-source PDF.js library — no Adobe software, account, or subscription is required, and the tool is completely free.
Word counters are accurate to within a small margin of each other, since different tools use slightly different rules for hyphens, decimals, and formatting artifacts. PDFcrest's tokenizer is Unicode-aware and documented, and counts should closely match Microsoft Word or Google Docs on the same text.
Not directly — a scanned PDF is just an image with no underlying text to extract. PDFcrest detects this and shows a notice suggesting OCR PDF first, which adds a searchable text layer that can then be counted.
Character count is the total length of the extracted text, reported two ways: including all spaces, and excluding whitespace. Both counts are Unicode-aware, so accented letters and non-Latin scripts are counted correctly.
Yes. Paragraphs are detected from vertical spacing between lines of text on the page — a gap noticeably larger than the normal line height is treated as a paragraph break, and the total count is shown alongside words and sentences.
Bold, italic, and font changes don't affect the count — only the underlying text matters. Unusual layouts (tables, multi-column text) can occasionally affect line and paragraph reconstruction, though word and character totals stay accurate.
With PDFcrest, yes — the file is read directly in your browser's memory and never transmitted anywhere. There's no upload step to intercept, since the tool makes no network requests once the page itself has loaded.
Yes. Open pdfcrest.com in any modern mobile browser (Safari on iPhone, Chrome on Android), upload your PDF, and the same full analytics dashboard appears — no app installation required.
Yes, completely free with no signup, no daily limit, and no watermark on exported statistics.
No. Your PDF is loaded and processed entirely in your browser's memory using PDF.js. It is never uploaded, stored, or seen by PDFcrest — closing the tab discards everything.
Yes. Export the full analysis as CSV (page table and summary), JSON (complete structured data including top keywords), or TXT (a readable report) — all generated locally, with nothing uploaded to produce them.
Not directly — a PDF that requires a password to open needs to be unlocked first with PDFcrest's Unlock PDF tool, since the browser can't read text it can't decrypt.
Yes, for word and character counts — the tokenizer is Unicode-aware and works across virtually any writing system. The readability score is an exception: it's tuned for English and less meaningful on non-Latin scripts.
The page needs an initial connection to load its HTML, CSS, and JavaScript, but once loaded, counting words and exporting results makes no further network requests — no data about your file is ever sent anywhere.
Yes. Pages are extracted incrementally with a live progress dashboard, and the browser yields control periodically so the tab never freezes. Files up to roughly 750MB are supported before a memory warning appears.
Yes. Choose Silent Reading (238 wpm), Presentation/Reading Aloud (130 wpm), or Speed Reading (450 wpm) from the dropdown, and the reading-time estimate updates instantly using that pace.
Speaking time estimates how long the document would take to read aloud at a typical conversational pace of about 150 words per minute — useful for estimating a presentation, voiceover, or narration length.
Using the standard Flesch Reading Ease and Flesch-Kincaid Grade Level formulas, based on average sentence length and average syllables per word — the same formulas referenced in U.S. federal plain-language writing guidance.
Word count measures distinct words; character count measures individual letters, digits, punctuation, and (optionally) spaces. Character count is often used for limits like tweet or SMS length, while word count is more common for essays and manuscripts.
Yes. The Keyword Search tool finds every occurrence of a word or phrase, shows which pages it appears on with a per-page count, and displays context snippets around up to 30 matches.
They're always included in the total word count. In the keyword frequency table specifically, common stop words are excluded by default so the list reflects meaningful content — toggle "Include common stop words" to see raw frequency instead.
Yes. Any modern browser on macOS or Windows — Chrome, Firefox, Safari, or Edge — works, since the tool runs as a web page with no installation required.
The tool detects this and shows a clear notice explaining the PDF appears to contain scanned images rather than selectable text, with a link to run OCR PDF first — it never silently shows a misleading zero-word result.
No. This is a read-only analysis tool — it only reads text out of your PDF to compute statistics. Your original file is never altered, re-saved, or uploaded.
Different tools use slightly different tokenization rules for hyphens, decimal numbers, and formatting artifacts, so small differences (typically under 1%) are expected and don't indicate an error in either count.

References & Further Reading

Sources this page is based on

Page last updated: · Written and maintained by the PDFcrest team. The counting logic shown above mirrors the real code running on this page.