Scanned PDF Too Big? Cut It to 1/5 the Size
You scan a contract or a stack of receipts and end up with a 50MB or 100MB monster of a PDF for what should be a few pages. Email rejects it, cloud uploads crawl, your phone freezes when you try to open it.
Why this happens is well-understood, and the fix is practical: with the right approach, you can shrink scanned PDFs to 1/5 to 1/10 of their original size while keeping them readable. Here's how.
Why scanned PDFs are so big
Regular PDFs (created from Word, for example) store text as text — that's lightweight. Scanned PDFs are different: every page is an embedded image, so the image size becomes the file size.
Three settings multiply that out:
- Resolution (DPI) — 600 DPI carries 4x the data of 300 DPI
- Color — full color is roughly 3x the size of grayscale
- Image format — uncompressed or PNG is 5–10x larger than JPEG
Multiply those out and you get the picture: 600 DPI, full color, PNG produces a file 60–120x larger than 300 DPI, grayscale, JPEG. And many scanners ship with the heavy defaults turned on.
Method 1: Optimize the scan settings (biggest impact)
If you're about to scan, the cheapest win is to fix it at the source. You can apply this to already-scanned files too — you just have to re-scan.
Recommended settings
| Use case | Resolution | Color | Format |
|---|---|---|---|
| Reading documents | 300 DPI | Grayscale | JPEG |
| Color is essential (figures, diagrams) | 300 DPI | Color | JPEG |
| Photos / artwork | 600 DPI | Color | JPEG (high quality) |
| OCR (text recognition) | 400 DPI | B&W or grayscale | JPEG |
300 DPI / grayscale is the sweet spot for documents. People often crank up to 600 DPI / color thinking "more is better", but for plain reading it's overkill — it just bloats the file.
Method 2: Re-compress an already-scanned PDF
If the scan already happened, you can still compress after the fact. Adobe Acrobat, online compression services, and browser-based tools all work — the differences are mostly about privacy and convenience.
Pros
- No re-scan needed — work with what you have
- Reductions like 50MB → 10MB are realistic
- Quality loss is usually invisible at normal viewing zoom
Watch out
- Already-compressed PDFs offer limited additional savings
- Online services upload your file — careful with confidential documents
Compress scans without uploading them
StayPDF's compression tool processes inside your browser. Even sensitive scans like contracts stay on your device.
Compress nowMethod 3: Optimize images individually, then build the PDF
If you have time and want maximum compression, the strongest approach is to optimize images one by one and then assemble them into a PDF:
- Resize and re-compress each scanned JPEG individually
- Use an image-to-PDF tool to combine them into one file
It's more work, but you get fine-grained control — useful when archiving large batches of scanned material. The image-to-PDF tool takes optimized images and stitches them into a single PDF in one step.
Pairing settings together
The savings come from stacking the right choices, not from any single one. As a rough rule of thumb for typical office documents:
- Half-letter / A5 receipts: 200 DPI grayscale JPEG is plenty. The text is large enough that 200 DPI stays sharp.
- Standard letter / A4 contracts: 300 DPI grayscale JPEG. Going from 600 DPI color to this combination usually delivers a 10–20x reduction.
- Documents with handwriting or signatures: 300 DPI grayscale is still fine. Pen strokes survive grayscale because they have strong contrast against the page.
- Documents with stamps or color-coded marks: drop to 250–300 DPI but keep color. The color information is the point of the document.
What about OCR?
Many people scan documents specifically to make them searchable. OCR (optical character recognition) extracts text from the image and embeds it as a hidden text layer. A few practical notes:
- Higher DPI helps OCR accuracy — 400 DPI is the sweet spot. Below 200 DPI, accuracy drops noticeably.
- Grayscale is fine for OCR — engines analyze contrast, not color. Don't waste storage on color unless the document needs it.
- Compress after OCR, not before — heavy JPEG compression introduces artifacts that hurt recognition.
- OCR adds very little to file size — the embedded text layer is tiny compared to the image. Don't skip OCR to save space.
Compression limits — when it just won't shrink more
Not every scanned PDF can be reduced further. Compression returns diminish in a few situations:
- PDFs that have already been heavily compressed
- Photos or artwork where preserving detail matters
- PDFs full of watermarks or complex backgrounds
In those cases, consider the alternative: only send the pages that actually need to go. The PDF split tool lets you keep the necessary pages and drop the rest, often a much bigger win than compression alone.
Pick by situation
Set 300 DPI / grayscale / JPEG. That alone cuts size to a fifth or less.
Run it through the compression tool. For confidential scans, pick a browser-based tool.
If compression alone isn't enough, use split to send only the pages that matter.
Common scanner mistakes that bloat files
If your scans regularly come out larger than they should, the cause is almost always one of these defaults left untouched on the device:
- "Highest quality" preset — most multifunction printers default to 600 DPI color. Switch the default profile to 300 DPI grayscale for everyday paperwork.
- PNG output — some scan apps save individual page images as PNG, then wrap them in a PDF. PNG is lossless and 5–10x larger than JPEG for the same visual content.
- Auto color detection — when a document has even one stamp or colored mark, auto mode flips the whole batch to color. Force grayscale unless you actually need color.
- Duplex blank pages — duplex scanning produces blank back-of-page images you don't need. Enable "skip blank pages" in the scanner driver, or trim them later with PDF organize.
Quick comparison: before and after
To set realistic expectations, here are typical real-world reductions on a 20-page contract scan:
| Source settings | Original size | After re-compression |
|---|---|---|
| 600 DPI color PNG | 120 MB | 8–15 MB |
| 600 DPI color JPEG | 40 MB | 6–10 MB |
| 300 DPI grayscale JPEG | 6 MB | 3–4 MB |
The lesson is that fixing the source beats post-hoc compression every time. Re-compression is for files you can't re-scan.
FAQ
- Are scanned PDFs always larger than text PDFs?
- Yes — every page is an embedded image, so the total comes down to image size. A one-page Word PDF might be 50 KB; the same page scanned at 600 DPI color easily passes 5 MB.
- Will compression damage the text?
- Aggressive settings can soften thin strokes. The balanced preset in StayPDF compress targets a level where body text stays sharp at 100% zoom. Always spot-check a couple of pages on important documents.
- Can I compress scans for free?
- Yes. StayPDF's tool is free with no usage cap and runs entirely in the browser, so confidential scans never leave your device.
- What if compression isn't enough?
- Drop unnecessary pages with PDF split or PDF organize. Sending only the relevant 5 of 30 pages beats compressing the full 30.
Bottom line
The "scanned PDF is huge" problem yields to a two-pronged approach: fix the scan settings going forward, and re-compress what you already have. 300 DPI grayscale is enough to read, and existing files can shrink by 2–5x with StayPDF's compression tool.
For sensitive scans — contracts, medical records, HR documents — choose a browser-based tool that doesn't upload. Your file stays on your device, with no third-party server involved.