Why scanned PDFs become huge and why text turns blurry
A scanned PDF is not text. It is a set of photographs of pages wrapped inside a PDF container. Every page is an image, often captured at 300–600 DPI. That resolution is useful for archival work but far more than needed for screen reading or ordinary printing. The result is files that routinely exceed email limits (typically 20–25 MB) or upload portals.
When compression is applied too aggressively, the tool downsamples those images or raises JPEG compression beyond the point where letter edges stay sharp. Text that looked crisp at 100 % zoom becomes soft or blocky. The goal is not maximum size reduction at any cost. The goal is a file small enough to send while remaining fully readable.
The practical method that protects readability
Two tools work together:
- OCR PDF adds an invisible text layer so the document becomes searchable and selectable. It can also deskew pages and improve contrast on low-quality scans.
- Compress PDF performs adaptive image downsampling and strips redundant metadata while leaving the visual layout intact.
For pure size reduction of a clean scan, compression alone is usually enough. For a scan that also needs to be searchable, run OCR first, then compress. OCR does not magically shrink the file, but a cleaner, deskewed image often compresses more efficiently and the resulting text layer survives moderate compression.
Step-by-step: compress a scanned PDF without blurry text
Both tools run in any modern browser on Windows, Mac, Linux, iPhone, or Android. Files transfer over TLS and are automatically purged from temporary storage within 60 minutes. No account means no profile is linked to your document.
What actually keeps text sharp
| Setting / Approach | Typical effect on scanned PDF | Readability impact |
|---|---|---|
| Downsample to ~150 DPI (screen) | Large size reduction | Usually invisible at normal viewing distance |
| Downsample to ~200 DPI | Moderate reduction | Safe for most printing |
| Aggressive JPEG quality < 60 | Maximum reduction | Visible blur and artifacts around letters |
| Grayscale conversion (if original is B&W) | Extra size saving | Often improves contrast for text |
| Metadata and unused font stripping | Small but free reduction | None |
The Compress PDF tool applies adaptive downsampling rather than a fixed low-quality preset. That is why most scanned pages stay readable after a single pass. Extreme reduction is possible but not the default goal.
Common problems and quick fixes
Text looks soft after compression
The image was downsampled too far or JPEG quality was set too low. Re-process the original with a lighter compression level, or accept a larger file.
Upload fails or times out
Files larger than the tool’s practical headroom (approximately 100 MB on Compress) should be split first with the Split PDF tool, compressed in sections, then merged again if needed.
OCR accuracy is poor
Low-contrast, skewed, or handwritten pages reduce recognition rates. The OCR tool includes deskew and contrast enhancement, but very poor source scans still limit results. Re-scan at 300 DPI in grayscale when possible.
File is still too large for email
After compression, further options are to remove non-essential pages, convert color pages to grayscale, or split the document into multiple smaller files.
Related tools that fit the same workflow
- Split PDF — divide an oversized scan before compression so each piece stays under portal limits.
- Merge PDF — reassemble compressed sections afterward.
- Rotate PDF — correct pages that were scanned upside-down or sideways before OCR or compression.
- JPG to PDF — turn individual photo scans into a multi-page PDF first, then compress.
FAQ
Does compressing a scanned PDF always make the text blurry?
No. Moderate downsampling (to roughly 150–200 DPI) and sensible JPEG quality leave text sharp for screen reading and ordinary printing. Blur appears only when settings are pushed too far.
Should I run OCR before or after compression?
Run OCR first on the highest-quality scan you have. A clean, deskewed image produces better recognition and often compresses more efficiently afterward.
Is the process free and private?
Yes. Both tools are free, require no account, transfer files over TLS, and automatically delete uploads and outputs within 60 minutes. No profile is linked to your document.
What file size limit should I expect?
Compress PDF handles files up to approximately 100 MB in normal use. Larger scans can be split, compressed in parts, then merged.
Can I compress on my phone?
Yes. The same browser tools work on iPhone and Android. File selection uses the device’s native picker.
Final takeaway
Scanned PDFs are large because they store high-resolution images. Careful compression reduces that weight by downsampling images and removing unused data while leaving letter edges intact. When searchability also matters, add an OCR pass first. The combination of the free OCR PDF and Compress PDF tools on PDF Convex gives you a complete browser workflow with no signup and automatic file deletion.
Open the tools, upload once, check the result at 150 % zoom, and you will have a file that travels easily without looking like a bad photocopy.
