Table of Contents
The 7 Culprits
1. High-Resolution Images
The #1 cause. A 12 MP phone photo (4000×3000) embedded at full resolution = 5–10 MB. A 20-page report with one photo per page = 100–200 MB. Most PDF creators (Word, Canva, InDesign, scanner apps) don't downsample by default. The image looks the same on screen at 150 DPI — but the file carries 10× the data.
2. Scans Saved as Images
Scanner apps and desktop scanners often save each page as a high-DPI image (300–600 DPI) wrapped in a PDF. A 50-page contract scanned at 300 DPI = 50–150 MB. At 600 DPI = 200–600 MB. The text is unsearchable (no OCR), and every page is a fat raster image.
3. Embedded Fonts (Not Subsetted)
A full OpenType font = 100–500 KB. If your document uses 5 fonts and embeds them fully = 0.5–2.5 MB just for glyphs. You might only use 20 characters per font. Subsetting (embedding only used glyphs) drops this to 20–50 KB per font. Many creators (especially older Word versions, some design tools) embed full fonts by default.
4. Metadata & Attachments
XMP metadata (Adobe private data), document info dictionary (author, title, keywords, timestamps), thumbnail images (one per page!), piece info (Illustrator/Photoshop round-trip data), embedded files (attachments), form field appearances — routinely 100 KB–5 MB of dead weight.
5. Redundant Content
Duplicate images (same logo on every page embedded separately instead of referenced once), unused form fields, hidden layers, annotation appearances, old revision streams (PDF supports incremental updates — every save appends, old data stays). A document saved 20 times in Acrobat can carry all 20 versions.
6. No Compression Applied
PDF supports Flate (zlib), LZW, JPEG, JPEG2000, JBIG2 compression for streams. Many creators write uncompressed streams by default — especially for text/content streams. A 500 KB text stream compresses to 50 KB with Flate. That's 90% savings for free.
7. Vector Graphics Bloat
CAD exports, Illustrator art, complex charts — thousands of path commands. A single detailed vector map can be 10–50 MB. PDF isn't great at compressing path data. For screen use, rasterize to image (PDF→JPG→PDF) — often 10× smaller.
How to Diagnose What's Making Your PDF Big
File Inspector Tips
You don't need forensic tools. Quick checks:
- Page count vs size: 10 pages, 50 MB? Almost certainly images/scans. 100 pages, 2 MB? Text-only, well compressed.
- Zoom test: Open in viewer, zoom 400%. Text sharp? Vector. Text pixelated? Raster image. Photos blurry at 100%? Low-res source (can't fix). Photos sharp at 400%? High-res source (compressible).
- Select text: Can you select? Text layer exists. Can't select? Scanned/image-only.
- Properties → Fonts: In Acrobat Reader (File → Properties → Fonts), check "Embedded" vs "Embedded Subset." Full embeddings = bloat.
Advanced: PDF Debuggers
For deep analysis: PDF Association tools, pikepdf (Python), qpdf (CLI). These show object-level breakdown: stream sizes, font sizes, image counts, compression filters.
How to Fix It
Quick Fix — Compress Online
Open Compress PDF. Choose Moderate (recommended). Download. Done. Handles causes 1, 3, 4, 6 automatically. Zero upload. Free.
Re-Encode Images via PDF→JPG→PDF
For stubborn image bloat (cause 1, 2, 7): PDF to JPG at 150 DPI → JPG to PDF. This re-rasterizes every page at controlled resolution. Vector bloat (cause 7) disappears — paths become pixels. Scans (cause 2) get downsampled. Result: predictable, small, screen-perfect.
Split the File
If a 200 MB file won't compress below 50 MB (too many high-res images), split it: Split PDF into 10 × 20-page files. Each ~5 MB. Email-friendly. Or extract only needed pages: Extract PDF Pages.
Remove Unnecessary Pages
Blank pages, duplicate covers, outdated appendices: Remove PDF Pages. Every page removed = its images/fonts gone.
Expected Sizes by PDF Type
{`// Rule-of-Thumb: Expected PDF Sizes (After Moderate Compression)
| Document Type | Pages | Expected Size | Notes |
|----------------------------|-------|---------------|----------------------------------|
| Text-only (contract, letter) | 10 | 200–500 KB | Fonts subsetted, streams compressed |
| Text + a few charts | 20 | 1–3 MB | Vector charts compress well |
| Mixed text + photos | 15 | 3–8 MB | Photos downsampled to 150 DPI |
| Scanned (300 DPI, B&W) | 50 | 5–15 MB | JBIG2 compression ideal |
| Scanned (300 DPI, Color) | 50 | 20–50 MB | JPEG compression, consider 150 DPI |
| Presentation (slides) | 30 | 2–6 MB | Rasterize for smaller |
| High-res portfolio | 20 | 10–30 MB | Consider lower DPI for web |
| CAD/Vector heavy | 5 | 5–20 MB | Rasterize via PDF→JPG→PDF |`}
If your PDF is 5–10× these numbers, something's wrong. Diagnose → fix.
FAQ
Why is my PDF so large?
Most likely: high-res images, unsubsetted fonts, metadata, or no compression applied. Run through Compress PDF (Moderate) — fixes 90% of cases.
How do I make a PDF smaller?
Compress PDF (Moderate). If still big: PDF to JPG (150 DPI) → JPG to PDF. Or Split PDF.
Why is a PDF bigger than the original Word doc?
Word stores content semantically (text + styles). PDF stores rendering commands (glyphs at coordinates). Fonts embed. Images don't downsample. Metadata adds. A 500 KB .docx → 5 MB PDF is normal without compression.
Does compressing a PDF reduce quality?
Moderate: no visible loss on screen. High: visible on photos. Low: lossless. See How to Compress a PDF Without Losing Quality.
Why is my scanned PDF so big?
Scans = full-page images at 300–600 DPI. 50 pages × 300 DPI color = 100+ MB. Fix: Compress PDF (High) or PDF to JPG (150 DPI) → JPG to PDF.
Why is my PDF 100 MB?
Almost certainly high-res images or scans. Check page count vs size. Run Compress PDF.
How do I reduce PDF file size for email?
Compress PDF → if still > 25 MB, Split PDF or re-encode. See How to Share Large PDFs by Email.
Got a bloated PDF? Drop it in Compress PDF — watch it shrink in seconds. Free. Private. No upload.