SSkrubly
en▼Open a tool
← All articles

September 29, 2026 · 7 min read

Shrink a Scanned Document Without Making It Unreadable

Government and bank portals often cap uploads at a few megabytes. How to get a scanned contract or ID under the limit while keeping the text legible.

A scanned contract being checked against an upload size limit

A four-page scanned contract comes out at 18 MB. The portal accepts 3 MB. You compress it hard, it goes through, and three weeks later you get a letter saying the document was illegible and you need to submit it again. That round trip is the real cost, so it's worth understanding what compression actually does to a scan before you pick a setting.

Why scans are so much bigger than documents

A PDF exported from Word is mostly vector text — instructions for drawing letters, which cost almost nothing. A scan is a photograph of a page. Every letter is pixels, and at 600 DPI in colour that is an enormous amount of data for what is, visually, black marks on white paper.

Fix it at the scanner before you compress anything

Rescanning takes two minutes and beats any compression, because you avoid throwing away detail twice. If you still have access to the paper, do this first.

DocumentScan atMode
Plain text contract, form, letter200–300 DPIGreyscale
ID card, passport page, driving licence300 DPIColour — it is usually required
Anything with a stamp, seal or signature in colour300 DPIColour
Page with photographs or diagrams300 DPIColour
Handwritten notes, faint carbon copies300–400 DPIGreyscale, raise contrast

The trade-off nobody explains

PDF compression comes in two fundamentally different flavours, and knowing which one you're using is the whole game.

In Skrubly's compressor that is the difference between the levels, and we'd rather say it plainly than let you find out after submitting: Light re-saves the PDF losslessly and keeps the text layer intact. Balanced and Strong rasterise each page — around 120 DPI and 96 DPI respectively — which shrinks a bulky scan dramatically but produces an image-only PDF. For a scan that was already an image, that's often a fair trade. For a PDF with real text in it, it is a downgrade you can't undo.

Try Light first, then step up only if you're still over the limit.

Open the File Compressor

A working order of attack

What "unreadable" looks like in practice

Judge the result on the details that matter to whoever reads it, not on the page as a whole at 25 percent zoom. Zoom to 100 percent and check: small print and footnotes, digits in reference numbers and dates, the difference between 8 and 3, accented characters, signatures, and the machine-readable lines on an ID document. Those lines are the first thing to smear, and they're often the first thing the other side checks.

If the text has to stay searchable

Some portals ask for a searchable PDF, or you may simply want to find a clause later. That means OCR: a pass that recognises the letters in the scan and stores them as a text layer behind the image. Most scanner apps do it automatically, Adobe Acrobat does it, and free options exist on the desktop. Run OCR before heavy compression — recognition on a mushy image is much worse — and if you rasterise afterwards, expect to lose the text layer and have to do it again.

ToolGood forNote
Skrubly compressorQuick size reduction in the browserNothing leaves your device
Adobe AcrobatOCR and fine controlPaid for most features
GhostscriptBatch work on a desktopCommand line, very effective
Scanner app OCRGetting a searchable scan in one stepQuality depends on the original

Privacy, because this is exactly the sensitive category

Contracts, bank statements and ID documents are the files you should be most careful about handing to a random website. A browser-based tool that processes locally never sends the document anywhere; a server-based one does, and you are trusting its retention policy. Whichever you use, check the file's own properties before submitting: PDFs carry an author name, a title and the software that produced them, and scans of ID documents sometimes carry the scanner model or the phone's details. Clearing those fields takes seconds.

Clear author, title and software fields from a PDF before you submit it.

Document metadata guide