How to make a scanned PDF searchable online free — without uploading
Converting a scanned PDF to searchable PDF online is free with Number My Document. Upload your scanned document, select OCR mode, click Process, and download a searchable PDF — all processed in your browser, your file never leaves your device. Unlike Adobe Acrobat (paid), OCR.space (3-page limit, watermark), or Smallpdf OCR (server upload), this tool is completely free with no restrictions.
How Think Brain 4-pass OCR works
Most free OCR tools run a single pass: render the image, send to Tesseract, output text. Think Brain OCR runs up to four preprocessing passes per page and picks the highest-confidence result — dramatically improving accuracy on difficult handwritten documents, faded inks, and poor-quality scans.
Pass 1 — CLAHE normalisation + Otsu binarization
The page image is divided into 32x32 pixel tiles. Each tile is independently contrast-stretched to use the full 0-255 range, cancelling uneven lighting, shadow, yellowing, or faded areas. A global Otsu threshold then binarizes the result into clean black-on-white ink strokes. A 2-pixel morphological dilation fills breaks in thin pen strokes before Tesseract sees the image.
Pass 2 — Strong contrast boost
If pass 1 confidence is below 75%, a second pass applies a 3x linear contrast boost centred at mid-grey. This targets washed-out ballpoint pen, light pencil, or overexposed mobile phone scans where standard binarization fails to separate ink from paper.
Pass 3 — Image inversion
Below 58% confidence, the image is inverted — turning dark backgrounds light and light ink dark. This catches white chalk on blackboard, white correction fluid notes, or pencil writing on cream or grey paper.
Pass 4 — Block-text PSM 6 fallback
If confidence remains below 48%, the page segmentation mode switches to PSM 6 (single uniform block of text), avoiding column or sparse-text detection errors on unusual handwriting layouts. The highest-confidence result across all four passes is used to build the final searchable PDF text layer.
What documents OCR best
Printed or typewritten PDF scans — attendance registers, government forms, invoices, printed letters — typically achieve 85-98% OCR confidence. Clear handwritten documents on white paper with blue or black ink reach 70-90%. Very poor handwriting, light pencil on grey paper, or heavily damaged documents typically achieve 40-70% with 4-pass mode — still enough to create a partially searchable text layer for keyword search.
When to use searchable PDF vs typed PDF output
Choose same-format searchable PDF when submitting to courts, government portals, archives, or any system that requires the original scan appearance with added searchability. Choose clean typed PDF when you need to copy-paste the content, share a text extract, or the document layout is simple enough that a typed reconstruction is useful without the scanned image.
OCR for legal, medical, and government documents
Legal professionals use searchable PDFs to make discovery documents keyword-searchable in document review software. Medical teams convert handwritten patient notes to searchable format for EMR integration. Government departments make scanned circulars and forms searchable for internal portals. All of these use cases benefit from browser-only processing — your sensitive documents never leave your device.
Compare: Number My Document OCR vs competitors
vs OCR.space: OCR.space adds a watermark on the free tier and limits conversions to 3 pages per request. No such restrictions here. vs Smallpdf OCR: Smallpdf uploads your file to their servers. This tool processes in your browser. vs Adobe Acrobat: Acrobat requires $14.99/month for OCR. This tool is free. vs ilovepdf OCR: ilovepdf uploads files and has daily limits. This tool has no daily limits and processes locally.