GetFormatted

OCR a scanned PDF

Read the text in a scanned PDF so it can be searched and copied, and download the text too.

How it works

  1. Add your scanned PDF.
  2. Press Read the text.
  3. Wait a few seconds a page.
  4. Download the searchable PDF, or the plain text.

What you get

  • A searchable PDF: the original pages, unchanged to look at, with the recognised words laid invisibly over them so you can search and select text.
  • A plain text file of everything that was read, page by page.
  • A note of how sure the reader was, so you know whether to trust the result or rescan.

Getting a good result

  • Scan straight, in good light or on a flatbed scanner, at 200 to 300 dpi.
  • Dark print on light paper reads best. Handwriting, faint print and curved pages do not.
  • English text is read. Other languages and alphabets are not read yet.
  • Tables and columns are not rebuilt, so copied text may come out in a different order from the page.

Check important numbers and names against the page. OCR can mistake characters such as 0 and O or 1 and l.

Your scan stays on your device

The text reader runs inside your browser and is served from this site. The first time you use it, about 15 MB is downloaded and kept by your browser, and after that it opens quickly. Your PDF is never sent anywhere.

Good to know

  • English text only. Other languages and alphabets are not read yet.
  • Accuracy depends on the scan: straight, sharp, high-resolution pages read well, and handwriting, curved pages and faint print do not.
  • The words are laid invisibly over the original pages, so the file looks the same. Tables and columns are not rebuilt, and the order of copied text can differ from the page.
  • The first use downloads the reader (about 15 MB) from this site, and your browser keeps it. It takes a few seconds a page.

Frequently asked questions

Is it free to OCR a PDF here?
Yes, for files up to 25 MB and 10 pages. GetFormatted Pro raises this to 100 MB and 200 pages.
Is my scan uploaded?
No. The text is read in your browser and the file stays on your device.
Which languages does it read?
English only for now.
Will the PDF look different afterwards?
No. The pages look the same. The recognised words are added invisibly on top, so the text can be searched and selected.
Why is some text wrong?
OCR depends on the scan. Blurry, tilted or faint pages, and handwriting, read poorly. A sharper, straighter scan helps.

Related