logo
menu

How to Translate a Scanned PDF When Text Isn’t Selectable

By Janet | July 28, 2026

If you cannot select the words in a PDF, you are probably dealing with a scanned PDF. That means each page is stored like an image. To translate it, you need one extra step: OCR first, translation second.

OCR, short for optical character recognition, turns image-based text into machine-readable text. Once the text is recognized, a translator can convert it into another language and, depending on the tool, keep part of the original layout.

This guide shows how to translate a scanned PDF without guessing your way through five different apps. You will learn how to check whether your PDF needs OCR, how to use an online scanned PDF translator, when desktop OCR software is safer, and how to avoid the common formatting problems that make translated documents hard to use.

Translate a scanned PDF with OCR and AI translation

Quick Answer: How to Translate a Scanned PDF

To translate a scanned PDF, use this workflow:

  1. Open the PDF and try to select a sentence.
  2. If the text cannot be selected, run OCR on the PDF.
  3. Translate the recognized text or upload the OCR-ready file into a document translator.
  4. Review names, numbers, tables, stamps, and headings manually.
  5. Download the translated file and compare it with the original before sharing.

For a simple scanned PDF under 10 MB, the fastest option is usually an online document translator with OCR support. For legal, medical, financial, or confidential documents, use a professional workflow or desktop OCR software and have a human review the output.

First, Check Whether the PDF Is Scanned or Digital

Before choosing a tool, do a 10-second test.

Open the PDF in your browser, Preview, Adobe Acrobat, or any PDF reader. Try dragging your cursor across one sentence.

If you can highlight individual words, the file is a digital PDF. You can usually translate it directly.

If the whole page behaves like one flat picture, or your cursor cannot select text at all, the PDF is scanned. You need OCR before translation.

This distinction matters because many translation tools can upload PDFs, but they may only translate text that already exists inside the document. A scanned page can look like a normal PDF while still being unreadable to a translator.

Method 1: Use an Online Scanned PDF Translator

For everyday documents, the simplest route is to use an online translator that supports PDF upload and OCR. This works well for class handouts, travel forms, product manuals, simple certificates, internal notes, and research pages where you mainly need a readable translation.

Lynote’s Document Translator supports PDF, DOCX, PPTX, and XLSX files, includes OCR for scanned PDFs, supports 135+ languages, and handles files under 10 MB. Use it when you want a direct upload-to-download workflow instead of manually copying OCR text into a separate translator.

Step 1: Upload the Scanned PDF

Choose the document upload mode and upload your PDF. If the scan is tilted, blurry, too dark, or full of shadows, clean it up first when you can. OCR is much more reliable when the letters are sharp and the page is straight.

Upload a scanned PDF in Lynote Document Translator

Good candidates for online OCR translation:

  • typed documents;
  • clean scans from a scanner app;
  • simple one-column pages;
  • school handouts;
  • manuals, reports, and forms with readable text.

Weak candidates:

  • handwriting;
  • heavily stamped documents;
  • very small text;
  • rotated pages;
  • old photocopies with noise;
  • complex tables where exact layout matters.

If the document contains sensitive personal, legal, or medical information, check your organization’s policy before uploading it to any online tool.

Step 2: Choose the Target Language

Set the source language to auto-detect if you are not sure, then choose the target language. If the scan contains two languages on the same page, review the result more carefully because OCR and translation may treat mixed-language sections unevenly.

Choose a target language before translating a scanned PDF

For best results, avoid translating a poor scan into a language with very different sentence length without checking the layout. For example, translating a compact form into German or a narrow table into English may cause line breaks or spacing changes.

Step 3: Download and Review the Translation

After the translation finishes, download the result and compare it with the original. Do not only check the first page. OCR errors often appear in the places people skip: footnotes, small table cells, page headers, names, dates, and numbers.

Download the translated document after OCR and translation

Use this review checklist:

  • Are names, addresses, dates, and numbers correct?
  • Did tables keep their row and column meaning?
  • Did headers, footnotes, and stamps translate or remain visible?
  • Are any words split incorrectly because of line breaks?
  • Does the translated file preserve enough layout for your purpose?

If you need a polished business or official version, treat the online result as a working draft and send it for human review.

Method 2: OCR First, Then Translate

If your tool cannot translate the scanned PDF directly, split the job into two steps.

First, use an OCR tool to convert the scanned PDF into searchable text or an editable document. Adobe Acrobat, ABBYY FineReader, Microsoft Lens, and other OCR tools can help with this. Then paste the extracted text into a translator, or save the OCR output as a document and upload it into a PDF/document translation tool.

This route is better when:

  • the scan quality is low;
  • you need to correct OCR text before translation;
  • the document has many tables;
  • you need to work offline;
  • you need stronger control over privacy;
  • you want to keep an editable copy for later.

It takes longer, but it gives you more control. For important documents, that extra control is usually worth it.

Can Google Translate a Scanned PDF?

Google Translate can handle many document uploads, but scanned PDFs are the tricky case. If the PDF is image-only, the text may not translate correctly unless OCR has already recognized it.

The practical rule is simple: if Google Translate gives you blank output, untranslated image text, or broken formatting, OCR the PDF first. After OCR, you can try Google Translate again or use a document translator built for uploaded files.

If your goal is specifically to compare Google’s options, see this related guide on how to translate a PDF with Google.

How to Improve OCR Accuracy Before Translation

Most scanned PDF translation problems begin before translation. The OCR engine cannot translate what it cannot read.

Improve the source file first:

  • straighten tilted pages;
  • crop out borders, fingers, or shadows;
  • rescan at a higher resolution when possible;
  • split double-page scans into single pages;
  • avoid photos taken at an angle;
  • remove password restrictions if you own the document and are allowed to edit it;
  • compress only after OCR if compression makes the text blurry.

If the scan is a photo of a page, a scanner app can often improve contrast and deskew the image before you convert it to PDF.

How to Preserve Layout When Translating a Scanned PDF

Layout preservation is the hard part. OCR can recognize text, and translation can convert meaning, but the translated words may not fit the same physical space.

This is why forms, tables, brochures, contracts, certificates, and multi-column reports need extra review. A translated sentence may become longer than the original. A table label may wrap onto two lines. A stamp or signature may remain as an image. A footnote may move.

For layout-sensitive documents, use this order:

  1. OCR the scanned PDF.
  2. Translate the document.
  3. Compare the translated file with the original page by page.
  4. Fix tables, labels, names, and line breaks manually.
  5. Export the final PDF only after review.

If the visual layout is more important than speed, use desktop OCR or a professional translation workflow instead of relying on a quick online conversion.

Which Method Should You Choose?

SituationBest methodWhy
You need a quick readable versionOnline OCR + translationFastest path for simple documents
You need the original file format preservedAI PDF Translator or document translatorBetter than copying text into a blank page
The scan is messyOCR first, then correct text manuallyTranslation will only be as good as the recognized text
The document is confidentialOffline OCR or approved internal toolReduces upload and compliance risk
The document is legal, medical, or officialProfessional translation reviewAccuracy matters more than speed
The file is large or many PDFs at onceDesktop OCR workflowMore control over pages and batches

For most casual use, start with the direct online workflow. If the result is messy, do not keep retrying the same translation button. Fix the scan or OCR text first.

Common Mistakes to Avoid

The most common mistake is uploading an image-only PDF to a regular translator and assuming the tool will read it. Some tools do; some do not. Always check whether OCR is part of the workflow.

The second mistake is trusting the output without checking names and numbers. OCR often confuses similar characters, especially in small text. A translated paragraph can look fluent while still carrying a wrong date, code, or name.

The third mistake is trying to preserve perfect layout from a poor scan. If the original is skewed or blurry, the translated file will probably need manual cleanup.

The fourth mistake is using online tools for documents that should not leave your device or organization. Convenience is useful, but privacy rules come first.

FAQ: Translating Scanned PDFs

How can I translate scanned documents?

Use OCR to recognize the text first, then translate the recognized text or upload the OCR-ready file to a document translator. If you use a tool with built-in OCR, the OCR and translation steps can happen in one workflow.

Why can’t I copy text from my scanned PDF?

Because the PDF likely stores each page as an image. A normal PDF reader can display the page, but it cannot select the words until OCR turns the image text into machine-readable text.

Can I translate a scanned PDF for free?

Some online tools offer free OCR or document translation with limits. Check file size, page count, language support, sign-up requirements, and whether the output keeps formatting before relying on it.

Will OCR keep the original formatting?

OCR recognizes text; it does not guarantee perfect layout. Some tools try to preserve fonts, images, tables, and spacing, but complex documents still need manual review.

What is the best way to translate a scanned PDF with tables?

Use OCR first, then check the table cells manually after translation. If row and column meaning matters, avoid copying plain text into a translator because it may destroy the table structure.

Can I translate handwritten scanned PDFs?

Handwriting is much harder than printed text. OCR may work for clear handwriting, but accuracy is usually lower. For important handwritten documents, manual transcription and human translation are safer.

Is it safe to upload a scanned PDF to an online translator?

It depends on the document and your rules. For school notes, public handouts, or non-sensitive materials, online tools may be fine. For contracts, IDs, medical files, financial records, or confidential work documents, use an approved secure workflow.

Conclusion

To translate a scanned PDF, do not start with the translation step. Start by checking whether the PDF has selectable text. If it does not, OCR is the missing step.

For simple documents, an online scanned PDF translator can save time. For messy scans, sensitive files, or layout-critical documents, OCR first, review carefully, and use a more controlled workflow.

The best result usually comes from a boring process: clean scan, accurate OCR, careful translation, and a final human check.