# Scanned PDF or text PDF: choose the right next step https://thedollscout.com/learn/scanned-pdf-vs-text-pdf Published by TDS Document Scout Updated: 2026-09-25 A PDF can contain actual text, page images or both. Being able to see words does not mean those words are extractable or usable with assistive technology. Extract a page of text. If the result is empty, inspect the page visually: it may be scanned, blank or contain unsupported content. Empty extraction alone does not prove it is a scan. For a scan, run OCR in a suitable editor, then check names, numbers, punctuation and reading order. OCR output can contain convincing errors. A text layer is only the first step. Review headings, alternatives and structure separately. This tool does not perform OCR or certify the resulting PDF. Reference material - W3C: PDF reading order: https://www.w3.org/WAI/WCAG22/Techniques/pdf/PDF3 - W3C: text alternatives in PDF: https://www.w3.org/WAI/WCAG22/Techniques/pdf/PDF1 - Mozilla PDF.js: https://mozilla.github.io/pdf.js/