How this tool computes its result
The checker opens the PDF locally, extracts the text content of each page, normalizes it into reading lines, and counts words and non-space characters. A page with fewer than 20 extracted non-space characters is labeled likely scanned or image-only; all other pages are labeled as having a text layer. The result reports document-wide searchability percentage, total extracted characters, and a page-by-page grid so mixed PDFs do not hide image-only appendices behind searchable front matter.
