Will a screen reader actually read your PDF?
Choose a PDF and this tool checks the structure screen readers depend on: tags, language, title, headings, alt text on figures, tables, bookmarks, form fields and security settings. Your file is read inside your browser and is never uploaded.
Free · No signup · Your PDF never leaves your computer
Choose a PDF to check
Drop a PDF here
or choose one from your computer
One PDF at a time, up to 100 MB. Checked in your browser, never uploaded.
Your PDF is opened by code running in this page. Its contents and its file name are never sent anywhere. To read some documents, the PDF reader asks this site for its own character-map and font files, named after the encoding or font the document uses (a Japanese document, for example, asks for a Japanese character map), so GotAlt can see what kind of file it was but not the file.
Results
What this check can't tell you
This is a structural check. It looks at whether the things a screen reader relies on exist, not whether they are right.
A clean result is not a pass
- Whether alt text is accurate. It can see that a figure has alternative text, not whether the text describes the picture. A figure described as "image" still counts as having alt text.
- Whether the reading order is right. It can see that tags exist, not whether they are in the order a person would read the page. Listening to the document with a screen reader is the only real test, and NVDA is free.
- Colour contrast. Text and background colours are not read. Use the contrast checker for a pair of colours, or the image contrast checker on a screenshot of a page.
- PDF/UA conformance. This is not a PDF/UA conformance test, and a clean result is not a statement that the file meets WCAG or any law.
- Whether the text comes out as real words. Text in a font saved without a Unicode map counts as text here, but a screen reader can read it as nonsense. Copy a sentence from the PDF and paste it into a text editor to see what comes out.
- Content outside the tags. Images that were never tagged, text that has been turned into outlines and anything else outside the tag tree and the file's settings is invisible to it.
- Files it cannot open. Password-protected PDFs, and the contents of XFA forms.
- Pages after the first 200. Page-level checks stop at the first 200 pages. A longer document gets a notice saying it was only partly checked.
For a full check, use PAC, a free checker built around PDF/UA and WCAG that runs on Windows only, and veraPDF, an open-source validator that covers PDF/UA. Both are free. This tool is the quick, private first look before you reach for them.
How it works
-
You choose a PDF
Drop a file on the box or pick one. Only the file you choose is read; the page cannot see anything else on your computer.
-
Your browser reads it
A copy of pdf.js, Mozilla's open-source PDF reader, is served from this site and runs in your browser. It pulls out the title, language, tags, text, figures, tables, bookmarks, form fields and security settings.
-
You get plain-language results
Each problem says what it means for someone using a screen reader, and how to fix it in Word, InDesign or Acrobat where we can say so accurately. Nothing is stored. Close the tab and the results are gone.
What it checks
| Check | Where it looks |
|---|---|
| Tag structure | Whether the file has a structure tree with tags in it. |
| Marked as tagged | The "Marked" flag in the file's mark information, reported separately from the tags themselves. |
| Language | The language set for the whole document. |
| Title | The title in the file's properties (or its XMP metadata), and whether the viewer is told to show it. |
| Text on each page | Whether each page has text a reader can extract. Pages with none are listed, and a document where no page has any is flagged as probably scanned. The exception is a page with no text whose figure tags all have alternative text and together cover at least 75% of the page, judged from the figure sizes recorded in the file. That page is listed separately as less serious, because a screen reader reads the descriptions, and it stops the document being flagged as scanned. A file that records no figure sizes cannot qualify. |
| Figures | Whether each Figure tag has alternative text. |
| Headings | Whether heading tags exist, and whether the levels skip. |
| Tables | Whether tables have header cells. This one is advisory. |
| Bookmarks | Whether a document of more than 20 pages has a bookmark outline. |
| Security settings | Whether encryption blocks extracting text for accessibility. |
| Form fields | Whether form fields have a tooltip, where this tool can read it. |
Questions
Is my PDF uploaded anywhere?
No. The file is read into your browser's memory and checked by code running in this page. It is never sent to GotAlt or to any other server, and nothing is kept once you close the tab. You can verify that yourself: open your browser's developer tools, switch to the Network tab and run a check. You will see this site's own files being loaded (the PDF reader, plus its character maps and standard-font files when a document needs them) but no request that carries your PDF or its name. Those map and font files are named after what the document uses, so a Japanese document asks for a Japanese character map: GotAlt can see that kind of detail in its request logs, never the file. Like every page on this site, this one also loads a web font from Google Fonts; that request is about the font and has nothing to do with your file.
How big a file can it check?
Up to 100 MB, because your browser has to hold the whole file in memory. Within that, page-level checks (text, tags, figures, tables and form fields) cover the first 200 pages. Document-level checks (title, language, bookmarks and security settings) always cover the whole file. If a document is longer than 200 pages the results say it was only partly checked.
What does "tagged" mean, and why does it matter most?
A tagged PDF contains a hidden structure that labels each piece of content as a heading, a paragraph, a list item, a table cell, a figure and so on, in reading order. Screen readers rely on it. An untagged PDF is closer to a picture of a page that happens to contain text: the reader has to guess the order. Getting tags right usually means exporting again from Word or InDesign with tagging switched on, which is far less work than repairing the tags afterwards.
Why does it say my document is scanned when it isn't?
The tool decides by whether it can extract any text. Text that has been converted to outlines (some design and print exports do this) looks like a picture to it. To test a flagged file yourself, open it in a PDF viewer and try to select a word with the mouse. If you cannot, a screen reader will not be able to read it either. The check can also be fooled the other way: text in a font with no Unicode map counts as text, even though what comes out may be nonsense. Copying a sentence and pasting it into a text editor shows which case you have.
Does a clean result mean my PDF is accessible?
No. It means the checks this tool can run found nothing wrong. Alt text can be present and useless, the reading order can be wrong and colour contrast is not checked at all. Treat a clean result as a good start, then run PAC (Windows only) and veraPDF and listen to the file with a screen reader.
Which tool should I use for a full PDF check?
PAC is a free checker built around PDF/UA and WCAG that runs on Windows only, and veraPDF is an open-source validator that covers PDF/UA. Adobe Acrobat Pro also has a built-in accessibility check, but it is a paid product. We built this tool for the quick, private first look, and we would rather you use those for the full check than pretend this replaces them.
Does this matter for ADA Title II?
The Department of Justice's rule for state and local governments treats a PDF on a public entity's website as web content. Its exception for older documents covers only documents that were posted before the compliance date, and not ones people still use to apply for, get access to or take part in a service. Read the DOJ's own fact sheet, and our guide to the Title II deadline, which tracks the dates. This is not legal advice.
Why does it only ask for bookmarks above 20 pages?
That number is our own rule of thumb, not something WCAG sets. A short document is easy to move around in; in a long one, bookmarks work as a table of contents that keyboard and screen reader users can jump through. If a shorter document is hard to find your way around, add bookmarks anyway. The W3C technique for PDF bookmarks explains how they help.
Related free tools
-
Image contrast checker
This tool does not read colours. Take a screenshot of a PDF page and check text over images or coloured backgrounds with this one.
-
Contrast checker
Check a foreground and background colour pair against the WCAG contrast ratios.
-
Accessibility bookmarklet
Most PDFs are reached from a web page. Run checks on that page, live in your browser.
The page that links to your PDF matters too
People find PDFs through web pages. The free scan checks your site's pages for the problems a machine can find, and tells you plainly what it can't.
Scan my site for free