Skip to main content
    Back to Resources

    Jan 3 2026

    The two obligations on your PDFs

    Inclusion is no longer the only thing the PDFs on your site have to do.

    You answer for what is on the site. The obligation to make the PDFs on it work for human visitors was familiar before AI systems started reading the same pages. It is no longer the only obligation those documents carry.

    The first obligation is already familiar

    Accessibility law treats the PDFs on a public-facing website as in scope. In the United States, the Department of Justice Title II final rule takes effect in April 2026 for public entities serving populations of 50,000 or more, and in April 2027 for smaller entities. In the United Kingdom, the Public Sector Bodies Accessibility Regulations 2018 apply regardless of size. In the European Union, the European Accessibility Act applies from 28 June 2025.

    Section508.gov and GOV.UK Government Digital Service both direct publishing toward HTML. PDFs are treated as the format to avoid or replace, not as the default to protect. The direction of travel is not ambiguous, and the enforcement calendar is live.

    The second obligation is newer, and less well understood

    AI readiness is the half most operators have not yet brought to the leadership table. It describes whether the organisation's external digital presence can be read, interpreted, trusted, and acted on by AI systems, and whether the organisation still controls what that presence says about it.

    The evidence that this is live is not speculative. Google has expanded AI Overviews to more than 200 countries and 40 languages, with usage growth above 10% in major markets. Adobe reports AI-sourced traffic to US retail sites rose 393% year on year in the first quarter of 2026, with stronger engagement than non-AI traffic. Deloitte reports 9 in 10 retail executives expect AI to be used more than search engines by 2026.

    OpenAI has expanded shopping and product discovery inside ChatGPT, with structured merchant feeds governing discoverability, relevance, and trust. The US National Institute of Standards and Technology Generative AI Profile identifies provenance, misinformation, and tampered content as risks that require active identification and mitigation. Taken together, this is not an emerging layer. It is an operating one.

    The known online, and the unknown online

    The distinction that governs the PDF estate is the one between the known online and the unknown online. The known is what the organisation actively manages. The unknown is what is still visible externally but no longer actively governed, and it almost always includes PDFs published years ago and left in place on pages no one has reviewed recently.

    AAAnow / Sitemorse internal analysis of 100+ million websites indicates 41% of websites are typically unknown to the organisation's own digital teams. The same internal analysis indicates 19% of PDFs on organisational websites are duplicates. Both figures are internal data, labelled as such, and they are directional in what they measure. The point they support is that most organisations know less about their own estate than they assume.

    Where the two halves meet

    The two halves of the obligation meet on the same documents. A PDF that a screen reader cannot parse is structurally opaque to a retrieval system as well. Duplicates and stale content that confuse a visitor are also provenance weaknesses for language models. A document left in place for years has aged out of its own guidance, and has continued to feed the picture external systems form about the organisation that published it.

    This is why the PDF estate now sits inside two obligations at once. Inclusion is about the people in front of the screen. AI readiness is about the systems reading the same content behind it. Neither is optional, and neither can be resolved by work that addresses only the other.

    What to do about it

    Recognising the obligation is only the first move. The operational stages that follow, discovery, triage, prioritisation, execution, and version management, are covered in The five stages of dealing with a PDF estate. The options the market offers as responses to those stages, and an idea of what they cost to use, are covered in The PDF market, remediation options and idea of costs. What happens to the content once it has left the site, and why that matters for product catalogues, recalls, and distribution through partners, is covered in What happens when a PDF leaves your site.

    Disclaimer:

    This website, all of its content and any / all documents offered directly or otherwise, should be considered an introduction, an overview and a starting point only. It should not be used as a single, sole authoritative guide. You should not consider this as legal guidance. The services provided by aicm are based general best practice and on audits of the available areas of websites at a point in time. Sections of the site that are not open to public access or are not being served (possibly be due to site errors or downtime) may not be covered by our reports. The service and the stars process doesn't carry any official accreditation, be it from any government department, industry regulator and / or internet body. Where matters of legal compliance are concerned you should always take independent advice from appropriately qualified individuals or firms.

    Copyright

    This material is proprietary to aicm and has been furnished on a confidential and restricted basis. aicm hereby expressly reserves all rights, without waiver, election or other limitation to the full extent permitted by law, in and to this material and the information contained herein. Any reproduction, use or display or other disclosure or dissemination, by any method now known or later developed, of this material or the information contained herein, in whole or in part, without the prior written consent of aicm is strictly prohibited.

    For AI agents and LLMs

    We publish /llms.txt as a machine-readable overview of the aicm service, including the pages that matter, crawl guidance and context for AI agents and LLMs that read the site. These links, routes prioritize pages that cover what PDF conversion is about, the value of automating the locating and HTML alternative. Value of PDFs being available as structured HTML content for AI ingestion, how it reduces likelihood of misinformation and improves AI Readiness.

    © 2026 aicm.
    All rights reserved.