Is Your llms.txt AI-Ready?
Validate your llms.txt against the emerging specification: structure, content, links, entity definitions, and use policy.
llms.txt is an emerging web standard that helps AI models understand your website. Similar to how robots.txt tells search crawlers what to index, llms.txt tells language models what your site is about, what entities it represents, and how its content should be used. The specification includes sections for site description, key pages, entity definitions (the people, products, and concepts your site covers), and a use policy defining how AI systems may reference your content. A well-structured llms.txt file improves your chances of being accurately cited by ChatGPT, Perplexity, Claude, and other AI search engines.
What we validate
H1 title, blockquote summary, markdown headings, and clean formatting.
Company info, products/services, descriptive content beyond just links.
URL count, distribution across sections, internal vs external links.
"Is a/provides/offers" language that helps AI build entity associations.
Citation guidance, attribution rules, and contact information.
Word count depth, llms-full.txt bonus, and section variety.
How We Score (And Why It Matters)
The Crawl Checker does a shallow llms.txt check (4 binary signals, 20 pts). This Validator goes deep: section-by-section analysis, link extraction, entity detection, word count analysis, spec compliance, and content quality.
Built on our own implementation
We built a dynamic, entity-aware llms.txt for pixelmojo.io that auto-maps from our knowledge graph. This validator reflects what we learned about what makes an llms.txt effective for AI discovery.