Skip to content
LLM Discovery Infrastructure

Publish llms.txt as a Navigation Aid, Not a Ranking Spell

Implement the llms.txt proposal with accurate scope, stable markdown, canonical sources, change ownership, access controls, and evidence that separates fetches from visibility.

Technical archivist preparing a concise field guide beside a large website blueprint

Field note

By XenGrowth EditorialPublished Reviewed 11 min read

Key takeaways

  • Treat llms.txt as a voluntary proposal for helping tools navigate useful material—not a universal search or citation standard.
  • Point to authoritative, public, canonical sources and keep markdown versions faithful to their HTML counterparts.
  • Do not expose private, draft, disallowed, or contradictory material through the file.
  • Measure fetches and downstream use separately; the file does not prove ranking, citation, or model training.

01

A map can help an agent without changing the terrain

A documentation site contains hundreds of pages wrapped in navigation and scripts. A concise file can point an agent toward stable overviews, concepts, references, and policies. It does not repair weak evidence, grant crawl permission, override a canonical, or guarantee use by a search or answer system.

The llms.txt site describes the format as a proposal to help agents use a website. Publish it when there is a real collection worth navigating and somebody will own the file as the site changes.

Swipe to compare every column

QuestionDefensible answerUnsupported leap
What is it?A proposed markdown navigation fileA universal web standard
What can it improve?Discovery by supporting clientsGuaranteed indexing or citation
What should it list?Stable authoritative public sourcesEvery URL on the domain
What proves value?Observed fetch and attributable useFile existence

02

Curate the file around decisions and authority

Organize links by the jobs an agent needs to complete. Prefer durable overview, concept, reference, pricing, policy, and support pages. Exclude thin variants, tracking URLs, expired campaigns, and pages whose authority is unclear.

If you offer markdown alternates, keep facts, examples, source notes, dates, and canonical identity aligned with the human page. A clean alternate that drops qualifications is easier to process and easier to misuse.

03

Keep access and indexing controls in their own systems

Robots.txt is an access-request protocol, and RFC 9309 says it is not authorization. Authentication protects private material. Meta or header directives address index preferences. llms.txt replaces none of them.

Do not create an attractive directory of sensitive URLs. When a listed page is removed, moved, restricted, or superseded, update the map and underlying URL behavior together.

04

Build evidence without inventing attribution

Log requests for the file and markdown alternates with verified user-agent and network context where practical. Track successful responses, freshness, broken links, and downstream referrals that actually expose a source.

Do not label every AI referral or citation an llms.txt win. Systems also discover pages through ordinary crawling, indexes, links, and feeds. Report controlled tests and their limits.

Primary sources and further reading

Use the source material to validate details against your own context and current platform configuration.

This field note follows the XenGrowth editorial policy: primary sources where available, visible limitations, material review dates, and no invented first-hand experience.

Stay with the problem

Explore AI search & GEO