SEO

llms.txt

Also called LLMs.txt, llms-full.txt

A proposed markdown file at a site's root pointing language models at its key pages; no major engine confirms reading it.

Quick facts: llms.txt

Category
SEO
Also called
LLMs.txt, llms-full.txt
Level
Intermediate
Affects
AI retrieval experiments, content maintenance load, internal documentation
Where to see it
Your web server root, server logs, any markdown editor
In this article4
  1. How llms.txt works
  2. Why llms.txt matters
  3. Where llms.txt goes wrong
  4. What to do about it

How llms.txt works

llms.txt is a proposed plain markdown file placed at the root of a domain, in the same position as robots.txt. Instead of telling crawlers what they may not touch, it is meant to do the opposite: point a language model at the pages that matter, with a line of context for each, so a model reading your site gets a curated map rather than a raw crawl.

The format is deliberately simple — a heading, an optional summary, then linked lists of the documents you consider canonical. Some sites publish a longer companion file holding the full text of those pages in one place. Nothing about it is enforced. It is a file you choose to write and a machine chooses whether to read.

Why llms.txt matters

The problem behind it is genuine. A model retrieving from a website meets navigation, cookie notices, related-post blocks and template furniture before it reaches the sentence that answers the question, and a short list of canonical documents would cut that noise considerably.

What matters just as much is what has not happened. No major search engine or assistant has confirmed that it reads the file, and Google’s search representatives have said publicly that Google does not use it. Any claim that llms.txt improves visibility in AI search should be treated as unverified. It is a proposal with a following, not a supported standard.

Where llms.txt goes wrong

The expensive failure is duplication. Publish one by hand and you now maintain a second, summarised version of your site, which begins drifting the day a price, a service or a URL changes. If anything ever does read it, it reads your out-of-date copy with exactly the confidence it would give your live pages.

The other failure is displacement. Hours spent writing an llms.txt are hours not spent on the things demonstrably read: the page text itself, structured data, internal links, and a site that states its facts without needing JavaScript. It also restricts nothing — access control belongs in robots.txt and on your server, never here.

What to do about it

My advice is short: publish it only if you can generate it. If your site can produce the file automatically from the same content that builds your pages, the upkeep is close to nothing and the experiment is harmless. If it means somebody editing markdown by hand every month, skip it and spend the time on the pages.

Either way, follow the evidence rather than the enthusiasm. Watch your server logs for requests to the file, so you know whether anything actually fetches it, and judge your standing in AI answers by putting your customers’ questions to the assistants themselves. If a supplier offers llms.txt as a deliverable with a promised result, ask what they will show you to prove it worked.

Do and do not

Do

  • Generate it from your site, never by hand
  • Check server logs to see whether anything fetches it
  • Remove it rather than let it go stale

Do not

  • Expect rankings or citations from publishing it
  • Use it to block or restrict crawlers
  • Let it disagree with your live pages

Questions people ask about this

Should I publish an llms.txt file?

Only if your site can generate it automatically. As a by-product of your existing content it costs nothing and harms nothing. As a hand-maintained document it becomes a second version of your site that quietly falls out of date. There is no confirmed benefit either way, so let maintenance cost decide it rather than hope.

Is llms.txt the same as robots.txt?

No. They share a location and a naming style and nothing else. robots.txt is a long-supported convention telling crawlers which paths they should not request, and the major crawlers honour it. llms.txt is a proposal suggesting which content a language model should prioritise. It grants no permission and blocks nothing.

Will llms.txt get my site cited in AI answers?

There is no evidence that it will. Citations in AI answers come from content the engines have crawled, indexed and judged relevant, so the dependable work is what it has always been: clear pages, accurate facts, structured data and a site that loads. If someone sells the file as a route to citations, ask what proof they will show you.

Related terms

Found this useful?

Share it, or ask an AI to summarise it

Back to the glossary

Knowing the term is the easy part

Applying it to your own site and budget is the work. Book a call and I will tell you what actually applies to you.