The llms.txt standard is one of the fastest-moving ideas in technical SEO. It is a simple text file with an outsized impact on how AI engines understand your site.

In this guide you will learn how to copy a working llms.txt and adapt it. We will keep it practical, with clear steps, visual breakdowns, and specific actions you can take today. The first step in any AI visibility project is to run a free AI crawler check on your website so you know exactly where you stand against the 196 bots we track across 8 categories.

Key Takeaways

  • llms.txt Examples and Copy-Paste Templates is a practical, repeatable process, not a one-time fix.
  • Most AI visibility problems trace back to access, not content.
  • You can verify every change with the free AI crawler check and the robots txt check.
  • Document your approach so the whole team applies it consistently.
Structure of an llms.txt file Annotated llms.txt example. An H1 heading holds the site name, a blockquote gives a one-line summary, H2 sections group content areas, and markdown links with descriptions point AI engines to your most important pages. /llms.txt # Acme Analytics > B2B analytics platform for retail teams ## Docs - [Quick Start](/docs/start): setup in 5 min - [API Reference](/docs/api): all endpoints ## Optional H1 = your site name (required) Blockquote = one-line summary AI engines read this first H2 sections group your content Links + short descriptions Point AI to your best pages Optional = safe to skip for context-limited models
Structure of an llms.txt file: site name, summary, and curated markdown links for AI engines.

Adding llms.txt Without Neglecting the File That Matters

llms.txt is a proposal for curating content, not a replacement for crawler access control, so the sequence keeps the two in the right order of importance.

5 Steps, in Order
1

Fix crawler access first

llms.txt has no effect if crawlers cannot reach your content. Run an AI crawler test and resolve any blocks before spending time on curation, because the curation file is read by a crawler that has to get in.

2

Confirm robots.txt is doing its job

Validate the live file with a robots.txt check and, where changes are needed, generate them with a robots.txt generator. This is the file with broad support today, and it stays the priority.

3

Select the pages worth pointing at

llms.txt is valuable in proportion to its curation. Choose the documentation, reference and canonical explanatory pages you would want quoted, and leave out everything you would not.

4

Write descriptions that state what each page answers

A link with a vague label adds nothing. A one-line description naming the specific question a page resolves is the entire value of the format.

5

Keep it current or remove it

A stale llms.txt pointing at moved or deleted pages is worse than none, because it advertises neglect. Set a review interval or do not publish the file.

The Part of llms.txt Structure That Trips People Up

Most llms.txt templates you can copy are a list of every URL on the site with the page title next to it. That is a sitemap with worse formatting, and it wastes the one thing the format is actually good for. If a model reads only one file to work out what your site is and what it is authoritative about, a flat inventory tells it almost nothing that the sitemap did not.

The format is a Markdown document, which means it can carry something a sitemap structurally cannot: judgement. Grouping, ordering, and one line of explanation per entry are all available to you, and they are the entire value. A short file that says what the site is for, names the handful of pages that matter most, and explains in one clause why each of those pages exists is more useful than a complete listing, and it is also much cheaper to keep accurate.

The llms.txt Details That Decide the Outcome

Open with what the site is, in plain language, before any list

The most valuable lines in the file are the first few, because they are the only place you get to state your own scope. Say what the site does, who operates it, and what subject it is credible about. Avoid marketing register: the audience for this paragraph is a system building a short description of you, and adjectives compress badly while specifics survive. If a reader could not tell from your opening lines what questions you should be trusted on, the rest of the file will not fix it.

Group by purpose, because the grouping is the signal

Headings cost nothing and carry real information. Tools, reference material, documentation and editorial content are different kinds of page and deserve different sections, because that structure tells a reader which parts of the site are canonical facts and which are opinion. A single undifferentiated list throws that distinction away. This is also the part of the file that stays useful longest, since your top-level structure changes far less often than any individual URL.

Annotate every entry with the reason it is listed, not a restatement of its title

The one-line description after each link is where the format earns its place. "Robots.txt Validator" adds nothing that the URL did not. A line saying what the page lets someone do, and what makes it different from the neighbouring entry, is genuinely new information. Write the descriptions so that someone who read only your llms.txt could correctly answer a question about which of your pages to use, because that is the realistic best case for the file.

Keep it short enough that it stays true

A file listing every URL is stale the week after you publish it, and a stale file is worse than a brief one, because it confidently points at pages that have moved. Prefer a short curated file you will actually maintain. If you want completeness, that is what sitemap.xml is for, and it can be generated. llms.txt should be the file you would be willing to have quoted back to you unaltered in a year, which is a strong argument for keeping it small.

The four pillars of AI visibility Four pillars supporting AI visibility. Pillar 1 Access: AI crawlers can reach your pages. Pillar 2 Infrastructure: llms.txt, sitemap and HTTPS in place. Pillar 3 Structure: clear headings, FAQs and schema markup. Pillar 4 Authority: expertise signals and citations from trusted sources. AI VISIBILITY: read, trusted, and cited by AI engines 1 ACCESS Crawlers can reach your pages: robots.txt, WAF, no JS walls 2 INFRASTRUCTURE llms.txt, XML sitemap, HTTPS, clean canonical URLs 3 STRUCTURE Clear H2/H3 headings, FAQs, schema markup, quotable paragraphs 4 AUTHORITY E-E-A-T signals, author pages, mentions on trusted sources Work the pillars in order: authority means nothing if crawlers cannot access your pages in the first place.
The four pillars of AI visibility: access, infrastructure, structure, and authority.

The One llms.txt Error Worth Auditing For

Pasting your sitemap into a Markdown file and calling it llms.txt

The reason this is the dominant mistake is that it is easy and it looks like completion: the file exists, it validates, a checker gives you the points. What it does not do is answer the only question the format was invented to answer, which is what a reader should understand about your site before deciding which page to trust. It also creates a maintenance liability out of a file that had no obligation to be exhaustive. The correction is smaller than people expect: cut the list to the pages you would name if someone asked in person, then spend the saved effort on the one-line descriptions, which is where the whole value sits.

How to Confirm llms.txt Structure Behaves the Way You Think

The check that matters here: Hand your llms.txt to somebody who does not know the site and ask them what you are authoritative about and which page they should use for a specific task. If they cannot answer from the file alone, the descriptions are doing no work.

Where to Go From Here

llms.txt is one piece of a larger picture. The AI bot directory documents every crawler we track with its operator, purpose and safety rating, and the bulk URL checker audits many sites in one pass if you manage a portfolio.

An llms.txt file only helps if crawlers can actually reach it. Confirm access with an AI crawler test once the file is live.

Your llms.txt Action Checklist

Five concrete steps, specific to what this guide covered. Work through them in order, changing one thing at a time so you can tell which change produced the result.

  • Establish a baseline with an AI bot access checker and write down the score before you change anything.
  • Apply the single highest-impact change from this guide, on its own, so you can attribute the result.
  • Validate the change with the robots checker before it reaches production.
  • Re-measure and compare against your baseline rather than against expectation.
  • Schedule a recurring re-check, because redesigns and security updates quietly undo this work.