ShowUp AIShowUp AI
All articles
Technical Guide7 min read

The honest guide to llms.txt

A clear-eyed look at llms.txt: what it is, what's confirmed (and not) about AI engine support, and how to set one up properly.

By ShowUp AI · Published

llms.txt is a proposed plain-text file, placed at the root of your domain, that gives AI systems a short, structured summary of your site and points them to your most important pages. It's worth adding because it's cheap and harmless, but be clear-eyed: as of now no major AI crawler has confirmed it reads or prioritises llms.txt, so treat it as a small hygiene item, not a fix for weak AI visibility.

Here's what the file actually is, what adoption really looks like today, what to put in it, and the mistakes that make people think it did something when it didn't.

What exactly is llms.txt supposed to do?

Answer first: llms.txt is a markdown file at yoursite.com/llms.txt that lists, in a short human-and-machine-readable format, what your site is, what it does, and links to the pages you consider most important — the idea being that an AI system fetching your site could read this one file instead of crawling everything.

It was proposed in 2024 as an equivalent to robots.txt or sitemap.xml, but for language models instead of search crawlers. robots.txt tells crawlers what they can't access; sitemap.xml lists every URL for indexing; llms.txt is meant to be curated and opinionated — a short brief, not a full index.

The typical structure looks like this (shown as plain text, not code, since the real file has no special syntax beyond markdown headings and links):

Title: your company name A one-line summary of what you do.

Section: Docs Link to your documentation homepage - short description Link to your API reference - short description

Section: About Link to your about/company page - short description

It's meant to be short — a few dozen lines, not a full sitemap dump.

Is llms.txt actually supported by ChatGPT, Perplexity, Gemini and Grok?

Answer first: no, not confirmed. There's no public statement from OpenAI, Anthropic, Google or Perplexity committing to fetch or prioritise llms.txt, and independent testing since it launched has found inconsistent or no evidence that it changes what gets retrieved or cited. Adoption on the publishing side has grown, but adoption on the reading side is the part that matters, and that's the unresolved half.

Being straight about where things stand:

  • Plenty of sites publish one now. Frameworks, developer tools and SaaS companies have added llms.txt files, partly as a hedge and partly because tooling makes it nearly free to generate.
  • No major AI lab has confirmed it's used in retrieval. These systems have not documented llms.txt as an input to how they crawl, index, or decide what to cite.
  • It doesn't replace normal crawling or indexing. Assistants that browse the web still rely primarily on standard web content, search indexes, and whatever their retrieval pipeline already does — llms.txt is not known to override or supplement that in a documented way.
  • It costs almost nothing to add, so the calculus is "small potential upside, near-zero downside" rather than "essential and skipping it hurts you."

So the honest framing for a client or reader is: add it because it's low effort and future-proofs you a little if support does arrive, not because it will visibly move your AI mentions this quarter. If you're choosing where to spend limited time on AI visibility, structured content, schema markup and third-party listings currently matter more — see the AI visibility checklist for the fuller priority order.

What should actually go in your llms.txt file?

Answer first: a one-paragraph description of what your company does, then a small number of curated sections linking to your highest-value pages — docs, pricing, key product pages, and an about/company page — each with a one-line description, kept short enough that the whole file could be read in under a minute.

A sensible structure:

  1. 1Title and one-line summary. What you do, for whom, stated plainly. No slogans.
  2. 2A short paragraph of context. Two or three sentences: category, main product, who it's for.
  3. 3A "Docs" or "Product" section linking to your most important functional pages — documentation, key features, pricing.
  4. 4An "About" section linking to your company/about page, so an entity-resolution step has somewhere authoritative to check facts like founding date or location.
  5. 5Optionally, a "Blog" or "Resources" section with two or three of your most canonical, evergreen articles — not a dump of every post.

What to leave out:

  • Marketing copy. This isn't a landing page; keep sentences factual and short.
  • Every page on your site. That's what sitemap.xml is for. llms.txt should be a curated shortlist, ideally under 50-100 lines.
  • Duplicate content from the file itself. Link to pages rather than pasting their content into the file — an llms-full.txt variant exists for that in some implementations, but it's a different, heavier file and not what most sites need.
  • Anything that changes weekly. Pricing numbers, promotions, and dates go stale in a file nobody automatically re-crawls on a schedule.

Where should the file live and how do you keep it accurate?

Answer first: llms.txt must sit at the root of your domain — yoursite.com/llms.txt — exactly like robots.txt, because that's the only location any current reader would know to check; subdirectories or CMS-generated paths won't be found.

Practical setup notes:

  • Static hosting: drop the file straight into your site's root/public folder alongside robots.txt and sitemap.xml.
  • CMS-based sites (WordPress, Webflow, etc.): check whether your platform allows arbitrary root-level files; some require a plugin or a redirect rule to serve a file that isn't a normal page.
  • Subdomains: if your docs live on a separate subdomain, that subdomain needs its own llms.txt if you want it covered — root-domain placement doesn't cascade down automatically.
  • Keep it under version control if you can, same as robots.txt, so changes are tracked and reviewable rather than edited ad hoc by whoever remembers it exists.
  • Revisit it quarterly. Update links when you launch a major product page, retire an old one, or rename navigation — a stale llms.txt pointing at dead pages is worse than a slightly outdated sitemap because there's no automated crawl to catch the broken links for you.

What are the common mistakes people make with llms.txt?

Answer first: the biggest mistake is treating llms.txt as a silver bullet and skipping the things that actually influence AI mentions today — clean structured content, schema markup, and third-party citations — while a close second is publishing a bloated or stale file that nobody maintains.

Specific mistakes to avoid:

  • Overclaiming impact. Don't tell stakeholders "we added llms.txt so we'll now show up in ChatGPT" — there's no evidence for that causal link yet.
  • Copy-pasting a template without editing it. Generic auto-generated llms.txt files (a common output of some website builders) that still contain placeholder text are worse than not having one — they signal neglect if anyone or anything does read it.
  • Listing too many pages. Defeats the purpose of being a curated summary; if everything is "important," nothing is.
  • Forgetting to update it. A file frozen at launch, still linking to a product you've since renamed, becomes actively misleading.
  • Putting it behind auth or a CDN rule that blocks bots. If your CDN or WAF blocks unfamiliar user agents by default, check that llms.txt (and robots.txt) are explicitly allowed through.
  • Skipping the fundamentals. Structured Organization and Product schema, FAQPage markup on genuinely helpful FAQ content, and consistent brand facts across third-party sites currently have a stronger evidence base for influencing AI answers than llms.txt does.

Quick reference: llms.txt vs the files it's often confused with

FilePurposeConfirmed AI-engine support
robots.txtTells crawlers what they can/can't fetchYes — long-established, widely respected
sitemap.xmlFull list of indexable URLsYes — standard for search indexing
llms.txtCurated summary + key links for AI systemsNo — proposed, not confirmed by major labs
Schema markup (JSON-LD)Structured facts about entities on a pagePartial — used by Google AI Overviews and increasingly cited as a strong signal elsewhere

Add llms.txt because it takes twenty minutes and can't hurt. Just don't let it eat the budget or attention that should go toward schema markup, third-party listings and content structure — the levers with an actual track record right now. If you want to see where your brand currently stands across ChatGPT, Perplexity and Gemini before deciding what to prioritise, run the free AI visibility check first.

Get named in AI answers — starting today

The $299 ShowUp AI Setup checks your live site and builds your report, technical kit (llms.txt, schema, meta, FAQs) and ready-to-publish content — with a 30-day re-check included. Monthly Watch plans are optional afterwards.

Written by the ShowUp AI team — we help brands get found by ChatGPT, Perplexity, Gemini and Google AI.

Related articles