STANDARD · §2.2 · DISCOVER · READ · LAST REVIEWED SEPTEMBER 2026

What is llms.txt?

A curated markdown index of your site, written for language models.


Agent Readiness Compare editors · Spec status checked September 2026

Answer

llms.txt is a markdown file, usually at /llms.txt, that gives language models a short description of a site and links to its most useful pages. Only an H1 with the site or project name is required; a summary blockquote and H2 sections of links are recommended. It helps agents read; it does not control crawling or expose actions.

Maintainer:
proposed by Jeremy Howard
Status:
community proposal

1.What is llms.txt?

llms.txt is a proposal, first published by Jeremy Howard in 2024, for a markdown file that helps language models use a website at inference time. Instead of making a model parse full HTML pages with navigation, scripts and layout, the site publishes a short, curated file that says what the site is and where the important content lives.

2.What goes in the file?

The proposal defines a fixed order:

  1. An H1 with the name of the project or site. This is the only required section.
  2. A blockquote with a short summary.
  3. Optional paragraphs or lists with more detail.
  4. Zero or more H2 sections, each containing a markdown list of links in the form [name](url), optionally followed by a colon and notes.
  5. An H2 section named 'Optional' for links an agent can skip when context is short.
/llms.txt (example)
# Example Analytics

> Example Analytics is a product analytics tool for SaaS teams. Plans start with a free tier; paid plans are billed per tracked event.

## Product
- [Pricing](https://example.com/pricing.md): plans, limits and billing
- [Features](https://example.com/features.md): what each plan includes

## Docs
- [Quickstart](https://example.com/docs/quickstart.md): install and first event
- [API reference](https://example.com/docs/api.md): endpoints and authentication

## Optional
- [Changelog](https://example.com/changelog.md)

3.Where does the file go?

At the site root (/llms.txt) to cover the whole site. The current proposal also allows a file at a sub-path, for example /docs/llms.txt, covering only the pages under that path. The proposal also asks sites to offer clean markdown versions of pages at the same URL with .md appended (page.html.md) or with the extension replaced by .md.

Note

Some sites also publish an llms-full.txt with the full text of their docs in one file. That is a common convention, not part of the llms.txt proposal we reviewed.

4.What does llms.txt not do?

  • It does not control access. Crawl permissions live in robots.txt.
  • It does not expose actions. For tasks, see MCP and WebMCP.
  • It is not guaranteed to be read. Whether a given agent or model fetches llms.txt is up to that agent.

5.How do you check it?

  1. Fetch https://yourdomain/llms.txt and confirm HTTP 200 and a text or markdown content type.
  2. Confirm the first line is an H1 and that the summary is accurate and current, including pricing wording.
  3. Open every link. Linked pages should return 200 and, ideally, markdown.
  4. Compare it with a readiness scanner result; ora.ai Scan and Cloudflare Is It Agent Ready both look for discovery files.

6.Where does it sit in the readiness stack?

Layers 1 and 2: it helps an agent find the right pages and read them cheaply.

Figure 1. The readiness stack: five layers, the standards that address each, and the tools that document support.
Text version of the diagram
Text version of the readiness stack
LayerQuestionStandardsTools that document support
L1 DISCOVERCan an agent find you, and is it allowed in?robots.txt (RFC 9309); Sitemaps; llms.txt; A2A Agent CardCloudflare AI Crawl Control (controls crawler access); ora.ai Scan (checks it); Cloudflare Is It Agent Ready (checks it)
L2 READCan it read and understand what you offer?llms.txt; Markdown negotiation (Accept: text/markdown); JSON-LDCloudflare Markdown for Agents (serves markdown); ora.ai Scan (checks it)
L3 ACTCan it complete a task on your site or API?MCP; WebMCP; agents.json; OpenAPICloudflare (hosts MCP servers); Vercel (hosts MCP servers); nekuda (builds WebMCP tools); ora.ai Journey and WebMCP audit (test it)
L4 PAYCan it pay you?x402; ACP; UCP; MPPCloudflare Pay per crawl (private beta, for crawlers); ora.ai Scan payments layer (checks it)
L5 TRUSTCan you tell which agent it is, and on whose behalf it acts?Web Bot Auth; OAuth 2.0Cloudflare (verifies signed bots); Vercel (verifies signed bots); Forter (links agentic shoppers to verified customer identities)

Frequently asked questions

Is llms.txt the same as robots.txt?

No. robots.txt says what crawlers may fetch. llms.txt describes the site and points to the best content. You can publish both.

Does llms.txt need to be in markdown?

Yes. The proposal specifies markdown with a required H1, then an optional blockquote summary and H2 sections of links.

How is llms.txt different from agents.json?

llms.txt describes content. agents.json describes API flows an agent can execute. See llms.txt and agents.json.

Update, August 2026: the proposal was revised to v2. See llms.txt v2: what changed.

Source: llmstxt.org · Reviewed Sep 2026