Техническое РУКОВОДСТВО

llms.txt

llms.txt is an informal, community proposal for a Markdown index that summarizes a site and points AI agents to selected resources.

  • 3 минуты чтения
  • Последнее обновление
На этой странице3 минуты чтения
  1. Обзор
  2. Глубокое погружение
  3. Стратегическое воздействие
  4. The Future of llms.txt
  5. Реальная реализация
  6. Риски и ограничения
  7. Дорожная карта реализации
  8. Продолжайте исследовать
  9. Часто задаваемые вопросы

Обзор

Publishing the file may help a tool that chooses to fetch it, but it is not a crawler directive, a search-ranking signal, or a guarantee that any particular assistant will use the content.

Глубокое погружение

The llms.txt idea was proposed by Jeremy Howard in 2024 as a way to address a real problem: large language models have limited context windows and web pages are often cluttered with navigation, ads, and JavaScript that make it hard for an AI tool to extract the actual content efficiently. The convention specifies a Markdown file at the site root (/llms.txt) with a required H1 title, an optional blockquote summary, and then sections of links, each ideally pointing to clean Markdown versions of pages rather than full HTML. Some documentation platforms and site owners also publish an llms-full.txt file that concatenates complete pages. This is a community extension, not part of the current llmstxt.org v2 specification. It is a proposal rather than a protocol requirement. Some documentation platforms generate these files and several AI labs publish their own documentation indexes, but publication does not mean every search crawler or AI product fetches, follows, or ranks by them. Treat the file as an optional author-curated index and verify a target tool’s behavior before relying on it. A key misconception is treating llms.txt as equivalent to robots.txt in enforcement — robots.txt is a widely respected directive about crawler access, while llms.txt is only a curated pointer file that a given AI tool may or may not choose to fetch and use. A site owner can experiment after checking for stale or sensitive links, and can measure whether known tools fetch the file. Do not promise an SEO or citation benefit, and do not use it to replace content quality, accessible navigation, or crawler controls.

Стратегическое воздействие

Стоимость и бюджет

Архитектурные решения влияют на производительность и эксплуатационные расходы на протяжении многих лет.

Более четкие решения

Техническое образование помогает командам выбрать правильный стек, а не только самый новый.

Контроль качества

Лучший инженерный выбор снижает вероятность возникновения проблем с надежностью на производстве.

The Future of llms.txt

The proposal remains open to community input, and its published format can change. Some documentation platforms generate files, but each crawler or assistant decides what to retrieve and how to use it. A site team should date its file, review links as documentation changes, and treat observed fetches as evidence of access only—not proof that a model used the content or that search visibility improved. Avoid making ranking promises; use official crawler controls for access rules and ordinary content quality for discoverability.

Реальная реализация

A documentation site publishes an llms.txt at its root listing its getting-started guide, API reference, and changelog as Markdown links with one-line descriptions.

An open-source project's llms.txt links directly to raw Markdown versions of its docs pages so an AI assistant can read clean text instead of parsing rendered HTML.

A company site includes an llms-full.txt variant that concatenates entire documentation contents into one file for tools that want to ingest everything at once.

A blog experiments with llms.txt by listing only its cornerstone explainer articles, hoping AI answer engines cite those pages more accurately.

Риски и ограничения

  • Оптимизация одного теста может скрыть более широкие недостатки системы.

  • Затраты на инфраструктуру и техническое обслуживание часто недооцениваются.

  • Пробелы в безопасности и наблюдаемости могут увеличиваться по мере усложнения систем.

Дорожная карта реализации

  1. Определите целевые показатели задержки, качества и стоимости перед внедрением.

  2. Тестирование при реалистичной нагрузке и условиях данных.

  3. Мониторинг прибора на наличие ошибок, дрейфа и влияния пользователя.

  4. Перед масштабированием подготовьте пути отката и реагирования на инциденты.

Продолжайте исследовать

Free newsletter

Get the daily AI briefing

Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.

One email each weekday. Unsubscribe in one click. We never sell or share your address.

Test yourself

Take the llms.txt quiz

Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.

Начать тест

Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation

Часто задаваемые вопросы

What is llms.txt?

llms.txt is an informal, community proposal for a Markdown index that summarizes a site and points AI agents to selected resources. Publishing the file may help a tool that chooses to fetch it, but it is not a crawler directive, a search-ranking signal, or a guarantee that any particular assistant will use the content.

According to the guide, what is the intended role of llms.txt in the proposal?

The proposal describes a curated Markdown overview and resource links; it does not force any agent to retrieve them.

How does llms.txt differ from robots.txt in terms of enforcement?

The guide explicitly warns against treating llms.txt as equivalent in enforcement to robots.txt, noting it is voluntary and adoption is uneven.

What does the companion file llms-full.txt do differently from llms.txt?

Some sites publish a full-content file as an informal extension; the current llmstxt.org v2 proposal defines the link-index format, not a required llms-full.txt companion.

Who proposed the llms.txt convention, and around when?

The guide attributes the proposal to Jeremy Howard in 2024, addressing limited context windows and cluttered HTML.

What content format does llms.txt recommend linking to, rather than full HTML pages?

The proposal recommends concise Markdown links and clean Markdown versions where available; complete site contents are not required.