GUIDA TECNICA

llms.txt

llms.txt is an informal, community proposal for a Markdown index that summarizes a site and points AI agents to selected resources.

  • 3 minuti di lettura
  • Ultimo aggiornamento
In questa pagina3 minuti di lettura
  1. Panoramica
  2. Immersione profonda
  3. Impatto strategico
  4. The Future of llms.txt
  5. Implementazione nel mondo reale
  6. Rischi e guardrail
  7. Tabella di marcia per l'implementazione
  8. Continua a esplorare
  9. Domande frequenti

Panoramica

Publishing the file may help a tool that chooses to fetch it, but it is not a crawler directive, a search-ranking signal, or a guarantee that any particular assistant will use the content.

Immersione profonda

The llms.txt idea was proposed by Jeremy Howard in 2024 as a way to address a real problem: large language models have limited context windows and web pages are often cluttered with navigation, ads, and JavaScript that make it hard for an AI tool to extract the actual content efficiently. The convention specifies a Markdown file at the site root (/llms.txt) with a required H1 title, an optional blockquote summary, and then sections of links, each ideally pointing to clean Markdown versions of pages rather than full HTML. Some documentation platforms and site owners also publish an llms-full.txt file that concatenates complete pages. This is a community extension, not part of the current llmstxt.org v2 specification. It is a proposal rather than a protocol requirement. Some documentation platforms generate these files and several AI labs publish their own documentation indexes, but publication does not mean every search crawler or AI product fetches, follows, or ranks by them. Treat the file as an optional author-curated index and verify a target tool’s behavior before relying on it. A key misconception is treating llms.txt as equivalent to robots.txt in enforcement — robots.txt is a widely respected directive about crawler access, while llms.txt is only a curated pointer file that a given AI tool may or may not choose to fetch and use. A site owner can experiment after checking for stale or sensitive links, and can measure whether known tools fetch the file. Do not promise an SEO or citation benefit, and do not use it to replace content quality, accessible navigation, or crawler controls.

Impatto strategico

Costo e budget

Le decisioni relative all'architettura determinano prestazioni e costi operativi per anni.

Decisioni più chiare

La formazione tecnica aiuta i team a scegliere lo stack giusto, non solo quello più nuovo.

Controllo di qualità

Migliori scelte ingegneristiche riducono gli incidenti legati all’affidabilità nella produzione.

The Future of llms.txt

The proposal remains open to community input, and its published format can change. Some documentation platforms generate files, but each crawler or assistant decides what to retrieve and how to use it. A site team should date its file, review links as documentation changes, and treat observed fetches as evidence of access only—not proof that a model used the content or that search visibility improved. Avoid making ranking promises; use official crawler controls for access rules and ordinary content quality for discoverability.

Implementazione nel mondo reale

A documentation site publishes an llms.txt at its root listing its getting-started guide, API reference, and changelog as Markdown links with one-line descriptions.

An open-source project's llms.txt links directly to raw Markdown versions of its docs pages so an AI assistant can read clean text instead of parsing rendered HTML.

A company site includes an llms-full.txt variant that concatenates entire documentation contents into one file for tools that want to ingest everything at once.

A blog experiments with llms.txt by listing only its cornerstone explainer articles, hoping AI answer engines cite those pages more accurately.

Rischi e guardrail

  • L'ottimizzazione di un benchmark può nascondere debolezze di sistema più ampie.

  • I costi delle infrastrutture e della manutenzione sono spesso sottostimati.

  • Le lacune in termini di sicurezza e osservabilità possono aumentare man mano che i sistemi diventano più complessi.

Tabella di marcia per l'implementazione

  1. Definire obiettivi di latenza, qualità e costi prima dell'implementazione.

  2. Benchmark in condizioni di carico e dati realistiche.

  3. Monitoraggio dello strumento per errori, deriva e impatto sull'utente.

  4. Preparare percorsi di rollback e risposta agli incidenti prima della scalabilità.

Continua a esplorare

Free newsletter

Get the daily AI briefing

Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.

One email each weekday. Unsubscribe in one click. We never sell or share your address.

Test yourself

Take the llms.txt quiz

Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.

Inizia il quiz

Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation

Domande frequenti

What is llms.txt?

llms.txt is an informal, community proposal for a Markdown index that summarizes a site and points AI agents to selected resources. Publishing the file may help a tool that chooses to fetch it, but it is not a crawler directive, a search-ranking signal, or a guarantee that any particular assistant will use the content.

According to the guide, what is the intended role of llms.txt in the proposal?

The proposal describes a curated Markdown overview and resource links; it does not force any agent to retrieve them.

How does llms.txt differ from robots.txt in terms of enforcement?

The guide explicitly warns against treating llms.txt as equivalent in enforcement to robots.txt, noting it is voluntary and adoption is uneven.

What does the companion file llms-full.txt do differently from llms.txt?

Some sites publish a full-content file as an informal extension; the current llmstxt.org v2 proposal defines the link-index format, not a required llms-full.txt companion.

Who proposed the llms.txt convention, and around when?

The guide attributes the proposal to Jeremy Howard in 2024, addressing limited context windows and cluttered HTML.

What content format does llms.txt recommend linking to, rather than full HTML pages?

The proposal recommends concise Markdown links and clean Markdown versions where available; complete site contents are not required.