概述
Publishing the file may help a tool that chooses to fetch it, but it is not a crawler directive, a search-ranking signal, or a guarantee that any particular assistant will use the content.
深入探討
The llms.txt idea was proposed by Jeremy Howard in 2024 as a way to address a real problem: large language models have limited context windows and web pages are often cluttered with navigation, ads, and JavaScript that make it hard for an AI tool to extract the actual content efficiently. The convention specifies a Markdown file at the site root (/llms.txt) with a required H1 title, an optional blockquote summary, and then sections of links, each ideally pointing to clean Markdown versions of pages rather than full HTML. Some documentation platforms and site owners also publish an llms-full.txt file that concatenates complete pages. This is a community extension, not part of the current llmstxt.org v2 specification. It is a proposal rather than a protocol requirement. Some documentation platforms generate these files and several AI labs publish their own documentation indexes, but publication does not mean every search crawler or AI product fetches, follows, or ranks by them. Treat the file as an optional author-curated index and verify a target tool’s behavior before relying on it. A key misconception is treating llms.txt as equivalent to robots.txt in enforcement — robots.txt is a widely respected directive about crawler access, while llms.txt is only a curated pointer file that a given AI tool may or may not choose to fetch and use. A site owner can experiment after checking for stale or sensitive links, and can measure whether known tools fetch the file. Do not promise an SEO or citation benefit, and do not use it to replace content quality, accessible navigation, or crawler controls.
戰略影響
成本與預算
多年來,架構決策決定著效能和營運成本。
更明確的決策
技術教育幫助團隊選擇正確的堆疊,而不僅僅是最新的堆疊。
品質管控
更好的工程選擇可以減少生產中的可靠性事故。
The Future of llms.txt
The proposal remains open to community input, and its published format can change. Some documentation platforms generate files, but each crawler or assistant decides what to retrieve and how to use it. A site team should date its file, review links as documentation changes, and treat observed fetches as evidence of access only—not proof that a model used the content or that search visibility improved. Avoid making ranking promises; use official crawler controls for access rules and ordinary content quality for discoverability.
現實世界的實施
A documentation site publishes an llms.txt at its root listing its getting-started guide, API reference, and changelog as Markdown links with one-line descriptions.
An open-source project's llms.txt links directly to raw Markdown versions of its docs pages so an AI assistant can read clean text instead of parsing rendered HTML.
A company site includes an llms-full.txt variant that concatenates entire documentation contents into one file for tools that want to ingest everything at once.
A blog experiments with llms.txt by listing only its cornerstone explainer articles, hoping AI answer engines cite those pages more accurately.
風險與防護欄
優化一項基準測試可以隱藏更廣泛的系統弱點。
基礎設施和維護成本常常被低估。
隨著系統變得更加複雜,安全性和可觀察性差距可能會擴大。
實施路線圖
在實施之前定義延遲、品質和成本目標。
在實際負載和資料條件下進行基準測試。
儀器監控錯誤、漂移和使用者影響。
在擴展之前準備回滾和事件回應路徑。
不斷探索
Free newsletter
Get the daily AI briefing
Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.
One email each weekday. Unsubscribe in one click. We never sell or share your address.
Test yourself
Take the llms.txt quiz
Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.
Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation
常見問題
What is llms.txt?
llms.txt is an informal, community proposal for a Markdown index that summarizes a site and points AI agents to selected resources. Publishing the file may help a tool that chooses to fetch it, but it is not a crawler directive, a search-ranking signal, or a guarantee that any particular assistant will use the content.
According to the guide, what is the intended role of llms.txt in the proposal?
The proposal describes a curated Markdown overview and resource links; it does not force any agent to retrieve them.
How does llms.txt differ from robots.txt in terms of enforcement?
The guide explicitly warns against treating llms.txt as equivalent in enforcement to robots.txt, noting it is voluntary and adoption is uneven.
What does the companion file llms-full.txt do differently from llms.txt?
Some sites publish a full-content file as an informal extension; the current llmstxt.org v2 proposal defines the link-index format, not a required llms-full.txt companion.
Who proposed the llms.txt convention, and around when?
The guide attributes the proposal to Jeremy Howard in 2024, addressing limited context windows and cluttered HTML.
What content format does llms.txt recommend linking to, rather than full HTML pages?
The proposal recommends concise Markdown links and clean Markdown versions where available; complete site contents are not required.
繼續學習
相關指南
為此主題精選的更多指南