Enter a domain and get a draft llms.txt written from your own pages: up to 25 of them, each listed by its title and meta description, in the format proposed at llmstxt.org. Whether AI assistants ever request the file is still an open question, so this page tells you what is known and leaves the decision to you.
Up to 25 pages, four at a time, each allowed 12 seconds. There is no live count to show: the draft arrives in one piece when the last page is in, usually inside a minute.
Up to 25 pages. The homepage first, then the URLs in your XML sitemap in the order it lists them, or the links on your homepage when there is no sitemap. One GET per page as VerandBot/1.0, 12 second timeout, served HTML only. Beside it, one GET for /llms.txt to see whether you already serve one.
An H1 with your domain, a neutral one-line summary, then up to three link lists: Key pages, the first 12 pages read, and Content, the rest. Each line is the page title, its URL and its meta description, with HTML entities decoded; a description over 160 characters is cut and ends with an ellipsis. Any URL containing privacy, terms, cookie, legal or disclaimer goes in a third list, Disclosures. Pages with no title are left out. No model writes any of it.
Know which pages matter most to you: the product ranks them by Search Console clicks and your pillar pages, and this free version has neither. Read a title or description that JavaScript adds after load. Write llms-full.txt or Markdown copies of your pages. Upload anything. Or tell you whether any assistant will request the file.
Find the pages, read what each one says about itself, lay it out in the proposed format, hand it to you. The card is the tool in motion on an example site, looped, and each step lights up while the card is doing it.
The homepage always comes first. The rest come from your XML sitemap, found through the Sitemap: line in robots.txt or the usual paths, in the order it lists them. No sitemap, and the tool follows the links on your homepage instead.
Four pages at a time, each a plain GET with a 12 second limit. From each one it keeps two things: the title and the meta description, exactly as your HTML serves them. A page that errors or times out is not listed.
An H1, a one-line blockquote, then Markdown link lists under H2 headings, which is the order llmstxt.org sets out. The first 12 pages go under Key pages, the rest under Content, and privacy, terms, cookie, legal and disclaimer pages under their own Disclosures heading. Pages with no title are skipped.
The card flags what a person should change before the file goes live: the placeholder summary, descriptions cut mid-sentence, the pages it could not know you care about. Then copy it or download it and serve it at /llms.txt.
A two-year-old proposal, a Markdown file at the root of your site, and one of the most oversold ideas in search right now. Here is what the file is, how it differs from robots.txt and a sitemap, what is actually known about who requests it, and how to decide whether yours is worth ten minutes.
An llms.txt file is a Markdown document served at https://yourdomain.com/llms.txt that gives an AI agent a short, curated map of your site: who you are, in a sentence, and a list of the pages worth reading, each with a line saying what is on it. Jeremy Howard of Answer.AI proposed it in September 2024 and published a second version, titled v2 on llmstxt.org, in August 2026. It is a proposal open for community input, not a standard: no standards body has adopted it, and nothing obliges a search engine or an AI company to read it.
The idea behind it is practical. A web page wraps its information in navigation, scripts and layout, and an agent that has to read a whole site to answer one question wastes most of its effort on that wrapping. A small file that says "start here" is cheap to read. The proposal's own words describe it as used "on demand, when an agent needs information about a topic while assisting a user", and says it has been used mostly for that, rather than for training.
The spec fixes the order of the sections and very little else. Only the first is required.
[name](url), optionally followed by a colon and a note about the page.v1 gave a section named Optional a special meaning, as the links a tool could drop when space was short. v2 removed that mechanical meaning; the heading is still a sensible home for secondary links. v2 also added two things this generator does not do for you: files at subpaths, where /docs/llms.txt covers only the pages under /docs/, and clean Markdown copies of pages at the same URL with .md added, announced with rel="alternate" and rel="describedby" links.
The clearest way to see what a generator can and cannot do is to put its draft beside a file a person wrote for the same site. Willowdale Equity, a multifamily syndication firm and one of the two sites Verand is tested on, serves a hand-edited file today. Its opening lines, fetched on 27 September 2026:
# Willowdale Equity > Willowdale Equity is a private real estate investment firm that acquires Class B and C value-add multifamily assets across the southern United States, then syndicates the equity alongside accredited (and occasionally non-accredited) passive investors. [...] [...] All content reflects the views of Willowdale Equity LLC, which is not a registered investment advisor; nothing on this site constitutes investment, tax, or legal advice. ## About the firm - [About Willowdale Equity](https://willowdaleequity.com/about/): firm overview, founders, mission, values, and track record summary - [Editorial Process](https://willowdaleequity.com/editorial-process/): how content is researched, written, reviewed, and updated; author credentials
And the opening of the draft this tool wrote for the same domain, on the same day:
# willowdaleequity.com
> Content and resources from willowdaleequity.com (willowdaleequity.com).
## Key pages
- [Passive Real Estate Investing | Willowdale Equity](https://willowdaleequity.com/): Private real estate investment firm acquiring Class B & C value-add multifamily across the Southeastern U.S. Join the Investor Club for first-look deals.
[the blog index, then:]
- [10 Year Treasury, How it relates to Multifamily CAPs and values | Blog](https://willowdaleequity.com/blog/10-year-treasury-what-it-is-and-how-it-relates-to-multifamily-cap-rates-and-values/): The 10-year Treasury and multifamily cap rates: how the benchmark rate flows through to commercial real estate valuations and what it means for investors.
The draft is correct and it is thin. It knows the site's pages and what each says about itself. It does not know that the firm wants its disclosure read first, which nine guides are its pillars, or that the tenth article in sitemap order is not the tenth most important. And the order it lists pages in is the order it read them, not the order the firm would choose. Those are the edits the card lists after every run, and they are the ten minutes that turn a draft into a file worth serving.
At the root of the host, as /llms.txt, served as plain text with a 200 status. Not in a folder, not behind a redirect to a login page, and not as an HTML page: a theme that answers every unknown path with a styled 404 page and a 200 status looks, to anything requesting the file, like a site that has one full of HTML. Verand's own checks treat an HTML body at that path as no file for that reason. Under v2 a file can also sit at a subpath and describe only the pages beneath it, which suits a site whose documentation or resource library lives in one section.
Three plain files at the root, three different jobs. They are often confused because they sit side by side.
| File | Its job | What it lists |
|---|---|---|
| robots.txt | Access rules | Which paths each named crawler may or may not request. An instruction crawlers choose to honour, and the one file of the three with a published standard (RFC 9309). |
| sitemap.xml | Discovery | Every indexable URL you want search engines to find, with optional dates. Complete by design, and far too long for anyone to read as a summary. |
| llms.txt | Orientation | A short, chosen set of pages with a note on each, plus a sentence on who you are. Curated by design, and advisory: it grants and blocks nothing. |
So an llms.txt never overrides robots.txt. A page listed in llms.txt but disallowed for a crawler stays disallowed for it, and a crawler you have blocked from the whole site will not be reading the llms.txt either. If you are unsure what your robots.txt allows, the Robots.txt Checker reads it by crawler name. If you are unsure your sitemap is found at all, the Sitemap Checker looks. This generator uses that sitemap to pick its pages, so a missing one changes which 25 pages you get.
This is the question every generator page answers with confidence, and the honest answer is mixed. Here is what the primary sources say, so you can weigh it yourself.
We have no request logs of our own to add to that list. The one source that answers the question for your site is your own server log: search it for requests to /llms.txt and note the user agent on each. If the answer matters to you, look there before and after you publish the file.
Most generators sell the file as a way into AI answers. That claim is unproven, and this page will not make it. There is a smaller case that holds regardless: an llms.txt is a list of pages you chose. If any agent does read it, it lands on the versions you have reviewed, the ones that carry your disclosures and your credentials, rather than on whatever it happened to crawl. The file costs one upload and grants nothing, so the downside is ten minutes.
For a firm whose content is regulated, that means leading with the pages that establish who is speaking and on what terms. An About page with credentials and licensing. The editorial or review process. Author pages. The disclosure or disclaimer page, and for a registered investment adviser, the page that links Form CRS and Form ADV. Then the pillar guides. Willowdale's file ends with a short note telling agents that offerings are made only by private placement memorandum to verified accredited investors, which is the kind of line only the firm can write.
One thing to know about this generator in particular: any URL containing privacy, terms, cookie, legal or disclaimer is listed under its own ## Disclosures heading at the end of the file, never mixed into the content lists and never dropped. On a regulated site that section may matter more than any other, so read it before you upload. It only holds the disclosure pages among the 25 the tool read; if yours sits deeper in the site, add it there by hand, and consider moving it to the top.
llmstxt.org lists several tools that generate the file on their own, including the Yoast SEO and AIOSEO plugins for WordPress, Wix for every Wix site, and the documentation platforms Mintlify and GitBook. If your site runs on one of them, you may already serve a generated file without having written it. The card checks for an existing /llms.txt on every run for that reason; compare before you replace one.
Generate the draft above, make the edits the card lists, and save the result as a UTF-8 text file named exactly llms.txt. On WordPress, either upload it to the web root with your host's file manager or SFTP, the folder that holds wp-admin and wp-content, or let an SEO plugin serve its own and paste your edits there. On a static site, put the file in the folder that becomes the root of the build, usually public/ or static/, and deploy. Then open https://yourdomain.com/llms.txt in a browser and confirm you see plain text. The llms.txt file needs no registration anywhere and no line in robots.txt.
Six things that are true of this tool, each one backed by a line in the code that runs it.
The request carries a domain and nothing else. There is no account, no session and no database behind the tool, so there is nothing for us to keep about you.
Every line is a title and a description your own pages already publish, laid out by fixed rules. No model writes or rewrites anything, so nothing appears in the file that your site does not already say.
This is the function Verand uses to write the fix when a customer's crawl finds no llms.txt. In the product it also ranks pages by Search Console clicks and pillar pages; here it has neither, and the card says so.
The 25-page cap, the placeholder summary, the descriptions it cut, whether a disclosure page made the list: each is stated on the card beside the draft, with the edit that fixes it.
Every run also asks your site for /llms.txt, and treats an HTML page there as no file. If you already serve one, the card tells you before you overwrite it with a draft.
The tool never pays for a rendered page, so a run costs nothing and is never metered. The one limit is a courtesy to the sites being read: 20 runs a minute per visitor.
What the file is, where it goes, and what is and is not known about who reads it.
A Markdown file at the root of a website that gives an AI agent a short, curated map of the site: an H1 with the site's name, a one-line summary in a blockquote, then lists of links to the pages worth reading, each with a note. It was proposed by Jeremy Howard in September 2024 and revised as v2 in August 2026 at llmstxt.org. It is a proposal, not a standard, and it controls nothing: unlike robots.txt it neither grants nor blocks access.
At the root of your domain, so it answers at https://yourdomain.com/llms.txt as plain text with a 200 status. On WordPress, upload it to the web root with your host's file manager or SFTP, or use an SEO plugin that serves one. On a static site, place it in the folder that becomes the root of the build and deploy. Then open the URL in a browser: you should see raw text, not a page in your site's design.
Not confirmed. Google's AI features documentation says you don't need new machine-readable or AI text files to appear in AI Overviews or AI Mode, and Google's John Mueller said in June 2025 that no AI system currently used llms.txt. OpenAI, Anthropic and Google publish llms.txt files for their own developer docs, and the proposal's author reports coding agents using them, but we have found no statement that ChatGPT, Claude or Gemini request the file from the sites they read. Your server logs are the reliable answer for your own site.
robots.txt sets access rules that crawlers choose to honour, and has a published standard. sitemap.xml lists every indexable URL so search engines can find them. llms.txt is a short, hand-picked list of pages with a note on each and a sentence about who you are, meant to orient an agent. It is advisory only: listing a page in llms.txt does not unblock it in robots.txt, and leaving a page out does not hide it.
It is not part of the llmstxt.org proposal. It is a convention from documentation platforms such as Mintlify, which describes it as a file that combines an entire documentation site into one. It suits software docs an agent might load whole. For a firm's marketing and advice site it is rarely worth it, and this generator does not write one: it produces llms.txt only.
Lead with the pages that establish who is speaking and on what terms: About with credentials and licensing, the editorial or review process, author pages, and the disclosure or disclaimer page, plus the page linking Form CRS and Form ADV for a registered investment adviser. Then your pillar guides. This generator lists URLs containing disclaimer, legal, privacy, terms or cookie under their own Disclosures heading; if your disclosure page was not among the pages it read, add it by hand before you upload.
An llms.txt only points at what you have already published. Verand writes articles from your own expertise and credentials, so the pages in the file say something only you can. It tracks where you rank on Google and where ChatGPT, Gemini, Google AI Overviews, Google AI Mode, Perplexity and Claude name you, and gates every draft so a claim your regulator would not allow never publishes.
Content built to rank in
Google and get cited by
ChatGPT
Perplexity
Gemini
Claude, with every claim checked before it goes live.