Skip to main content

llms.txt for law firms: does it actually matter?

It's the file everyone suddenly wants to sell you. Here's what llms.txt actually is, why no major AI engine reads it yet, and the honest reason to ship one anyway.

FirmForte field-guide hero card for the article: llms.txt for law firms: does it actually matter?

The short answer

No major AI provider fetches llms.txt from an ordinary business website — not OpenAI, not Anthropic, not Google — and server logs rather than opinion are what show it. It takes ten minutes to publish if you want one, and it will not get you cited. What does get a firm cited is crawler access, answers written to be lifted, and corroboration from sources you don't control.

If an agency has pitched you a "llms.txt setup" as the thing that gets your firm into ChatGPT, save your money. As of 2026, no major AI engine reads llms.txt from a law firm's website. The file is real, it's cheap to make, and there's a modest case for having one. Getting cited is not part of that case, and anyone selling it that way is selling you the 2026 version of a meta keywords tag.

Here's what it is, what it does, and the honest reason to spend an afternoon on it and then forget about it.

What is llms.txt, exactly?

llms.txt is a plain-text file you put at the root of your domain that hands an AI model a curated reading list of your site. It was proposed in September 2024 by Jeremy Howard of Answer.AI as a way to point language models at your most important pages in clean Markdown, instead of making them wade through navigation, scripts, and cookie banners.

The idea borrows its shape from robots.txt and sitemap.xml. You write an H1 with your firm's name, a short blurb, then a Markdown list of links: your practice-area pages, your attorney bios, your contact page, your best guides, each with a one-line description. Some sites also publish an llms-full.txt that inlines the full text of those pages. The llms.txt proposal at llmstxt.org lays out the format. It's a sensible idea on paper. The problem is what's happened since.

Do ChatGPT, Claude, and Google actually read it?

No. As of 2026, no major AI provider fetches llms.txt from an arbitrary business website. Not OpenAI, not Anthropic, not Google. This isn't a hedge or a "results may vary." It's what shows up in server logs: the AI crawlers that visit your site don't request the file at all.

Ahrefs ran the check in mid-2025 and reported flatly that no major LLM provider supports llms.txt. Google's own search reps have said the same about Search: John Mueller compared it to the keywords meta tag, the one search engines stopped trusting a decade ago because the site owner controls it and can say anything. One log study covered by PPC Land found that 97% of published llms.txt files received zero AI requests, even as the number of files kept climbing.

97% Share of published llms.txt files that got zero requests from AI crawlers in a 2025 server-log study reported by PPC Land. The file is being made. It isn't being read.

Where does the confusion come from? OpenAI, Anthropic, and Perplexity all host an llms.txt on their own developer-documentation sites. That gets passed around as proof the standard has arrived. It hasn't. Those files exist so coding assistants like Cursor and Claude Code can pull structured API docs. That has nothing to do with whether GPTBot or ClaudeBot checks your firm's llms.txt when it's answering "best estate planning attorney in Tulsa." It doesn't.

So why would a law firm publish one at all?

Because it's close to free and it's good hygiene, not because it earns citations. It takes an afternoon, it can help IDE agents and internal tools that do read the file, and it might matter later if adoption ever turns real. That's the whole upside. Ship it with those expectations and you won't be disappointed.

This is the same stance we take on schema markup, and it's worth repeating because the sales pitch is identical. Structured data is hygiene and rich-result eligibility, not a trick that makes an engine cite you. llms.txt sits one rung below even that, because schema is at least read and used today. Publishing an llms.txt is fine. Paying a monthly retainer line item for it is not. If a proposal has "llms.txt optimization" as a recurring deliverable, you're being charged for a file that takes less time to write than the invoice does to read.

What are the common misconceptions about it?

Most of the confusion comes from treating llms.txt like a switch. It isn't one. A few beliefs worth clearing up before you spend any time on it.

"It controls how AI sees my site." It doesn't. llms.txt is a suggestion, not an instruction. Even a model that did read it would still crawl your actual pages, and your real HTML is what it quotes. The file can point at a page; it can't change what that page says or force anyone to read it.

"It's like robots.txt, so engines have to respect it." robots.txt works because crawlers agreed, years ago, to check for it and honor it. That agreement doesn't exist for llms.txt. Borrowing the filename and the root-of-domain placement doesn't borrow the obedience. A convention only matters when the other side has opted in, and for a law firm's site, they haven't.

"llms-full.txt feeds the model my whole site, so it'll know everything." Inlining every page into one giant file mostly creates a maintenance problem. Now you've got a second copy of your content that drifts out of date the moment you edit a page, and you're trusting a file almost nobody fetches to carry it. If the content matters, it belongs on the page itself, marked up cleanly, where engines are already looking.

"Publishing one is a ranking signal." There's no evidence it moves anything in search or in AI answers, and Google's own comparison to the keywords meta tag cuts the other way. A signal the site owner fully controls and can fill with anything is exactly the kind engines learn to ignore.

Who is it actually worth it for?

The honest answer is that it depends on how you'd spend the afternoon otherwise. For most law firms, llms.txt is a fine ten-minute chore and a bad use of anything more than that.

It's genuinely worth doing if you already have your fundamentals in order, your pages answer real questions, your SEO foundation is solid, your entity is consistent, and you just want the hygiene box ticked. In that case the file is close to free and there's a small chance it pays off later if adoption turns real. It's also more defensible if you run internal tools or IDE agents against your own content, since those do read the file, though that's a narrow case for a law firm and has nothing to do with getting cited.

It's not worth it, and it's an active distraction, if your site still has the problems that actually keep firms out of AI answers: thin practice-area pages, content locked behind JavaScript an engine won't run, an inconsistent name and address across the web, or no real authority off your own domain. Shipping a tidy llms.txt while those sit unfixed is polishing the mailbox on a house with no address. Fix the house first. If you're not sure which category you're in, the free audit tells you plainly.

A quick illustrative example

Here's a hypothetical to make the priority concrete. It's made up, not a client result.

Say two estate-planning firms in the same city both decide to "do AEO" this quarter. Firm A pays an agency for a monthly "llms.txt optimization" line item and calls it done. Firm B ignores llms.txt entirely and instead rewrites its three main practice-area pages to answer the questions people actually type, fixes a name-and-address mismatch between its site and its Google Business Profile, and makes sure the pages render without JavaScript so a crawler can read them.

When someone asks an AI engine for an estate-planning attorney in that city, the engine does what it always does: it looks for a page that answers the question cleanly, from a source it can verify is real. Firm B built exactly that. Firm A published a file nobody fetched. Neither firm can force a citation, and we'd never claim a specific engine will name either one. But only one of them did the work that engines are actually reading. That's the whole point of the example, the retainer went to the wrong file.

What actually gets a firm cited instead?

The work that was already working before llms.txt existed. Answer-first content that responds to the questions people actually ask, an entity that's consistent everywhere your firm is named, clean HTML an engine can read without running JavaScript, and genuine authority signals off your own site. None of that is a file you drop at your root.

We've written the long version of this in how AI engines decide which law firm to cite, and the practical checklist in how to get cited by ChatGPT. The short version: engines quote the page that answers the question cleanly and comes from a source they can verify is real. The markup, including the schema types every law firm site needs, makes your content legible. The content is what earns the citation. Skip the content work and the best-formatted llms.txt in the world changes nothing.

How to make one in ten minutes (and then move on)

If you want one anyway, here's the whole job. Open a text file, write your firm's name as an H1, add a sentence describing what you do, then list your 8 to 15 most important URLs in Markdown with a short description each. Save it as llms.txt, upload it to the root of your domain, and you're done.

Pick the pages you'd actually want an AI to read: each practice-area page, your about and attorney pages, your contact page, and two or three of your strongest guides. Skip the thin stuff. If you'd rather not hand-write the Markdown, our free llms.txt generator builds a clean one from your site in a couple of minutes, with no email required. Make it, publish it, and put your real attention on the things engines are reading right now.

That's the honest position. llms.txt is a tidy idea that the industry got out ahead of. Have one if you like the hygiene. Just don't confuse a file nobody's fetching with the work that gets a firm into the answer. If you want to see which of that real work your site is missing, that's what the free audit is for, and what our AEO service builds in from launch.

Questions we get about this

  • What is llms.txt?

    A proposed plain-text file at the root of a website, intended to give AI models a curated map of the site's most useful content — conceptually similar to robots.txt or a sitemap, but aimed at language models. It's a proposal from the community rather than a standard any engine has committed to. The idea is reasonable and the adoption is the problem. Understanding what it is matters mainly so you can recognize it being oversold.

  • Do ChatGPT, Claude, or Google read llms.txt?

    No. As of 2026 no major AI provider fetches llms.txt from an arbitrary business website, and this is visible in server logs rather than being a matter of opinion — the AI crawlers that do visit your site don't request the file. Google has publicly compared the idea to the keywords meta tag, which search engines stopped trusting because the site owner controls it and can say anything. If a vendor tells you otherwise, ask them to show you the requests in your own logs.

  • Should a law firm publish an llms.txt file anyway?

    Only if it costs you ten minutes and you have no illusions about it. There's a reasonable argument for publishing one as a low-cost hedge in case adoption arrives, and no argument for paying anyone to produce one or treating it as an AEO deliverable. What it must not do is displace the work that does matter — crawler access, extractable answers, entity consistency, third-party corroboration. Publish it, forget it, and spend the attention elsewhere.

  • What actually gets a law firm cited instead?

    Being reachable by AI crawlers, answering real questions directly enough to be quoted, being consistent enough across the web that an engine can tell which firm you are, and being referenced by sources you don't control. Those four are unglamorous and none of them are a file you upload. That's precisely why a one-file solution is appealing and why it doesn't work. Start by checking that your robots.txt isn't blocking the crawlers, which takes about a minute.

Share