llms.txt, Honestly
Everyone is telling you to add one file to your website so ChatGPT starts recommending your business. The file is real, it takes five minutes, and I'll give you the template. But it is almost certainly not the thing that gets you recommended — and the stuff that does work is nobody's favorite thing to post about.
Here's the pitch you've seen a hundred times: "ChatGPT started recommending my business and I didn't pay for it. I set it up with one file called llms.txt."
I've made that claim myself. Then I went and checked it against the crawler logs, the platform statements, and the visibility studies — and the honest version is more useful than the pitch. So this page is the honest version: what the file actually does, what got you recommended if it wasn't the file, and the specific order I'd work in if I wanted a business to show up when someone asks ChatGPT for a recommendation in their city.
First: what llms.txt actually is
It's a plain text file at the root of your site — yoursite.com/llms.txt — written in Markdown. It was proposed in September 2024 by Jeremy Howard of Answer.AI as a way to hand a language model a clean, curated map of your site instead of making it crawl through navigation, cookie banners, and JavaScript to find the good stuff.
Think of it as the difference between robots.txt and a menu. Robots.txt says what you're allowed to look at. llms.txt says here's what matters and here's where it lives.
It is a proposed community standard. It is not a W3C spec, it is not a search engine requirement, and — this is the part that matters — it is not something any major AI company has committed to reading.
Six myths worth killing
Myth 1 — "llms.txt is why ChatGPT recommends me"
This is the big one, and the data is not close.
Ahrefs ran server-log analysis across roughly 137,000 domains and found that 97% of llms.txt files received zero requests — not "few," zero. In a separate look at AI bot traffic, the file is almost never fetched: GPTBot, ClaudeBot, PerplexityBot, and OAI-SearchBot overwhelmingly skip it and crawl your HTML directly. When llms.txt does get requested, the largest single category of requester is SEO audit tools checking whether the file exists — not AI systems reading it.
Meanwhile adoption has genuinely exploded. Originality.ai's monitoring of 3M+ websites clocked llms.txt going from about 4,000 instances in June 2025 to just over 36,000 by May 2026 — 8.8x growth. So we have a lot of people putting up a sign that nearly nobody is reading.
Myth 2 — "It's an official standard the AI companies support"
As of mid-2026, no major provider — OpenAI, Google, Anthropic, Meta, or Mistral — has publicly committed to reading llms.txt in production. Google went further and said it out loud: on June 15, 2026 it added a note to its AI-optimization guidance stating llms.txt files are not required for Google Search, and that the file "won't negatively or positively impact your visibility or rankings." OpenAI and Anthropic both point site owners at robots.txt for controlling crawler behavior; llms.txt shows up in their developer docs, which is a different thing entirely.
Myth 3 — "It controls whether AI can use my content"
It does not. llms.txt is a suggestion about organization, not a permission system. If you want to allow or block AI crawlers, that is robots.txt, and the user-agent names are specific — more on that below. Anyone telling you llms.txt keeps your content out of training data is confusing two unrelated files.
Myth 4 — "It's a ranking factor"
There is no retrieval system currently known to weight it. Ranking inside AI answers runs on the same substrate as the rest of the web: whether you're in the index, whether the page answers the question cleanly, and whether other sources corroborate that you exist and are good.
Myth 5 — "Just block the AI crawlers, they're stealing my content"
Blocking is a real choice with a real cost, and most people making it don't know they're making two decisions at once. GPTBot is OpenAI's training crawler. OAI-SearchBot is the one that builds the index ChatGPT Search reads from when it answers a live question and cites sources. Block GPTBot and you've opted out of training. Block OAI-SearchBot and you have removed yourself from ChatGPT's search results. A lot of sites did the second one by accident with a blanket rule and then wondered why they were invisible.
Myth 6 — "Ranking #1 on Google means ChatGPT will find me"
Helps, doesn't carry you. Pages holding Google position 1 do get cited by ChatGPT meaningfully more often than pages outside the top 20 — but only a small minority of the URLs ChatGPT cites are Google top-10 results at all. And in AI Overviews, the share of citations coming from Google's top 10 has been falling, not rising. These are overlapping systems, not the same system. You have to work both.
How ChatGPT actually decides who to recommend
When someone types "best vehicle wrap shop in Salt Lake City" into ChatGPT, roughly this happens:
- It expands your question. One prompt becomes several sub-queries — "vehicle wrap shops Salt Lake City," "car wrap reviews Utah," "vinyl wrap cost SLC." You are not competing for one query, you're competing for a fan of them.
- It retrieves candidates from its search index — built by OAI-SearchBot and, for ChatGPT, drawing heavily on Bing's real-time index rather than Google's.
- It picks what to cite based on how cleanly the page answers, how authoritative and corroborated the source looks, and how fresh it is.
- It writes the answer — and here the model's training knowledge blends in, which is why some businesses get named with no citation at all. That part comes from being written about across the open web over years, not from anything on your server.
What actually works, in the order I'd do it
This is ranked by impact-per-hour for a local or service business, not by what's fun.
1. Google Business Profile, filled out to the last field
For local recommendations this is the single highest-leverage asset, and it's free. It has become a primary structured data source for AI systems answering "near me" style questions. Fill in every field: exact categories (primary and secondary), service areas, hours including holiday hours, services with descriptions and prices, attributes, and a real Q&A section you seed yourself. Post to it. Upload real photos with real filenames.
The scale of the opportunity here is stark. SOCi's 2026 Local Visibility Index looked at 350,000+ locations and found only 1.2% were recommended by ChatGPT, versus 35.9% appearing in Google's local 3-pack. Almost nobody has done this work yet.
2. Reviews — volume, recency, and across more than one platform
A rating around 4.3+ that appears consistently on Google, Yelp, Facebook, and the vertical sites for your trade does more for AI recommendation than your backlink profile does. Recency matters as much as the average; a 4.9 with nothing new in fourteen months reads as a business that may not be operating. Ask every finished customer, every time, with a link. That's the whole system.
3. Get mentioned on sites that aren't yours
This is the one people skip because it can't be done in an afternoon, and it's why the businesses that do it win. Analyses of what ChatGPT cites when making recommendations consistently show third-party sources dominating — Wikipedia and Reddit sit near the top of the list, with editorial sites and directories behind them. For local trades specifically, the platforms that reliably feed AI recommendations are Google Business Profile, Yelp, Angi, HomeAdvisor, Nextdoor, and the BBB.
Practical version: claim and complete every relevant directory listing. Get into local "best of" roundups — email the writer, most will take a well-written pitch. Participate honestly in the subreddit for your city or trade under a real identity; do not astroturf, it gets caught and it makes you a liability in the exact corpus you're trying to influence.
4. Identical NAP everywhere
Name, address, phone — byte-identical across every listing. "Ste 4" in one place and "Suite #4" in another is enough ambiguity to cost you an entity match. This is boring and it is load-bearing.
5. Let the right crawlers in
Check yoursite.com/robots.txt right now. If it blocks OAI-SearchBot you are not in ChatGPT Search, full stop. A reasonable default for a business that wants to be found:
# Search crawlers that produce citations — allow these User-agent: OAI-SearchBot Allow: / User-agent: PerplexityBot Allow: / User-agent: Claude-SearchBot Allow: / User-agent: Google-Extended Allow: / # Training crawler — your call. Allowing it helps future # models know you exist. Blocking it costs you nothing today. User-agent: GPTBot Allow: / Sitemap: https://yoursite.com/sitemap.xml
Note that these are separate user-agents needing separate entries — a rule for GPTBot does nothing for OAI-SearchBot. Also make sure your host or CDN's bot protection isn't blocking them at the edge, where robots.txt never gets consulted. Cloudflare's AI bot blocking is on by default on some plans and has quietly removed a lot of businesses from AI search.
6. Be in Bing's index
Because ChatGPT's retrieval leans on Bing, and most business owners have never once thought about Bing. Sign up for Bing Webmaster Tools, submit your sitemap, verify the site, and turn on IndexNow so new pages get picked up in hours instead of weeks. It takes twenty minutes and it is the least contested surface in this entire document.
7. Write pages that answer one question, immediately
The structural pattern that gets extracted and cited:
- One page, one question. "How much does a full vehicle wrap cost in Utah?" is a page. "Services" is not.
- Answer in the first two sentences. Before the story, before the trust-building, before the hero image. If a model has to read 400 words to find the answer, it'll cite the competitor who put it up top.
- Real specifics. Actual price ranges, actual turnaround times, actual neighborhoods you serve, actual brands you carry. Specific numbers get quoted. "Competitive pricing" gets skipped.
- Headings that are questions, short paragraphs, and a comparison table where a comparison is warranted — tables get lifted into answers constantly.
- Say who you are and where. "We're a vehicle wrap shop in Salt Lake City serving Davis and Utah County" beats every clever tagline for this purpose.
8. Schema markup
Structured data isn't magic but it removes ambiguity about what your business is. Minimum viable set: LocalBusiness (or the specific subtype), Service for each service, FAQPage on your FAQ, and sameAs links from your Organization schema out to every profile you own — Google Business Profile, Yelp, Facebook, Instagram, LinkedIn. That sameAs block is how you tell a machine that all those listings are one entity.
9. Keep it fresh
Freshness is a consistent citation driver across every study of this. A page updated this quarter beats an identical page from 2023. Put a visible "Last updated" date on your key pages and make it true.
Now make the llms.txt anyway — here's the template
After all that: I still recommend making the file. Three honest reasons.
- Agents are the actual use case. The emerging thing worth being early on isn't AI search — it's autonomous agents that navigate sites to book appointments, pull docs, or compare options on someone's behalf. A clean site map in Markdown is genuinely useful to those, and that traffic is coming.
- It costs five minutes, once. Expected value is low but the cost is nearly zero and the downside is zero. Google has explicitly said it can't hurt you.
- Writing it is a diagnostic. If you can't state in eight lines what your business does, where, for whom, and why someone picks you — that's not a file problem, and you just found the actual problem.
Fill this in, save it as llms.txt, and drop it in the root of your site so it lives at yoursite.com/llms.txt:
# Summit Vehicle Wraps > Vehicle wrap, paint protection film, and window tint shop in > Salt Lake City, Utah. Family owned since 2015, 3M Preferred > installer, 8,000 sq ft climate-controlled install bay. ## What we do - Full and partial vehicle wraps — commercial fleets and personal - Color change wraps - Paint protection film (PPF) - Ceramic coating - Window tint ## Where we work Salt Lake City, Sandy, Draper, Lehi, Provo, Ogden, Park City. Shop at 123 Example St, Salt Lake City, UT 84101. ## Pricing - Full wrap, standard sedan: $3,200–$4,500 - Partial wrap: $1,400–$2,600 - Full front PPF: $1,800–$2,400 - Turnaround: 3–5 business days ## Why customers choose us - 3M Preferred installer with certified installers on staff - 5-year warranty on materials and labor - 4.9 stars across 340+ Google reviews - In-house design team, no outsourced print ## Key pages - [Vehicle wrap pricing](https://example.com/pricing): full cost breakdown by vehicle type - [Fleet wraps](https://example.com/fleet): volume pricing and scheduling for commercial fleets - [Wrap care guide](https://example.com/care): how to wash and maintain a wrap - [FAQ](https://example.com/faq): durability, removal, resale value, insurance ## Contact Phone: (801) 555-0100 Email: hello@example.com Hours: Mon–Fri 8am–6pm, Sat 9am–2pm
How to find out where you actually stand
Fifteen minutes, no tools, do it today. Open ChatGPT with web search on and run the queries a real customer would — not your business name:
| Ask this | What you're testing |
|---|---|
| "Best [your service] in [your city]" | Are you in the consideration set at all? |
| "Who should I hire for [specific problem] near [city]?" | Problem-first queries — where most real intent lives |
| "How much does [your service] cost in [city]?" | Whether your pricing page is the source being quoted |
| "[Your business name] — what do people say about them?" | What the model already believes about you, and from where |
| "[Competitor] vs [you]" | How you're framed head-to-head |
Two things to write down each time: who got named, and which sources got cited. That citation list is your actual to-do list — those are the pages and platforms you need to be on. Re-run the same five prompts monthly. Answers vary between runs, so track the trend across several months, not any single result.
The 30-day version
| When | Do this | Time |
|---|---|---|
| Day 1 | Run the five prompts above. Screenshot everything. Check robots.txt and your CDN bot settings. | 1 hr |
| Week 1 | Google Business Profile to 100% — every field, categories, services with prices, 20+ photos, seeded Q&A. | 3 hrs |
| Week 1 | Bing Webmaster Tools, sitemap submitted, IndexNow on. Write the llms.txt while you're in there. | 45 min |
| Week 2 | NAP audit across every listing you can find. Fix every mismatch. Claim what's unclaimed. | 2 hrs |
| Week 2 | Add LocalBusiness, Service, and FAQPage schema. Wire up sameAs to every profile. |
2 hrs |
| Week 3 | Write three answer-first pages: pricing, your top service by revenue, and a real FAQ. Answer in the first two sentences. | 4 hrs |
| Week 3 | Turn on a review request that fires automatically after every completed job. This is the compounding one. | 1 hr |
| Week 4 | Pitch two local roundups. Answer questions in your city/trade subreddit as yourself. | 2 hrs |
| Day 30 | Re-run the same five prompts. Compare to your Day 1 screenshots. | 30 min |
Under twenty hours, nearly all of it free. Expect movement in Bing and Google Business Profile within weeks; expect the third-party mention work to take a couple of months to show up in answers. Anyone promising you AI visibility in seven days is selling the llms.txt story again in a new hat.
Or we do it for you
Everything on this page is genuinely DIY-able, and if you've got the twenty hours you should just go do it. Most owners I talk to don't — they're the one answering the phone, running the estimates, and texting the crew at 9pm.
So we do this for businesses. It's part of how we work as a fractional Chief AI Officer: the AI search visibility work above, plus the thing that makes it worth doing — catching every lead that comes in from it, following up on every quote automatically, and getting the business out of your head and into a system your team can run.
Book a free 30-minute audit and I'll run the prompts against your business live, show you exactly where you're showing up and where you're not, and give you the fix list in priority order. You keep the roadmap whether we ever work together or not.
Sources
Because a page arguing about evidence should show its own. Checked July 2026:
- PPC Land — llms.txt adoption rises 8.8x but 97% of files get zero AI requests (Originality.ai monitoring of 3M+ sites; Ahrefs server-log analysis of ~137,000 domains)
- Digital Strategy Force — Does your site need llms.txt to get cited by AI search in 2026? (Google's June 15, 2026 guidance note; AI bot traffic analysis)
- CrawlerCheck — OAI-SearchBot user-agent reference and No Hacks — The AI user-agent landscape in 2026
- PushLeads — AI search visibility for local service businesses (SOCi 2026 Local Visibility Index, 350,000+ locations; BrightLocal 2026 Local Consumer Review Survey)
- Stackmatix — SearchGPT ranking factors and The Digital Bloom — LLM ranking factors in 2026 (query fan-out, citation drivers, Google-position correlation)
- The original llms.txt proposal — Jeremy Howard / Answer.AI, September 3, 2024