Free WordPress Plugin
A free llms.txt plugin that does not stop at llms.txt. It generates all 10 AI Discovery Files, including ai.txt, identity.json and brand.txt, then logs which AI bots come and read them. Helps ChatGPT, Claude, Gemini and Perplexity discover, understand and correctly cite your website. Free under the GPL, with no pro tier and no feature gating. No coding required.
Most AI visibility plugins only generate llms.txt. This plugin generates all 10 files defined in the AI Discovery Files Specification, the complete set that AI systems need to accurately represent your business.
AI Discovery Files are machine-readable files placed in your website's root directory that tell AI systems (ChatGPT, Claude, Gemini, Perplexity, and Copilot) exactly who you are, what services you offer, and how to refer to your brand. Think of them as robots.txt for AI, except instead of telling crawlers what to ignore, you tell AI systems what to get right.
Every one of the ten has a published specification behind it, with required fields and validation rules. Without them, AI systems fall back on guesswork: confusing businesses with competitors, inventing services, and getting brand names wrong. This plugin generates all ten from your WordPress dashboard, with no coding required.
llms.txtYour business identity in Markdown, the file most AI systems look for first.llm.txtThe singular filename, redirected to llms.txt so either request resolves.llms.htmlThe same identity as a readable web page, with Schema.org markup.ai.txtHow AI systems may interact with, represent and cite your business.ai.jsonThose same permissions as strict JSON, so nothing is left to interpretation.identity.jsonCanonical facts about the business, aligned to Schema.org Organisation.brand.txtHow your name is written, and what it must never be shortened to.faq-ai.txtCitation-ready answers to the questions people actually ask about you.developer-ai.txtTechnical context: your platform, your APIs and your integration points.robots-ai.txtAccess rules for AI training and inference crawlers, bot by bot.AI systems like ChatGPT, Claude, and Gemini are answering questions about your business right now. The question is whether they are getting it right.
AI systems guess who you are based on whatever fragments they find. They confuse your business with competitors, fabricate services you do not offer, and cite your brand name incorrectly.
You give AI systems a single, authoritative source of truth: your name, services, location, permissions, brand rules, and FAQs. They stop guessing and start citing you accurately.
Four steps from install to AI-visible. No coding required.
Search for "AI Discovery Files" in your WordPress plugins directory and click Install.
Fill in your business details. The plugin auto-detects your site name, tagline, pages, and more.
See exactly what each file contains before enabling it. Copy and inspect every line.
Enable the files you want. They are served instantly at the correct root URLs.
Everything a free llms.txt WordPress plugin should do, and the nine other AI Discovery Files besides.
New in v2.0. Reads your own pages to pre-fill FAQs, services and key people as reviewable drafts. Needs WordPress 7.0.
Pulls your brand boilerplate, tagline and service descriptions from pages you already published, using the AI connectors added in WordPress 7.0. It copies rather than invents, and every import arrives as a draft you approve first.
Pulls your site name, tagline, pages, theme, and WordPress version automatically. Less typing, fewer mistakes.
Reads site options and post types rather than parsing your layout, so it behaves the same on classic themes, block themes and Gutenberg. On a standard install there is nothing to configure before the first file is generated.
Generates every AI Discovery File defined in the specification, from llms.txt to robots-ai.txt.
It is an llms.txt generator that does not stop at llms.txt. Most plugins produce that one file. All ten means your permissions, identity, brand rules and crawler policy agree with each other instead of one file speaking alone.
See exactly what each generated file contains before enabling it. Copy to clipboard with one click.
If you have been working out how to create llms.txt by hand, this is the shortcut: read the finished file, adjust the source fields, publish only when it says what you meant. Nothing is served until you switch it on.
Checks every file against the AI Discovery Files specification and flags formatting or content issues.
A built-in validator for required fields, section headings and attribution, run against the published ADF-001 to ADF-010 specifications. For answer engine optimisation the parse matters: a malformed file is a file an AI system skips.
Warns if physical files already exist at the same URLs. No silent overwrites, no surprises.
If you already uploaded an llms.txt file to WordPress by hand, or a caching or SEO plugin is serving one, you are told before anything changes. Your existing file is never deleted and never overwritten.
Start with Essential (2 files) and expand to Recommended (6) or Complete (all 10) when ready.
Essential is the fastest way to add llms.txt to WordPress and be done with it. Recommended adds permissions and identity. Complete publishes the full set. Move up whenever you like; nothing already written is rewritten.
Logs which AI bots read your files, and when. Our own site recorded 43 reads from 11 bots in 30 days.
Named agents, not estimates: GPTBot, ClaudeBot, the Claude web crawler, PerplexityBot, OAI-SearchBot and the rest, with timestamps, response codes and CSV export. Being read is the step before being cited, and this is the part you can verify.
Multisite-aware with per-site settings, plus filters on every generated file and every collected data point.
Roll AI visibility out across a portfolio of client sites: per-site settings on a network, generation scriptable from WP-CLI, and the aidf_generated_content and aidf_collected_data filters to enforce your own house style.
Disable the plugin and the files stop being served. No residual data, no side effects. Your content stays yours.
Free on WordPress.org with no pro tier, no subscription and no account to create. Every feature on this page is in the free download, the crawler log included.
Every other AI visibility tool asks an AI system what it thinks of you, then sells you the answer as data. This plugin does the one thing you can actually verify: it logs each AI crawler that fetches your AI Discovery Files, names the bot, timestamps the request, and records the response code. No prompts, no sampling, no guessing.
Every screenshot on this page is my own live dashboard on mcneece.com, covering the 30 days to 15 July 2026. Not a mockup, not demo data.
in 30 days
named and counted
OpenAI to ByteDance
bots hit robots.txt
I expected llms.txt to dominate. It did not. brand.txt and faq-ai.txt tied on 8 reads each, ahead of llms.txt on 7. OAI-SearchBot's single most-requested file was /brand.txt: OpenAI's search crawler wanted to know what to call me before it wanted my page list.
In a June window on this same site, 5 AI operators showed up. By this July window it was 8. ByteDance's Bytespider, PerplexityBot, GoogleOther and ClaudeBot all appeared for the first time. If you checked once in spring and concluded AI ignores you, your conclusion has expired.
Every logged request returned 200 and nothing was blocked. That sounds unremarkable until you learn how often it goes the other way: a stray robots.txt rule or a host-level bot filter turns AI crawlers away, and without a log you never find out.
/brand.txt before anything else.llm.txt sits honestly at zero: the plugin reports what happened, not what you hoped happened.200.Which bots read which files, with access bars, per-file bot tags, and 7, 30 or 90-day views. Logged at the moment the plugin serves the file, so the data holds up regardless of CDN or caching.
Page-level analytics break behind edge caches, because the edge answers and your server never hears about it. This tracks discovery file responses the plugin serves directly, so it reports honestly on shared hosting, managed WordPress and edge-cached setups alike.
Catches the contradiction where your robots.txt blocks an AI crawler while your permissions say "allow all", and flags it with a link to the fix. This is the check that turns "0 blocked" from a hope into a fact.
Click any bot for its trend, file list and response codes. The activity log filters by bot, date range, status code and filename, and exports to CSV so the evidence leaves the dashboard with you.
A log line proving PerplexityBot fetched your faq-ai.txt is not proof that Perplexity quoted your answer. Retrieval and citation happen inside systems nobody outside those companies can observe, and any tool claiming otherwise is inferring from a sample and calling it measurement. This is deliberately inputs data: it verifies your files exist, are reachable, and are being fetched by named bots at recorded times. That is the part you control, and it is the part that is actually checkable. If you want the reasoning behind that split, read why we think AI mention trackers are a waste of money.
Logging is off by default, so there is zero overhead until you switch it on. It recognises 43 AI crawlers across 8 operator groups, including GPTBot, ClaudeBot, PerplexityBot, GrokBot, Applebot, Bytespider and Gemini-Deep-Research, and you choose which ones to track. Want the longer write-up? We published 11 days of raw crawler logs from this same site, and 30 days of dashboard evidence answering the sceptics who say nothing reads these files.
A complete settings interface for managing all 10 AI Discovery Files.
Start with the essentials and expand when you are ready. Every tier adds more signals for AI systems.
Improve personality, identity, permissions, brand control, and FAQs.
ai.json Machine-parseable permissionsidentity.json Structured business identitybrand.txt Naming and terminology rulesfaq-ai.txt Pre-answered questions for AIFull coverage with developer context, crawler directives, and compatibility.
llm.txt Compatibility redirectllms.html Human-readable referencedeveloper-ai.txt Technical contextrobots-ai.txt AI crawler directivesThen you have a file at that address. Whether it tells an AI system anything useful about your business is a separate question.
Yoast SEO has generated llms.txt free since version 25.3, Rank Math and All in One SEO do the same, and between them they run on 17 million sites. So it is worth being precise about what they actually put in that file, because it is not what most people assume.
They generate a list of your content. Rank Math writes each post's title, URL and a short description, with options for post types and taxonomies. Yoast picks recent and cornerstone pages. What you get is an index of links: a sitemap, written in Markdown, for language models.
That is a fair reading of the original llms.txt proposal, which was written for documentation sites where a curated list of pages is exactly the useful thing. It is much less useful for a business. A list of forty blog post URLs does not tell ChatGPT what you do, which towns you cover, what you charge, or which jobs you turn down.
This plugin publishes a description of the company instead. Not a list of pages, and not limited to what happens to be written on your website. Most of the following appears nowhere on a normal site, because there has never been a reason to put it there:
That is the real difference. One plugin tells an AI which of your pages to read. This one tells it what your company is, who runs it, what it will and will not do, and how it expects to be quoted. Then it logs which AI systems came and took that information. Same filename, an entirely different job.
| Plugin | Installs | What its llms.txt contains | Tells you which AI bots read them | Cost |
|---|---|---|---|---|
| AI Discovery Files | New | A description of the company company facts, key people, brand and citation rules, exclusions, plus 9 more files |
Yes, per file | Free |
| Yoast SEO | 10 million+ | A list of pages recent and cornerstone content |
No | Freemium |
| Rank Math SEO | 4 million+ | A list of pages title, URL and short description |
No | Freemium |
| All in One SEO | 3 million+ | A list of pages plus a full-text llms-full.txt |
No | Freemium |
You can run both. If your SEO plugin already writes a physical llms.txt, this plugin detects the clash and tells you, so you choose which one serves the file. Nothing breaks, and nothing is silently overwritten.
Want the evidence rather than the claim? We published 30 days of AI crawler logs showing which bots read which file on a live site.
Everything you need to know about the AI Discovery Files WordPress plugin.
Installation, setup, and first steps with the plugin.
Install it directly from your WordPress dashboard. Go to Plugins → Add New, search for "AI Discovery Files", and click Install Now. Alternatively, download it from WordPress.org and upload the ZIP file.
The plugin requires WordPress 7.0 or later and PHP 7.4 or later, and is tested up to WordPress 7.1. WordPress 7.0 matters because the optional AI import uses the AI Client that shipped in that release. Everything else works without it.
Yes, completely free and open source under the GPL v2 licence. There are no premium tiers, no upsells within the plugin, and no feature gating. Every capability is available to every user.
llms.txt is a plain-text file placed in your website root that tells large language models who you are, what you do, and what pages matter most. It is rapidly becoming the standard way websites communicate with AI systems, adopted by companies like Stripe, Cloudflare, and Dell. This plugin generates llms.txt plus 9 additional AI Discovery Files directly from your WordPress data, giving AI systems the complete picture rather than just one file.
Install this plugin, fill in your business details, and enable the files. The plugin generates machine-readable files at URLs like yoursite.com/llms.txt and yoursite.com/ai.txt that ChatGPT, Claude, Gemini, and Perplexity read when they encounter your site. You can verify it is working with the free AI Visibility Checker.
The AI Discovery Files Pack is a done-for-you service where experts research your business and write the files for you. The WordPress plugin lets you create and manage the files yourself through a settings interface. The plugin is for WordPress sites; the Pack works for any website.
Caching, conflicts, performance, and developer features.
Yes. AI Discovery Files are served via WordPress rewrite rules with appropriate cache headers. They are compatible with WP Super Cache, W3 Total Cache, LiteSpeed Cache, and other popular caching plugins.
The plugin detects existing physical files and warns you on the Status tab. It will not override physical files. You should remove the physical file before enabling the plugin-generated version.
Yes. Every generated file can be filtered using WordPress hooks. Use the aidf_generated_content filter to modify any file's output, or aidf_collected_data to adjust the data before generation.
robots.txt tells search engine crawlers which pages to index or ignore. llms.txt tells AI systems who you are, what you do, and what pages contain the most important information. They serve different purposes and both should be present. This plugin generates llms.txt alongside robots-ai.txt, which handles AI-specific crawler permissions, giving you complete control over both discovery and access.
No. File generation only runs when specific URLs are requested (e.g. /llms.txt), so there is zero overhead on normal page loads. AI Crawler Analytics only tracks access to discovery file URLs that the plugin serves, so normal page requests are completely unaffected. When logging is disabled (the default), no tracking runs at all.
AI Crawler Analytics logs which AI bots read the AI Discovery Files this plugin generates: llms.txt, ai.txt, identity.json, and the rest. You see which bots accessed which files, how often, when they last called, and whether any are being blocked by your robots.txt. That gives you direct evidence that GPTBot, ClaudeBot, PerplexityBot and other AI crawlers are fetching your files, rather than an inference from prompting an AI system. On our own site it recorded 43 file reads from 11 distinct bots across 8 AI operators in the 30 days to 15 July 2026, and 55 reads from 9 bots in the 30 days to 4 August. We published the full log and what it does and does not prove. The feature works reliably on every hosting platform, including sites behind CDN edge caching, because the plugin controls the discovery file responses. It includes bot detail drill-downs, a filterable activity log with CSV export, and a WordPress dashboard widget. See the AI Crawler Analytics screenshots for what it looks like with real data.
When AI Crawler Analytics is switched on, the plugin writes one row per request to a discovery file: the bot name matched from the user agent, which file was requested, the timestamp and the HTTP status code. Rows are stored in your own database on your own hosting. Nothing is sent to us or to any third party.
A daily task deletes anything past your retention window, which defaults to 90 days and cannot be set below 30. Logging is off by default, so no data is collected at all unless you opt in.
Because crawler logs can constitute personal data under UK GDPR, treat this as you would your server access logs: set a retention period you can justify, and mention the logging in your privacy policy. If you would rather not store it, leave the feature switched off. Every other part of the plugin works without it.
Deactivating stops the files being served, so yoursite.com/llms.txt and the rest return 404 again. Nothing else on your site changes, because the plugin never writes files to disk and never edits your posts or pages.
Deleting the plugin removes its settings and its crawler log tables, so no orphaned data is left behind. Export the crawler log to CSV first if you want to keep the history.
The crawler registry covers 43 AI crawlers grouped into 8 categories: Major AI Assistants (GPTBot, ChatGPT-User, OAI-SearchBot, ClaudeBot, Claude-User, Claude-SearchBot, PerplexityBot, Perplexity-User, GrokBot, xAI-Grok, MistralAI-User), Google AI (GoogleOther, Google-CloudVertexBot, Gemini-Deep-Research, Google-NotebookLM), Apple and Microsoft (Applebot, Applebot-Extended, bingbot, AzureAI-SearchBot), Meta, Amazon, AI Search Engines, Training Crawlers (including CCBot and Bytespider), and Chinese AI (including DeepSeek and PanguBot). The default selection covers the major platforms, and you can tick or untick any individual crawler.
How the plugin affects your visibility to AI systems and search engines.
AI Discovery Files focus on the infrastructure side of generative engine optimisation (GEO). Rather than tracking how often AI mentions your brand, this plugin ensures AI systems have accurate, structured data about your business to work with. It complements traditional SEO plugins like Rank Math or AIOSEO. They handle search engine optimisation, while AI Discovery Files handle AI visibility.
AI Discovery Files are designed for all major AI systems including ChatGPT (OpenAI), Claude (Anthropic), Gemini (Google), Perplexity, Copilot (Microsoft), and other large language models. The llms.txt format in particular has been adopted by companies including Stripe, Cloudflare, and Dell.
AI systems draw from whatever information they can find about your business. AI Discovery Files give you a way to provide authoritative, structured data: your correct business name via brand.txt, your actual services via identity.json, and pre-answered FAQs via faq-ai.txt. This does not guarantee what AI says, but it gives AI systems a reliable source to cite rather than guessing from third-party content.
The files simply stop being served. Your settings are preserved so you can reactivate later without re-entering anything. If you delete the plugin entirely, all settings are removed cleanly.
Use the free AI Visibility Checker to scan your site after enabling your files. It validates every AI Discovery File, checks for conflicts, and gives you a score with actionable recommendations.
Yes. Once your files are live, submit your site to the AI Visibility Directory, the verified registry of websites implementing AI Discovery Files. It provides additional visibility and a dofollow backlink to your website.
Install the free plugin, configure your business details, and start being accurately represented by AI systems today.
Free under GPL v2 · No lock-in · No tracking
" and &. llms.html keeps proper HTML escaping; the JSON files were never affected.Lang: header, ai.json and identity.json carry a language property, and llms.html keeps its lang attribute. BCP 47 format, zero setup.$schema reference, permissions and restrictions as structured action lists, plus attribution, scope, licensing, and metadata. ai.txt and ai.json render from the same internal mapping, so they can never contradict each other.legalName, alternateName, sameAs, contactPoints, identifier, areaServed, a $schema reference, and the new Organisation Type field. Country names are converted to ISO codes for addressCountry.User-agent: / Allow: / Disallow: stanzas per RFC 9309 plus a Sitemap: line, replacing the previous shorthand.index, follow).URL: line in faq-ai.txt so AI systems can cite the page behind the answer.ClaudeBot (model training), Claude-User (live visits when a Claude user asks about your site), and Claude-SearchBot (search indexing).OAI-SearchBot (OpenAI search indexing), Perplexity-User (Perplexity live visits), and Applebot-Extended (Apple AI training). Applebot-Extended follows your training policy in the same way as Google-Extended, CCBot, Bytespider, and Meta-ExternalAgent.wp_get_connectors() returns core's connector definitions regardless of whether an API key is actually set. v2.0.1 asks the WP AI Client's own provider registry whether any provider is configured, so the buttons now reflect whether an AI call can actually succeed.# llms.txt: https://example.com/llms.txt) instead of as a custom directive. Parser-safe with every robots.txt implementation, including strict validators. AI tools that deliberately scan for AI Discovery hints still read the URLs out of the comments.<link rel="alternate"> reference in the HTML <head> for each active discovery files-maxage=0 on discovery file responses so CDN edge caches pass through to originai-visibility-verify.txtai-visibility-verify.txt with user-supplied verification code