Free tool

Free AI Readiness Checker

Check whether ChatGPT, Claude, Gemini, Perplexity and other AI assistants can access, read and understand your website. Free, with plain-language results and the official sources behind every rule.

Free check of the home page. Only public information is used.

Checked against the official documentation of the AI vendors, RFC 9309 and schema.org.

  • Free, no account needed
  • Only public information: nothing is changed on your site
  • Results in plain language, with the evidence

What we check

What does the AI Readiness Checker check?

The checker answers three questions about your website: may and can AI crawlers access it, can they read the content, and can they understand what it is about. Every rule comes from an official source: the documentation of the AI vendors, the robots.txt standard RFC 9309, schema.org and Google Search Central.

AreaWhat we checkIncluded in
Access (35 points)What robots.txt allows each of the 27 official AI crawlers, on the home page and the pages linked from it; a test with each crawler’s User-Agent against your firewall (for example Cloudflare); meta robots and X-Robots-Tag directives such as noindex or nosnippet.Free check
Reading (35 points)How much of the main text is in the HTML without JavaScript, compared with the page rendered in a real browser; status code, redirects and canonical; the sitemap and whether it is listed in robots.txt; response time.Free check
Understanding (25 points)Structured data (JSON-LD, Microdata, RDFa) validated against the schema.org vocabulary and the fields Google requires for Product, LocalBusiness and BreadcrumbList; structured data added only by JavaScript.Free check (home page)
llms.txt (5 points)Whether /llms.txt exists and follows the proposed format of llmstxt.org, and whether its links work.Free check
Deep checkUp to 25 pages from your menu and sitemap, each rendered and checked: headings, language, hreflang, titles and descriptions, image alt texts, About / Contact / legal pages, authors and dates, FAQ sections.Deep check (verified websites)

The check only reads public pages and files, like a visitor or a crawler does. It never logs in and never changes anything on your website.

How it works

How does the AI readiness check work?

You enter an address and pass a short automatic check against bots. Our server then reads your robots.txt, sitemap, home page and llms.txt with a clearly named scanner (EasySoftScanner/1.0), repeats the request with the User-Agent of each AI crawler, renders the page in a headless browser and shows the report on this page, usually in about a minute.

  1. Enter the address. Please check websites you own or manage.
  2. We read the rules. robots.txt is parsed exactly as RFC 9309 describes, for every official AI crawler, with the rule and line that applies.
  3. We test the access. The same page is requested as a browser and as each crawler, so a firewall block or a challenge page becomes visible.
  4. We compare and validate. The raw HTML is compared with the rendered page, and the structured data is checked against schema.org.
  5. You get the report. A summary in plain words, what to do first, and the technical details with the evidence and the sources.
Close-up of chips on a dark circuit board
Long library aisle between tall shelves full of books

Reading the results

How do I read the results?

The crawlers are grouped by what they do, because blocking them has different consequences. Blocking training is a choice, not a problem: it does not lower your score. We do not tell you to allow everything; the report explains what each rule means so that you can decide.

  • Training Crawlers that collect content to train AI models, for example GPTBot, ClaudeBot or Google-Extended. Blocking them is legitimate and has no effect on search.
  • Search Crawlers that index pages for AI search and answers, for example OAI-SearchBot, Claude-SearchBot, PerplexityBot, Googlebot and Bingbot. Blocking them usually means your pages cannot be shown as a source.
  • On user request Fetches that happen when a person asks an assistant to open your page, for example ChatGPT-User or Claude-User. Some vendors state that these may not follow robots.txt.

Each result also has a confidence level. Confirmed means direct proof, for example the exact robots.txt rule. Probable means strong clues; the firewall test is never more than probable, because firewalls can also check the IP addresses of the real crawlers. Needs verification means only you can decide, for example whether a canonical link is intended. The score from 0 to 100 weighs access 35, reading 35, understanding 25 and llms.txt 5.

Example report

What does an AI readiness report look like?

Below is the real report of a test website from our lab. Its robots.txt blocks some search and user-request crawlers, a firewall rule blocks PerplexityBot, Cloudflare shows a challenge to another crawler, and there is no llms.txt and no structured data. It is not a customer website.

Example report Test site ai-block.corpus.test from our scanner test lab, built with deliberate robots.txt and firewall rules for AI crawlers. Not a customer website. Excerpt: 9 of 17 results (the informational ones are left out; the score is that of the full scan).

AI Readiness report · Basic scan

ai-block.corpus.test

Scanned on 30.09.2026 00:35 · http://ai-block.corpus.test/

81/100

How the score is calculated
CategoryPoints
Access for AI agents 21 / 35
Readable content 33 / 35
Structured data and identity 23 / 25
llms.txt 4 / 5

Scoring formula version ai-1. Each category starts with its maximum; points are deducted only for findings in this report.

Some things to improve

AI agents can reach the site, but some things make it harder to read or understand.

Score 81/100: the share of the checked points that are in order (higher is better). It covers what can be checked from outside, not how AI services decide to use your site.

What to do first

  1. robots.txt blocks 1 AI search / answer engines agent(s) on the homepage Medium If this is not intended, remove or narrow the rules for these agents in robots.txt (the exact rule is in the technical details).
  2. Server/WAF refuses 1 AI search / answer engines agent(s) Medium If this is not intended, allow these agents in the firewall, CDN or security plugin, preferably by the IP ranges the vendors publish.
  3. 4 key page(s) missing from the sitemap Low Add them to the sitemap, or check that they should indeed not be indexed.
How sure are we? The labels explained
Confirmed
We saw direct proof.
Probable
Strong clues, but no direct proof. The reason is shown.
Needs verification
Cannot be seen from outside. Please check it as described.
Actively exploited
The flaw is on the CISA list of vulnerabilities that attackers are known to use. Deal with these first.

AI agents by purpose

AI training 5 of 8 allowed by robots.txt

Blocking AI training is a legitimate choice: it only says that your content should not be used to train AI models. AI search and answer engines use other agents.

AgentOperatorrobots.txtServer access
GPTBotOpenAIBlocked Disallow: / line 5OK
ClaudeBotAnthropicAllowed Allow: /$ line 10OK
Google-ExtendedGoogleBlocked Disallow: / line 5Not tested
Applebot-ExtendedAppleAllowedNot tested
CCBotCommon Crawl FoundationBlocked Disallow: / line 5OK
Meta-ExternalAgentMetaAllowedOK
AmazonbotAmazonAllowedChallenge (Cloudflare)
MistralAI-TrainingMistral AIAllowedOK
AI search and answer engines 10 of 11 allowed by robots.txt

AI search and answer engines may not be able to find, quote and link your site.

AgentOperatorrobots.txtServer access
OAI-SearchBotOpenAIBlocked Disallow: / line 13OK
Claude-SearchBotAnthropicAllowedOK
GooglebotGoogleAllowed Allow: / line 28OK
Google-CloudVertexBotGoogleAllowed Allow: / line 28OK
PerplexityBotPerplexityAllowedRefused (403)
BingbotMicrosoftAllowedOK
ApplebotAppleAllowed Allow: / line 28OK
Meta-WebIndexerMetaAllowedOK
Amzn-SearchBotAmazonAllowedOK
DuckAssistBotDuckDuckGoAllowedOK
MistralAI-IndexMistral AIAllowedOK
User-triggered AI fetchers 6 of 6 allowed by robots.txt

When someone asks an AI assistant about your page, the assistant cannot read it for them. Some vendors say these fetchers may not follow robots.txt at all.

AgentOperatorrobots.txtServer access
ChatGPT-UserOpenAIAllowedOK
Claude-UserAnthropicAllowedOK
Google-AgentGoogleNo robots.txt name (identified by User-Agent only)OK
Google-GeminiNotebookGoogleNo robots.txt name (identified by User-Agent only)OK
Perplexity-UserPerplexityAllowedOK
Meta-ExternalFetcherMetaAllowedOK
Amzn-UserAmazonAllowedOK
MistralAI-UserMistral AIAllowedOK

What we measured

Pages scanned
1
Text readable without JavaScript
/: 100%
robots.txt
found
llms.txt
not found
  • 0 Critical
  • 0 High
  • 2 Medium
  • 5 Low
  • 2 Info

Medium (2)

robots.txt blocks 1 AI search / answer engines agent(s) on the homepage

Medium Confirmed

What it means

Your robots.txt asks these agents for AI search and answer engines not to read your home page: OAI-SearchBot.

Why it matters

AI search and answer engines may not be able to find, quote and link your site.

What to do

If this is not intended, remove or narrow the rules for these agents in robots.txt (the exact rule is in the technical details).

Technical details

Confirmed: robots.txt fetched and evaluated with RFC 9309 group matching and precedence; the matching rule is quoted.

Blocked: OAI-SearchBot. Blocking search agents reduces the chance that AI search and answer engines (and, for Googlebot or Bingbot, classic search and the AI features built on it) can find, cite and link the site.

How to fix it

If you want to be found and cited by these services, remove or narrow the Disallow rules for these agents (exact rule per agent in the evidence). You can keep training agents blocked: the settings are independent.

Evidence

Checked URL
GET http://ai-block.corpus.test/robots.txt
What was found
OAI-SearchBot: Disallow: / (line 13)
Rule
air.robots.blocked
Observed
30.09.2026 00:35

Server/WAF refuses 1 AI search / answer engines agent(s)

Medium Probable

What it means

Your server or firewall refuses requests from agents for AI search and answer engines: PerplexityBot.

Why it matters

AI search and answer engines may not be able to find, quote and link your site.

What to do

If this is not intended, allow these agents in the firewall, CDN or security plugin, preferably by the IP ranges the vendors publish.

Technical details

Probable: Tested with the agent's User-Agent from the scanner's IP. WAFs and CDNs also verify real crawlers by IP address (reverse DNS or published ranges), so real crawlers may be treated differently: at most probable.

Requests carrying the User-Agent of PerplexityBot got an error (status per agent in the evidence) while a normal browser request got 200. Blocking search agents reduces the chance that AI search and answer engines (and, for Googlebot or Bingbot, classic search and the AI features built on it) can find, cite and link the site.

How to fix it

If this is not intended, allow these agents in your firewall/WAF/CDN bot rules (preferably verified by the vendors' published IP ranges, not only by User-Agent).

Evidence

Checked URL
GET http://ai-block.corpus.test/
What was found
PerplexityBot: 403
Rule
air.access.blocked
Observed
30.09.2026 00:35

Low (5)

4 key page(s) missing from the sitemap

Low Confirmed

What it means

Some pages from your main menu are not in the sitemap.

Why it matters

Crawlers may find them later or not at all.

What to do

Add them to the sitemap, or check that they should indeed not be indexed.

Technical details

Confirmed: Compared the main-menu links with every URL in the sitemap.

Pages linked from the main menu but not listed in the sitemap: http://ai-block.corpus.test/about-internal, http://ai-block.corpus.test/news/?page=2, http://ai-block.corpus.test/private/press, http://ai-block.corpus.test/private/archive.

How to fix it

Add the pages to the sitemap (or check that they should not be indexed).

Evidence

Checked URL
GET http://ai-block.corpus.test/sitemap.xml
Rule
air.index.sitemap_missing_pages

No llms.txt file

Low Confirmed

What it means

Your site has no llms.txt file.

Why it matters

Optional. Major AI vendors have not confirmed that they use llms.txt, so the effect is uncertain.

What to do

Optional: publish /llms.txt with your site name, a short summary and links to your key pages.

Technical details

Confirmed: Fetched /llms.txt directly (the server answered with an HTML page).

http://ai-block.corpus.test/llms.txt was not found. llms.txt is a proposed standard (llmstxt.org); major AI vendors have not confirmed using it, and Google states it is not needed for Google Search. Impact uncertain.

How to fix it

Optional: publish /llms.txt with an H1 (site name), a short summary and lists of links to your key pages.

Evidence

Checked URL
GET http://ai-block.corpus.test/llms.txt
Rule
air.llms.missing
Observed
30.09.2026 00:35

robots.txt blocks key pages for 7 AI search / answer engines agent(s)

Low Confirmed

What it means

Your home page is open, but robots.txt closes some important pages to agents for AI search and answer engines: /about-internal, /news/?page=2, /private/archive.

Why it matters

AI search and answer engines may not be able to find, quote and link your site.

What to do

Check that the rules for these pages are intended (the exact rule is in the technical details).

Technical details

Confirmed: robots.txt evaluated for the key pages linked from the homepage menu.

The homepage is allowed, but these agents are blocked on some key pages: Claude-SearchBot, PerplexityBot, Bingbot, Meta-WebIndexer, Amzn-SearchBot, DuckAssistBot, MistralAI-Index. Pages: /about-internal, /news/?page=2, /private/archive. Blocking search agents reduces the chance that AI search and answer engines (and, for Googlebot or Bingbot, classic search and the AI features built on it) can find, cite and link the site.

How to fix it

Check that the Disallow rules affecting these pages are intended (the exact rule is in the evidence).

Evidence

Checked URL
GET http://ai-block.corpus.test/robots.txt
Rule
air.robots.partial
Observed
30.09.2026 00:35

robots.txt blocks key pages for 6 user-triggered AI fetchers agent(s)

Low Confirmed

What it means

Your home page is open, but robots.txt closes some important pages to agents for user-triggered AI fetchers: /about-internal, /private/archive.

Why it matters

When someone asks an AI assistant about your page, the assistant cannot read it for them. Some vendors say these fetchers may not follow robots.txt at all.

What to do

Check that the rules for these pages are intended (the exact rule is in the technical details).

Technical details

Confirmed: robots.txt evaluated for the key pages linked from the homepage menu.

The homepage is allowed, but these agents are blocked on some key pages: ChatGPT-User, Claude-User, Perplexity-User, Meta-ExternalFetcher, Amzn-User, MistralAI-User. Pages: /about-internal, /private/archive. These agents fetch a page when a user asks an AI assistant about it. Blocking them means the assistant cannot read the page for that user. Several vendors state these fetchers may not apply robots.txt at all (see evidence).

How to fix it

Check that the Disallow rules affecting these pages are intended (the exact rule is in the evidence).

Evidence

Checked URL
GET http://ai-block.corpus.test/robots.txt
Rule
air.robots.partial
Observed
30.09.2026 00:35

No structured data (schema.org) found

Low Confirmed

What it means

Your pages contain no structured data (schema.org).

Why it matters

Structured data describes in a machine-readable way who you are and what the page contains.

What to do

Describe your organisation, your site and your main content types with schema.org (most SEO plugins can do it).

Technical details

Confirmed: JSON-LD, Microdata and RDFa extracted from the raw HTML.

Homepage: no JSON-LD, Microdata or RDFa with schema.org types was found.

How to fix it

Describe the organisation (Organization/LocalBusiness), the site (WebSite) and the main content types (Product, Article, BreadcrumbList) with schema.org JSON-LD.

Evidence

Checked URL
GET http://ai-block.corpus.test/
Rule
air.schema.none

Info (2)

Site served through Cloudflare: check the AI crawler settings

Info Confirmed

What it means

Your site is served through Cloudflare.

Why it matters

Cloudflare has its own settings that can block or challenge AI crawlers, independently of robots.txt.

Technical details

Confirmed: Cloudflare response header observed.

Responses carry Cloudflare headers (cf-ray). Cloudflare can block or challenge AI crawlers with dashboard settings (Block AI Bots, AI Crawl Control), independently of robots.txt.

How to fix it

In the Cloudflare dashboard, review Security → Bots (Block AI Bots) and AI Crawl Control; decide per crawler (training vs search) instead of one global switch.

Evidence

Checked URL
GET http://ai-block.corpus.test/
What was found
cf-ray
Rule
air.access.cloudflare
Observed
30.09.2026 00:35

robots.txt: 21 of 25 AI agents allowed on the homepage

Info Confirmed

What it means

21 of the 25 AI agents we know may read your home page according to robots.txt.

Why it matters

The table shows the decision and the exact rule for each agent, grouped by what the agent is used for.

Technical details

Confirmed: robots.txt fetched and evaluated per agent.

Each agent from the official vendor documentation was evaluated against robots.txt with RFC 9309 precedence (the exact matching rule per agent is in the evidence). Blocked: 4.

Evidence

Checked URL
GET http://ai-block.corpus.test/robots.txt
Rule
air.robots.summary
Observed
30.09.2026 00:35

Data sources — last synchronisation

  • AI agent list (official vendor documentation, last checked): 29.09.2026 02:00
  • schema.org vocabulary: 29.09.2026 02:00

AI agents were tested from our server with the User-Agent their vendors publish; real crawlers may be treated differently. Results marked “Probable” or “Needs verification” are indicative. No check can guarantee that a site appears in AI answers.

Data sources

Where do the rules and the crawler list come from?

Only from official sources. The names, robots.txt tokens and purposes of the 27 crawlers come from each vendor’s own documentation; third-party bot lists are not used. A monthly job checks these pages again and reports every change to us for review, and each report shows when the list was last verified.

Glowing blue optical fibers against a dark background

Prices

What does it cost? Free check, deep check with credits or monitoring

The basic check of the home page is free and needs no account. A deep check of up to 25 pages is available for websites whose ownership you have verified: the first one of each domain and account is free, after that it costs one credit. With Monitoring, the AI readiness check runs every month, together with the weekly security scan.

Free

Basic check

€0

  • Home page: robots.txt for every official AI crawler, firewall test, content without JavaScript, schema.org, sitemap, llms.txt
  • No account needed
  • Up to 3 scans per hour and 10 per day per website

Free deep scan One per module for each verified domain and each account (confirmed email address).

Pay per scan

Deep scan with credits

from €1.96 per scan

PackPrice
1 credit€4.90
3 credits€9.90
10 credits€24.90
25 credits€49.00
  • 1 credit = 1 deep scan of a verified website, on either module (security or AI readiness)
  • Deep AI check: up to 25 pages from the menu and the sitemap, headings, language, identity pages and FAQ
  • Credits are valid for 12 months from purchase
  • If a scan fails because of a problem on our side, the credit is returned

Subscription

Monitoring

from €7.90 per month

PlanWebsitesManual rescans / monthPrice
Monitoring Starter1—€7.90 / month
or €79.00 / year
Monitoring Agency10100€24.90 / month
or €249.00 / year
Monitoring Agency 2525250€49.00 / month
or €490.00 / year
  • Weekly security scan and an email when a newly published vulnerability affects a component we detected
  • Monthly AI readiness check
  • History and comparison of scans
  • Included rescans reset every month and are not carried over; cancel anytime, the plan runs until the end of the paid period

Credits and plans are bought in your GetEasySoft account. Customer accounts for the scanner open soon; the free scan is available now.

Prices in EUR.

Close-up of a server rack with small green status lights

Privacy

What happens to my data?

We store the address you checked, the results and the date, so that the report can be shown and shared. Your IP address is used only as a hashed value to limit abuse. The AI features that would send your content to an AI provider are switched off; the check itself uses no AI model.

  • Operator: Einzelunternehmen Alin Ghiorghiu, Kloten, Switzerland.
  • Hosting of the scanner: HOSTON SRL, Romania (EU).
  • Reports: kept for up to 12 months, then deleted automatically. Private links expire after 30 days.
  • Bot protection: Cloudflare Turnstile, loaded only when you use the form.
  • Details: privacy policy.

Help

Want it fixed without spending your evening on it?

We can adjust robots.txt and your firewall rules to the choices you make, add structured data, make the main content visible without JavaScript and write an llms.txt. Send us the address of your website, and you get a clear answer and a fixed price before we change anything.

Let us fix it Free security scan

Laptop showing program code in a dark room lit in blue

Frequently asked questions

Is the AI readiness check free?

Yes. The basic check of your home page is free and needs no account. There is a limit of 3 checks per hour per visitor and 10 per day per website. A deep check of up to 25 pages is available for websites you have verified: the first one is free, then it costs one credit.

Which AI crawlers are checked?

27 agents of 11 operators, taken only from the official documentation of each vendor: for example GPTBot, OAI-SearchBot and ChatGPT-User (OpenAI), ClaudeBot, Claude-SearchBot and Claude-User (Anthropic), Googlebot and Google-Extended, PerplexityBot, Bingbot (used by Copilot), Applebot, Meta, Amazon, DuckDuckGo and Mistral AI.

Should I let AI crawlers train on my content?

That is your decision, and blocking training crawlers is a legitimate choice: it does not lower your score. Blocking search crawlers usually means your pages can no longer be found and quoted as a source in AI answers. The report shows each purpose separately.

Does a good score mean ChatGPT or Google will recommend my website?

No. Nobody can promise that. The check measures technical conditions that you can verify and fix: whether AI crawlers may and can reach your pages, whether the content is readable without JavaScript and whether it is described with structured data.

Why does the firewall test say “Probable” and not “Confirmed”?

We send requests with the User-Agent of each AI crawler from our scanner and compare them with a normal browser request. Firewalls can also check the real IP addresses of the crawlers, which we cannot use, so a block or an allow is at most probable. Your server logs are the confirmation.

Is llms.txt required?

No. llms.txt is a proposed standard (llmstxt.org). The large AI vendors have not confirmed that they use it, and Google says it is not needed for Google Search. It has the smallest weight in the score: 5 of 100 points.

Do FAQ sections still help?

Clear questions with direct answers help people and AI assistants understand a page. The FAQ rich result in Google is a different thing: Google limited it in 2023 and no longer shows it since 7 May 2026. The report says both.

My website uses Cloudflare. What should I check?

Cloudflare can block AI crawlers with “Block AI Bots” or AI Crawl Control, and it can show a challenge page instead of your content. When the check detects Cloudflare or a challenge, the report links to the official Cloudflare settings.

Does the check change anything on my website?

No. It only reads public pages and files, like a visitor or a crawler does: robots.txt, the sitemap, your pages and llms.txt. It never logs in and never sends forms.

How much does a deep check cost?

The first deep check of each verified domain and account is free (confirmed email address). After that one deep check costs one credit, from 4.90 EUR for one credit to 49 EUR for 25. Credits are valid for 12 months and also work for the security scanner.

Check your website now

The free check takes about a minute. If something needs changing and you do not have the time, we can do it for you.

Last updated: · Published by GetEasySoft

Photos: Jakub Pabis on Pexels, Zetong Li on Pexels, Georgie Devlin on Pexels, panumas nikhomkhai on Pexels, Nemuel Sereti on Pexels