No reachable sitemap.xml. Publish one and reference it in robots.txt.
具体的な手順を見る
これは何を意味するか
A sitemap is a machine-readable list (sitemap.xml) of every page on your site. It is how crawlers find pages that aren’t linked from your homepage.
なぜ重要か
Agents discovering your site crawl outward from the pages they know. Without a sitemap, anything not linked from a top-level page (docs, older articles, product variants) may simply never be read, so it can’t be quoted or recommended.
やること
Most platforms generate one automatically: Next.js (app/sitemap.ts), WordPress (Yoast/Rank Math), Shopify and Wix (built in). Turn it on rather than hand-writing.
Add a "Sitemap: https://yoursite.com/sitemap.xml" line to robots.txt so crawlers find it.
Submit it once in Google Search Console and Bing Webmaster Tools for the search side.
No reachable robots.txt at the site root. Publish one so agents can discover your crawl policy and sitemap.
0/100
pass
robots.txt allows AI agents
All 9 answer-time access agents allowed.
100/100
fail
Content-Signal directivesemerging
No Content-Signal directives. Add e.g. `Content-Signal: ai-train=no, ai-summarize=yes` to declare granular AI usage policy beyond binary allow/disallow.
0/100
pass
llms.txt present
Found /llms.txt but missing H1 header and markdown links. Not scored.
100/100
fail
Sitemap present
No reachable sitemap.xml. Publish one and reference it in robots.txt.
0/100
skip
Link response headers
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Canonical URL
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
MCP server cardemerging
Not measured: the origin answered HTTP 403 for this resource, which says we were refused or could not be served, not that the resource is missing.
—
skip
OAuth authorization metadataemerging
Not measured: the origin answered HTTP 403 for this resource, which says we were refused or could not be served, not that the resource is missing.
—
skip
OAuth resource metadataemerging
Not measured: the origin answered HTTP 403 for this resource, which says we were refused or could not be served, not that the resource is missing.
—
skip
API catalog / OpenAPIemerging
Not measured: the origin answered HTTP 403 for this resource, which says we were refused or could not be served, not that the resource is missing.
—
skip
Web Bot Authemerging
Not measured: the origin answered HTTP 403 for this resource, which says we were refused or could not be served, not that the resource is missing.
—
可読性
skip
Markdown negotiationemerging
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Text-to-markup ratio
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Content without JavaScript
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Heading hierarchy
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Cookie wall blocks content
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Declared language
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Page title
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Meta description
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
構造化データ
skip
JSON-LD present
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Structured data validates
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Schema type coverage
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
整合性:各チャネルは一致していますか?
skip
JSON-LD price matches the visible priceemerging
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
JSON-LD name appears on the pageemerging
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Declared language matches the contentemerging
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
実行可能性
skip
Paywall / login wall
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Agent Skills manifestemerging
Not measured: the origin answered HTTP 403 for this resource, which says we were refused or could not be served, not that the resource is missing.
—
skip
WebMCP actionsemerging
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Primary action reachable
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
fail
Reachability
Terminal HTTP 403.
0/100
skip
Primary action is clickableemerging
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Filters reachable by URLemerging
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
パフォーマンス
skip
Time to first byte
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Full render time
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
—
skip
Page weight
Not measured: the origin returned HTTP 403 for this URL, so the document received is an error or challenge response rather than the page. Scoring it would describe that response, not your site.
Before reading anything, an agent has to reach your page: one request, a clean answer. Yours didn’t answer cleanly. The finding above says exactly what happened: an error status, a connection failure, or a chain of redirects before any content arrived.
なぜ重要か
Agents give up much faster than browsers. A human waits through three redirects and a retry; an agent on a time budget records "unreachable" and answers from other sources. Every other measurement on this page happens only after this one succeeds, so this failure hides everything else your site does right.
やること
Read the finding above: it names the status code and counts the redirects we followed.
If it is a redirect chain, collapse it to one hop. The classic chain is http → https → www → final page; configure your host to jump straight from any variant to the final address in a single redirect.
If it is a 4xx or 5xx status, your server is refusing or failing the page. Fix the route or the error before anything else on this page.
If the fetch failed outright, check DNS and the TLS certificate; an expired or mismatched certificate turns every agent away at the door.
リポジトリを開いた Claude Code、Cursor などのアシスタントに貼り付けてください。この検出結果、手順、避けるべき近道が含まれます。
確認すること
Run: curl -I https://yoursite.com and confirm a 200, or a single 301 followed by a 200.
Hit Re-scan above: "Reachability" flips to PASS, and checks that were skipped get measured for the first time.
避けること
Don’t serve an error page with a 200 status to look reachable. Agents then read the error text as your content and summarise your business from it.
Don’t block by IP reputation so aggressively that assistant crawlers are refused while browsers pass; the finding above reflects what an agent gets, not what you see.
No reachable robots.txt at the site root. Publish one so agents can discover your crawl policy and sitemap.
具体的な手順を見る
これは何を意味するか
Your site has no robots.txt file at all. That file is the front-door sign for every robot: what to read, what to skip, and where your sitemap lives.
なぜ重要か
Without one, every crawler guesses. Most treat a missing file as "allowed", but you lose the one place to point them at your sitemap and to state your policy explicitly. An explicit welcome is a clearer signal than silence.
やること
Create a plain text file named robots.txt.
Start permissive: "User-agent: *" then "Allow: /", plus a "Sitemap:" line pointing at your sitemap URL.
Serve it at exactly yoursite.com/robots.txt (site root, not a subfolder).
→ https://website.com/robots.txt
リポジトリを開いた Claude Code、Cursor などのアシスタントに貼り付けてください。この検出結果、手順、避けるべき近道が含まれます。
確認すること
Open yoursite.com/robots.txt. You should see your file, not a 404 page.
Re-scan: "robots.txt present" flips to PASS, and if you added the Sitemap line, "Sitemap present" often flips with it.
避けること
Don’t copy a robots.txt from another site. Theirs encodes their policy, including blocks you may not want.
Don’t return an HTML error page at that URL with status 200; robots parsers read it as garbage rules.
No Content-Signal directives. Add e.g. `Content-Signal: ai-train=no, ai-summarize=yes` to declare granular AI usage policy beyond binary allow/disallow.
具体的な手順を見る
これは何を意味するか
Content-Signal is a newer robots.txt directive (pushed by Cloudflare) that lets you state a granular AI policy: for example "don’t train on my content, but summarising it in answers is fine". Classic robots.txt only offers all-or-nothing per crawler; this adds the middle ground. We looked for a Content-Signal line in your robots.txt and response headers and found none.
なぜ重要か
Without a granular signal, crawlers infer your intent from blunt allow/block rules, and sites often block more than they mean to just to avoid training use. A declared signal lets you keep answer visibility (summaries, citations) while opting out of what you object to. This is an emerging standard: honoring is voluntary and adoption is early, which is why it is flagged as emerging and weighs little.
やること
Decide your actual policy first: are you fine with AI training on your content? With summarisation in answers? With search indexing?
Express it as one line in robots.txt, for example: Content-Signal: ai-train=no, ai-summarize=yes.
Keep your existing User-agent rules; the signal adds nuance on top, it doesn’t replace them.
リポジトリを開いた Claude Code、Cursor などのアシスタントに貼り付けてください。この検出結果、手順、避けるべき近道が含まれます。
確認すること
Open yoursite.com/robots.txt and confirm the Content-Signal line is live.
Hit Re-scan above: "Content-Signal directives" flips to PASS.
避けること
Don’t declare signals that contradict your User-agent rules, such as ai-summarize=yes while blocking every AI crawler; conflicting instructions get you treated as unreliable.
Don’t add the line just for the score. It is a public policy statement; say what you mean, because compliant crawlers will act on it.