Documentation · v0.2.0

Everything you need to run Crawlwise.

Install it, understand the rubric, apply reversible fixes, and — if you want — hand the whole loop to an agent over MCP.

Last updated Aug 2026 · rubric-v1
01 · Getting started

Installation

Crawlwise requires WordPress 6.7+ and PHP 7.4+.

  1. 1 Upload Upload the crawlwise folder to /wp-content/plugins/, or install it through Plugins → Add New.
  2. 2 Activate Activate the plugin through the Plugins screen.
  3. 3 Scan Open Crawlwise in the admin menu and click Scan to grade your site.
  4. 4 Fix & verify Apply a fix from any failing check, then click Verify to confirm the change over a real HTTP request.

Activation creates one custom table ({prefix}_crawlwise_scans) for score history and flushes rewrite rules so /llms.txt, *.md and .well-known/ routes can resolve.

02 · Quick start

The three-beat loop

The core workflow is a three-beat loop you can repeat any time your site changes, producing an overall 0–100 score and a readiness level.

Scan

Crawlwise fetches your live URLs and grades five dimensions against the rubric.

Fix

Each failing check offers a one-click fix. Applying one writes plugin settings only — it never edits your theme or rewrites posts.

Verify

Re-fetch over a real request to confirm the score moved. A scheduled re-scan then watches for drift.

Because every check is an HTTP request against your own site, the score reflects what a crawler actually sees — including cases where a CDN or a physical robots.txt is overriding WordPress.

03 · Concepts

The rubric

Scoring is driven entirely by a versioned rubric (rubric-v1) — the single source of truth for what gets scored and how much it’s worth. Dimensions roll up to the overall score using a weighted average; info-tier results and inapplicable dimensions are excluded from the denominator.

Level Name Score Meaning
5 Agent-Native 90–100 Fully legible and actionable to agents.
4 Interactive 75–89 Capabilities & APIs are discoverable.
3 Governed 55–74 Explicit rules for which bots may do what.
2 Readable 30–54 Content available in machine-friendly form.
1 Discoverable 10–29 Basic robots & sitemap signals present.
0 Invisible 0–9 Agents can’t reliably find or read the site.

A new rubric version is a new file — the engine stays data-driven, so scoring rules can evolve without changing engine code.

04 · Concepts

Dimensions & checks

Fifteen checks across five dimensions. Each is graded independently over HTTP.

Discoverability 11 pts
D1 robots.txt present & sane Core 4 pts
D2 XML sitemap Core 4 pts
D3 Link header for discovery Core 3 pts
Content 18 pts
C1 Markdown content negotiation Core 8 pts
C2 /llms.txt Core 6 pts
C3 Structured HTML (headings, landmarks, FAQ, tables) Core 4 pts
Bot Access 15 pts
B1 Content Signals in robots.txt Core 5 pts
B2 Explicit AI bot rules Core 5 pts
B3 Web Bot Auth key directory Advanced 5 pts
Capabilities 12 pts
K1 API catalog Advanced 3 pts
K2 MCP card Advanced 3 pts
K3 Agent skills Advanced 3 pts
K4 OAuth metadata Advanced 3 pts
Commerce · Recommended only 10 pts
M1 UCP signals Advanced 5 pts
M2 x402 signals Advanced 5 pts
05 · Concepts

Fixes & reversibility

This is the part that sets Crawlwise apart from a stateless scanner: it doesn’t just tell you what’s wrong, it fixes it — safely.

Settings only Every fix writes plugin settings and nothing else. No theme files are edited, and no posts are rewritten.
Genuinely reversible Because state lives in settings, “Turn off” restores the exact prior behavior. Applying the C3 fix, for example, derives the meta description at render time rather than rewriting excerpts.
Non-destructive exports Static export refuses to overwrite a robots.txt or llms.txt it didn’t create, so a hand-maintained file is never clobbered.

Enabled static export and nothing happened? It won’t overwrite an existing robots.txt / llms.txt. Remove or rename the existing file if you want the plugin to manage it.

06 · Advanced

Agent (MCP) server

Crawlwise ships an optional MCP server that exposes scan, fix, verify and revert as tools an AI agent can call directly. It is off by default and requires a bearer token you generate yourself.

  1. 1 Enable Go to Crawlwise → Settings, tick Enable the MCP server, and click Save Changes.
  2. 2 Copy the token It is stored hashed and cannot be displayed again — regenerate it if you lose it.
  3. 3 Register with your agent Add the endpoint and bearer token to Claude Code, Cursor, or any MCP-capable client:
Terminal
# Register Crawlwise with your agent
$ claude mcp add --transport http crawlwise \
    https://your-site.com/wp-json/crawlwise/mcp \
    --header "Authorization: Bearer <token>"
✓ connected · tools: scan, fix, verify, revert

The token is stored as a SHA-256 hash, never in plain text, and is excluded from settings exports.

07 · Advanced

Privacy & security

Privacy

Crawlwise makes HTTP requests to your own site only. It has no telemetry — nothing is sent to us or any third party.

/llms.txt, the Markdown output and the JSON-LD graph only expose already-public content: password-protected, private and draft content is never included.

Security
The MCP bearer token is stored as a SHA-256 hash and cannot be recovered.
The token is never included in a settings export.
Failed MCP authentication attempts are rate-limited and rejected.
All dashboard REST routes require the manage_options capability.
08 · Reference

FAQ

Does it conflict with Yoast, Rank Math, or AIOSEO?

No. Crawlwise detects them and defers per feature — you’ll never get duplicate meta descriptions or two JSON-LD graphs.

Why does a check say “couldn’t verify”?

Some hosts block a site from making HTTP requests to itself. Crawlwise reports an informational result rather than guessing a misleading pass or fail.

I enabled static export and nothing happened.

Static export refuses to overwrite an existing robots.txt or llms.txt — your file is never clobbered. Remove or rename it if you want the plugin to manage that path.

What happens when I uninstall?

Uninstalling removes the custom table, all options and every emitted route. Your site is left exactly as Crawlwise found it.

09 · Reference

Changelog

Every release, newest first.

0.1.0 — initial release The scan → fix → verify loop across five dimensions. Password-protected content is excluded from Markdown output, /llms.txt and JSON-LD descriptions. The MCP bearer token is stored hashed, shown once, and excluded from settings export. *.md and .well-known/* URLs return a real 404 when a feature is off, and scan history is capped per scope.
Still stuck? Open a thread on WordPress.org or reach us through wpaiscanner.com — we usually answer within a day.
Docs for Crawlwise v0.2.0 · rubric-v1 — corrections welcome via the support forum. wpaiscanner.com/docs