Agent Tools / MCP install guides / Site Check in Pi

Site Check in Pi

Site SEO and landing checks. The MCP server is https://sitecheck.openkrill.app/mcp. No key. The steps below are only for Pi. A call to check_page_tags was checked against that server on 2026-10-01.

What this server answers

Look at what your website or online store tells crawlers, link previews and AI agents. Ask "can ChatGPT and Claude crawl example.com?", "is this product page set up for search?", "can AI shopping agents read this store?", "are any pages in this store's sitemap broken?" or "what would make this landing page convert better?".

Site Check fetches public pages and files and reports what it finds. For AI crawler access it evaluates robots.txt for fourteen published AI crawler names (GPTBot, OAI-SearchBot, ChatGPT-User, ClaudeBot, Claude-SearchBot, Claude-User, PerplexityBot, Google-Extended and others) against the page path, shows the rule that decided each one, and adds the X-Robots-Tag header and the robots meta tag of the page. For page tags it reads the title, description, canonical link, language and icons, plus the Open Graph and Twitter share tags, and lists the common ones that are missing. For redirects it follows the chain hop by hop and shows each status code, the final address and the main response headers.

For online stores, one product page can be checked for its title and description length, h1 heading, canonical link and Product structured data (name, brand, SKU, GTIN, MPN, price, currency, availability, variants). One domain can be checked for llms.txt, agents.md, the Universal Commerce Protocol file that Shopify stores publish, the sitemap and AI crawler rules in robots.txt. A sample of up to 40 addresses from the sitemap can be checked for broken pages and redirects. These work on any public store and add Shopify-specific fields where the store is on Shopify. They are not a speed test.

For landing pages, one page can be checked for the items a conversion review looks at, each with the evidence found and a fix: one clear h1 and a supporting sentence under it, a primary call to action that is near the top and repeated, whether the call-to-action links work, the number of form fields, trust signals (testimonials, customer logos, review markup, guarantees), a visible price or next step, the mobile viewport tag, HTML size and render-blocking hints, and a way to contact the business. The checklist topics follow the conversion-rate-optimization skill in coreyhaines31/marketingskills (MIT licence); the checks and wording are our own. Contact details, review text and page copy are not returned, only counts and the headline.

Only public websites on the standard web ports are checked. Private, local and internal addresses, other schemes such as file or ftp, and redirects into any of those are refused, and pages are fetched with a User-Agent that names Agent Tools. The store checks read robots.txt first and do not request a page that it disallows for Agent Tools. A check makes a handful of requests, and the sitemap sample or the landing page check at most 46.

It shows what the site publishes to a request from Agent Tools servers right now. It does not test whether a firewall or bot filter blocks real crawlers, does not run JavaScript, does not return page text, does not sign in to anything and does not scan for security vulnerabilities. robots.txt is a request that crawlers follow by choice, not access control.

It stores nothing about you or the site, and no fetched page is kept: answers are cached for five minutes, and the only lasting record is a daily count of calls per tool.

What it can do

Tools

Add Site Check in Pi

Pi does not speak MCP on its own. The pi-mcp-adapter extension (these steps match version 2.34.0) loads remote servers and keeps their schemas out of the prompt until you ask. Install it once, then restart Pi:

pi install npm:pi-mcp-adapter

Add the server to a shared MCP file. Use .mcp.json in the project if only that project needs Site Check, or ~/.config/mcp/mcp.json if you want it in every project. Pi also reads ~/.agents/mcp.json. A Pi-only override lives in ~/.pi/agent/mcp.json or .pi/mcp.json and wins over the shared file.

{
  "mcpServers": {
    "site-check": {
      "url": "https://sitecheck.openkrill.app/mcp"
    }
  }
}

The url is enough. Do not set command: that field is for a local process. No API key header. After restart, run /mcp and confirm the server is listed. The first call connects lazily. Ask: Can GPTBot and ClaudeBot crawl https://example.com/blog? The adapter should offer check_ai_crawler_access, check_page_tags, trace_redirects, product_page_seo, ai_readiness_check, link_sample_check, landing_page_check.

Leave directTools off unless you want those schemas in every turn. With a handful of tools the cost is small. The default proxy is the better fit: search for the tool, then call it. /mcp disable site-check writes a project override in .pi/mcp.json without editing the shared file. Run /reload after that change.

Pi will not find this server by scanning Cursor or Claude config unless you run /mcp setup and import it. Putting the url in .mcp.json is the direct path and the one this page tests against. The server is the same public endpoint the other clients use, with the same limits and the same privacy rules.

A call checked on 2026-10-01

example.com is a public page. The tool reads the head only. The request below was posted to https://sitecheck.openkrill.app/mcp as tools/call. HTTP 200. Source: the Site Check server, read 2026-10-01.

{
  "method": "tools/call",
  "params": {
    "name": "check_page_tags",
    "arguments": {
      "url": "https://example.com"
    }
  }
}

Repeat the call yourself if you need a newer reading. Cached answers expire. A rate limit is not a result: wait and try again. Nothing in the call is a ranking, a filing, or advice.

Same server, other clients

Roles that use Site Check

Terms used on this page

Source: plugin listing for site-check in this repository, and a live tools/call on 2026-10-01. As of 2026-10-01.