PluginBench
MCP Server
Active
Apache-2.0

io.github.houtini-ai/seo-crawler-mcp MCP Server

io.github.houtini-ai/seo-crawler-mcp

Crawl websites for SEO errors and store results in SQLite (now superseded by SEO Audit Console).

What is the io.github.houtini-ai/seo-crawler-mcp MCP server?

The SEO Crawler MCP server crawls websites and analyzes them for SEO errors using Crawlee, storing results in SQLite. This project is retired and has been superseded by SEO Audit Console, which adds search query data and ranks findings by potential click recovery rather than severity alone.

SEO Crawler crawled entire websites into a SQLite database and provided 25+ analysis queries to identify SEO issues. While the project is archived, the npm package remains available and functional. The crawler has been integrated into the newer SEO Audit Console tool, which combines crawl data with Google Search Console history and DataForSEO data for more actionable insights.

How to install io.github.houtini-ai/seo-crawler-mcp

Copy-paste configuration for popular MCP clients.

transport: stdio
Config generated by PluginBench — verify against the source before use.
Environment / auth
  • OUTPUT_DIR
    required

    Directory where crawl results are saved

  • DEBUG

    Enable verbose debug logging (set to 'true' to enable)

~/Library/Application Support/Claude/claude_desktop_config.json
{
  "mcpServers": {
    "seo-crawler-mcp": {
      "command": "npx",
      "args": [
        "-y",
        "@houtini/seo-crawler-mcp"
      ],
      "env": {
        "OUTPUT_DIR": "<YOUR_OUTPUT_DIR>",
        "DEBUG": "<YOUR_DEBUG>"
      }
    }
  }
}

Tools & capabilities

Tools this server exposes to the agent.

  • Website crawling — Crawl websites and extract SEO-relevant data into SQLite storage
  • SEO analysis queries — 25+ pre-built analysis queries to identify SEO errors and issues

Use cases

  • Identify broken links, missing metadata, and crawl errors across a website
  • Analyze on-page SEO factors like title tags, meta descriptions, and heading structure
  • Detect redirect chains and off-host redirect issues
  • Discover pages via sitemaps and known URLs for comprehensive site coverage
  • Export SEO audit results to SQLite for further analysis

io.github.houtini-ai/seo-crawler-mcp MCP server FAQ

What is the SEO Crawler MCP server?

It's a retired MCP server that crawled websites and analyzed them for SEO errors, storing results in SQLite. It has been superseded by SEO Audit Console, which adds search query data and ranks findings by potential click recovery.

Is this tool still maintained?

No, the project is archived and retired. The npm package (@houtini/seo-crawler-mcp) remains available and functional, but there are no further releases or updates.

How do I install it?

Install via npm with `npm install @houtini/seo-crawler-mcp`. However, the maintainers recommend migrating to SEO Audit Console for new projects.

What replaced this tool?

SEO Audit Console (github.com/houtini-ai/seo-audit) is the successor. It includes the crawler plus integration with Google Search Console and DataForSEO for ranking findings by potential click recovery.

Will my existing installation break?

No, the npm package is deprecated but not unpublished. Existing installations will continue to work without changes.

Does it require authentication?

The README does not specify authentication requirements. SEO Audit Console may require Google Search Console or DataForSEO credentials for full functionality.

README (reference)

Source of truth, from the repository.

<div align="center"> <img src="https://raw.githubusercontent.com/houtini-ai/seo-audit/master/assets/logo.png" width="120" height="120" alt="SEO Audit Console" /> </div>

SEO Crawler MCP - retired

This project has been superseded by SEO Audit Console.

SEO Crawler crawled your whole site into SQLite and gave Claude 25+ analysis queries over the result. The crawler survived the move - it was rewritten to be stdio-safe and it is now the crawl half of a bigger tool.

SEO Audit Console keeps the crawl and adds the thing the crawl could never tell you on its own: what people actually searched for to get there. It merges the crawl with your Google Search Console history and on-demand DataForSEO data, then ranks every finding by the clicks it could recover rather than by severity.

A flat crawler sells you severity. This ranks by yield.

Where to go

git clone https://github.com/houtini-ai/seo-audit.git

Docs and setup: github.com/houtini-ai/seo-audit

What this means for you

  • The npm package still installs. @houtini/seo-crawler-mcp is deprecated, not unpublished. Nothing you have running will break.
  • This repository is archived. The code stays readable and every existing link keeps working, but there are no further releases, and issues and pull requests are closed.
  • The crawler got better on the way over. It no longer depends on Crawlee (which logged to stdout and corrupted MCP's JSON-RPC framing), it seeds discovery from sitemaps and known GSC URLs rather than links alone, and it guards against redirects that wander off-host.

Migrating

Start a crawl on the new tool and it builds its own database:

start_crawl for https://example.com/

Run refresh_property instead if you want the crawl, the Search Console sync and URL inspection in one pass.


Part of the Houtini open-source MCP set. Questions: hello@houtini.com.

Related MCP servers

Amazon product search returning paste-ready affiliate deal rows for your blog

0
TypeScript
MIT
View repository →

Analyze content gaps and discover missing queries using AI search engine techniques

12
TypeScript
Apache-2.0
View repository →

Access financial data and market information via Financial Modeling Prep API

2
JavaScript
MIT
View repository →

Google Gemini image generation, video, and search-grounded chat inside Claude

30
TypeScript
Apache-2.0
View repository →

Use Grok as a peer code reviewer and second-opinion consultant inside Claude, Cursor, and Cline.

10
TypeScript
MIT
View repository →

MCP bridge for autonomous x402 payments in HPP USDC.e — discover and pay for services per call.

0
TypeScript
View repository →