io.github.houtini-ai/seo-crawler-mcp MCP Server
io.github.houtini-ai/seo-crawler-mcp
Crawl websites for SEO errors and store results in SQLite (now superseded by SEO Audit Console).
What is the io.github.houtini-ai/seo-crawler-mcp MCP server?
The SEO Crawler MCP server crawls websites and analyzes them for SEO errors using Crawlee, storing results in SQLite. This project is retired and has been superseded by SEO Audit Console, which adds search query data and ranks findings by potential click recovery rather than severity alone.
SEO Crawler crawled entire websites into a SQLite database and provided 25+ analysis queries to identify SEO issues. While the project is archived, the npm package remains available and functional. The crawler has been integrated into the newer SEO Audit Console tool, which combines crawl data with Google Search Console history and DataForSEO data for more actionable insights.
How to install io.github.houtini-ai/seo-crawler-mcp
Copy-paste configuration for popular MCP clients.
OUTPUT_DIRrequiredDirectory where crawl results are saved
DEBUGEnable verbose debug logging (set to 'true' to enable)
Tools & capabilities
Tools this server exposes to the agent.
Website crawling— Crawl websites and extract SEO-relevant data into SQLite storageSEO analysis queries— 25+ pre-built analysis queries to identify SEO errors and issues
Use cases
- Identify broken links, missing metadata, and crawl errors across a website
- Analyze on-page SEO factors like title tags, meta descriptions, and heading structure
- Detect redirect chains and off-host redirect issues
- Discover pages via sitemaps and known URLs for comprehensive site coverage
- Export SEO audit results to SQLite for further analysis
io.github.houtini-ai/seo-crawler-mcp MCP server FAQ
It's a retired MCP server that crawled websites and analyzed them for SEO errors, storing results in SQLite. It has been superseded by SEO Audit Console, which adds search query data and ranks findings by potential click recovery.
No, the project is archived and retired. The npm package (@houtini/seo-crawler-mcp) remains available and functional, but there are no further releases or updates.
Install via npm with `npm install @houtini/seo-crawler-mcp`. However, the maintainers recommend migrating to SEO Audit Console for new projects.
SEO Audit Console (github.com/houtini-ai/seo-audit) is the successor. It includes the crawler plus integration with Google Search Console and DataForSEO for ranking findings by potential click recovery.
No, the npm package is deprecated but not unpublished. Existing installations will continue to work without changes.
The README does not specify authentication requirements. SEO Audit Console may require Google Search Console or DataForSEO credentials for full functionality.
README (reference)
Source of truth, from the repository.
SEO Crawler MCP - retired
This project has been superseded by SEO Audit Console.
SEO Crawler crawled your whole site into SQLite and gave Claude 25+ analysis queries over the result. The crawler survived the move - it was rewritten to be stdio-safe and it is now the crawl half of a bigger tool.
SEO Audit Console keeps the crawl and adds the thing the crawl could never tell you on its own: what people actually searched for to get there. It merges the crawl with your Google Search Console history and on-demand DataForSEO data, then ranks every finding by the clicks it could recover rather than by severity.
A flat crawler sells you severity. This ranks by yield.
Where to go
git clone https://github.com/houtini-ai/seo-audit.git
Docs and setup: github.com/houtini-ai/seo-audit
What this means for you
- The npm package still installs.
@houtini/seo-crawler-mcpis deprecated, not unpublished. Nothing you have running will break. - This repository is archived. The code stays readable and every existing link keeps working, but there are no further releases, and issues and pull requests are closed.
- The crawler got better on the way over. It no longer depends on Crawlee (which logged to stdout and corrupted MCP's JSON-RPC framing), it seeds discovery from sitemaps and known GSC URLs rather than links alone, and it guards against redirects that wander off-host.
Migrating
Start a crawl on the new tool and it builds its own database:
start_crawl for https://example.com/
Run refresh_property instead if you want the crawl, the Search Console sync and URL inspection in one pass.
Part of the Houtini open-source MCP set. Questions: hello@houtini.com.
Related MCP servers
Amazon product search returning paste-ready affiliate deal rows for your blog

io.github.houtini-ai/fanout
Analyze content gaps and discover missing queries using AI search engine techniques

io.github.houtini-ai/fmp
Access financial data and market information via Financial Modeling Prep API

Google Gemini image generation, video, and search-grounded chat inside Claude

Use Grok as a peer code reviewer and second-opinion consultant inside Claude, Cursor, and Cline.

HPP x402 MCP Bridge
MCP bridge for autonomous x402 payments in HPP USDC.e — discover and pay for services per call.