io.github.thunderbit-open/thunderbit-mcp-server MCP Server
io.github.thunderbit-open/thunderbit-mcp-server
AI-powered web scraping with Markdown distillation and JSON Schema extraction.
What is the io.github.thunderbit-open/thunderbit-mcp-server MCP server?
The Thunderbit MCP server is an AI-powered web scraping tool that converts web pages into Markdown format or extracts structured data using JSON Schema. It enables AI agents to efficiently gather and transform web content for analysis and integration.
This server provides web scraping capabilities optimized for AI workflows, allowing you to distill web pages into clean Markdown or extract structured data according to JSON Schema specifications. It's useful for content aggregation, data extraction, and feeding web information into AI pipelines.
How to install io.github.thunderbit-open/thunderbit-mcp-server
Copy-paste configuration for popular MCP clients.
THUNDERBIT_API_KEYrequiredsecretYour Thunderbit API key (tb_...). Get one at https://app.thunderbit.com/console
Tools & capabilities
Tools this server exposes to the agent.
Web page to Markdown conversion— Convert web pages into clean, readable Markdown formatJSON Schema-based data extraction— Extract structured data from web pages using JSON Schema definitions
Use cases
- Extract structured data from websites using JSON Schema definitions
- Convert web pages to Markdown for AI processing and analysis
- Aggregate content from multiple web sources into a unified format
- Automate data collection workflows for research or business intelligence
- Feed web content into AI pipelines for further processing
io.github.thunderbit-open/thunderbit-mcp-server MCP server FAQ
It's an AI-powered web scraping tool that distills web pages to Markdown or extracts structured data via JSON Schema, designed for integration with AI agents.
The README does not specify pricing; check the GitHub repository for licensing details.
Install via npm with: npm install @thunderbit/mcp-server, then configure it in your MCP client (Cursor, Claude, etc.).
The README does not specify authentication requirements; refer to the GitHub repository for setup details.
Yes, it can scrape and extract data from web pages, though you should respect robots.txt and terms of service.
It supports Markdown output for page distillation and JSON output for structured data extraction via JSON Schema.
Related MCP servers
Build, test, deploy and run AI phone agents: agents, numbers, calls, tests, knowledge, campaigns.

io.github.thuupx/lunge
Agent-native API client: execute/test REST/GraphQL/WS/SSE with assertions, extraction, collections

io.github.thuupx/memory-mcp-lite
A lightweight, structured, token-efficient local-first MCP memory server

Algenta MCP Server
Governed data discovery, exact queries, decisions, simulations, and runtime utilities over MCP.
Repo intelligence: triage, fix PRs, SARIF reachability, recall. Free codna login to execute.
View repository →SQAI MCP server — four governed, deterministic, read-only data tools. Free login to execute.