PluginBench
MCP Server
Maintained
Apache-2.0

ai.smithery/IlyaGusev-academia_mcp MCP Server

ai.smithery/IlyaGusev-academia_mcp

Search arXiv, ACL Anthology, and Hugging Face datasets; retrieve citations and analyze academic papers with web search and LLM tools.

What is the ai.smithery/IlyaGusev-academia_mcp MCP server?

The Academia MCP server provides tools to search, fetch, analyze, and report on scientific papers and datasets from arXiv, ACL Anthology, and Hugging Face. It integrates citation tracking via Semantic Scholar, web search capabilities, PDF/LaTeX processing, and optional LLM-powered document analysis for research workflows.

Academia MCP accelerates academic research by aggregating access to major paper repositories (arXiv, ACL Anthology), citation networks (Semantic Scholar), datasets (Hugging Face), and web search. It enables downloading papers, extracting text, answering questions over documents, generating research proposals, and compiling LaTeX—all from a single interface. Ideal for researchers, literature review automation, and AI-assisted paper analysis.

How to install ai.smithery/IlyaGusev-academia_mcp

Copy-paste configuration for popular MCP clients.

transport: http
Config generated by PluginBench — verify against the source before use.
Environment / auth
  • Authorization

    Bearer token for Smithery authentication

~/.cursor/mcp.json
{
  "mcpServers": {
    "IlyaGusev-academia_mcp": {
      "url": "https://server.smithery.ai/@IlyaGusev/academia_mcp/mcp"
    }
  }
}

Tools & capabilities

Tools this server exposes to the agent.

  • arxiv_search — Query arXiv with field-specific queries and filters.
  • arxiv_download — Fetch a paper by ID and convert to structured text (HTML/PDF modes).
  • anthology_search — Search ACL Anthology with fielded queries and optional date filtering.
  • hf_datasets_search — Find Hugging Face datasets with filters and sorting.
  • s2_get_citations — List papers citing a given arXiv paper (Semantic Scholar Graph).
  • s2_get_references — List papers referenced by a given arXiv paper.
  • visit_webpage — Fetch and normalize a web page.
  • web_search — Unified search wrapper; available when at least one of Exa/Brave/Tavily keys is set.
  • exa_web_search — Provider-specific web search via Exa.
  • brave_web_search — Provider-specific web search via Brave.
  • tavily_web_search — Provider-specific web search via Tavily.
  • get_latex_templates_list — Enumerate built-in LaTeX templates.
  • get_latex_template — Fetch a built-in LaTeX template.
  • compile_latex — Compile LaTeX to PDF in WORKSPACE_DIR.
  • read_pdf — Extract text per page from a PDF.
  • download_pdf_paper — Download a paper PDF (requires WORKSPACE_DIR).
  • review_pdf_paper — Download and optionally review PDFs with LLM (requires LLM + workspace).
  • document_qa — Answer questions over provided document chunks (requires LLM).
  • extract_bitflip_info — Research proposal helper (requires LLM).
  • generate_research_proposals — Generate research proposals (requires LLM).

Use cases

  • Search and download academic papers from arXiv and ACL Anthology, then extract and analyze their content.
  • Retrieve citation and reference networks for a paper to map research landscapes and identify key works.
  • Conduct web searches and fetch web pages to supplement academic research with current information.
  • Generate and score research proposals based on literature, with LLM-powered analysis.
  • Compile LaTeX documents to PDF and extract text from PDFs for document QA and automated summarization.

ai.smithery/IlyaGusev-academia_mcp MCP server FAQ

What is the Academia MCP server?

Academia MCP is an MCP server that provides tools to search, fetch, analyze, and report on scientific papers from arXiv, ACL Anthology, and Hugging Face datasets. It includes citation tracking, web search, PDF/LaTeX processing, and optional LLM-powered document analysis.

Is Academia MCP free to use?

The core server is free and open-source. However, some features require API keys: web search (Exa, Brave, or Tavily), LLM tools (OpenRouter), and Semantic Scholar access. These are optional and depend on your use case.

How do I install Academia MCP in Claude Desktop?

Install via pip (`pip3 install academia-mcp`), then add to Claude Desktop config: `{"mcpServers": {"academia": {"command": "python3", "args": ["-m", "academia_mcp", "--transport", "stdio"]}}}`

What authentication is required?

Authentication is optional and disabled by default. To enable token-based auth for HTTP transports, set `ENABLE_AUTH=true` and manage tokens via the `academia_mcp auth` CLI commands.

What environment variables do I need to set?

Core tools work without configuration. Optional: `OPENROUTER_API_KEY` (LLM tools), `EXA_API_KEY`/`BRAVE_API_KEY`/`TAVILY_API_KEY` (web search), `WORKSPACE_DIR` (LaTeX/PDF output), `PORT` (HTTP port, default 5056).

Can I run Academia MCP in Docker?

Yes. Build with `docker build -t academia_mcp .` or use the pre-built image `phoenix120/academia_mcp`. Run with environment variables for API keys and mount a volume for `WORKSPACE_DIR`.

README (reference)

Source of truth, from the repository.

Academia MCP

PyPI CI License smithery badge Verified on MseeP

MCP server with tools to search, fetch, analyze, and report on scientific papers and datasets.

Features

  • ArXiv search and download
  • ACL Anthology search
  • Hugging Face datasets search
  • Semantic Scholar citations and references
  • Web search via Exa, Brave, or Tavily
  • Web page crawler, LaTeX compilation, PDF reading
  • Optional LLM-powered tools for document QA and research proposal workflows

Requirements

  • Python 3.12+

Install

  • Using pip (end users):
pip3 install academia-mcp
  • For development (uv + Makefile):
uv venv .venv
make install

Quickstart

  • Run over HTTP (default transport):
python -m academia_mcp --transport streamable-http
# OR
uv run -m academia_mcp --transport streamable-http
  • Run over stdio (for local MCP clients like Claude Desktop):
python -m academia_mcp --transport stdio
# OR
uv run -m academia_mcp --transport stdio

Notes:

  • Transports: stdio, sse, streamable-http.
  • host/port are used for HTTP transports; ignored for stdio. Default port is 5056 (or PORT).

Authentication

Academia MCP supports optional token-based authentication for HTTP transports (streamable-http and sse). Authentication is disabled by default to maintain backward compatibility.

Enabling Authentication

Set the ENABLE_AUTH environment variable to true:

export ENABLE_AUTH=true
export TOKENS_FILE=/path/to/tokens.json  # Optional, defaults to ./tokens.json

Managing Tokens

Issue a new token:

academia_mcp auth issue-token --client-id=my-client --description="Production API client"

# Issue token with 30-day expiration
academia_mcp auth issue-token --client-id=test-client --expires-days=30

# Issue token with custom scopes
academia_mcp auth issue-token --client-id=admin --scopes="read,write,admin"

List active tokens:

academia_mcp auth list-tokens

Revoke a token:

academia_mcp auth revoke-token mcp_a1b2c3d4e5f6...

Using Tokens

Include the token in the Authorization header with the Bearer scheme or as a query parameter apiKey.

Security Notes:

  • Tokens are displayed only once during issuance. Store them securely.
  • Use HTTPS in production to protect tokens in transit.
  • The tokens.json file is automatically created with restrictive permissions (mode 600).
  • Tokens are stored in plaintext (standard practice for bearer tokens) - protect the tokens file.

Claude Desktop config

{
  "mcpServers": {
    "academia": {
      "command": "python3",
      "args": [
        "-m",
        "academia_mcp",
        "--transport",
        "stdio"
      ]
    }
  }
}

Available tools (one-liners)

  • arxiv_search: Query arXiv with field-specific queries and filters.
  • arxiv_download: Fetch a paper by ID and convert to structured text (HTML/PDF modes).
  • anthology_search: Search ACL Anthology with fielded queries and optional date filtering.
  • hf_datasets_search: Find Hugging Face datasets with filters and sorting.
  • s2_get_citations: List papers citing a given arXiv paper (Semantic Scholar Graph).
  • s2_get_references: List papers referenced by a given arXiv paper.
  • visit_webpage: Fetch and normalize a web page.
  • web_search: Unified search wrapper; available when at least one of Exa/Brave/Tavily keys is set.
  • exa_web_search, brave_web_search, tavily_web_search: Provider-specific search.
  • get_latex_templates_list, get_latex_template: Enumerate and fetch built-in LaTeX templates.
  • compile_latex: Compile LaTeX to PDF in WORKSPACE_DIR.
  • read_pdf: Extract text per page from a PDF.
  • download_pdf_paper, review_pdf_paper: Download and optionally review PDFs (requires LLM + workspace).
  • document_qa: Answer questions over provided document chunks (requires LLM).
  • extract_bitflip_info, generate_research_proposals, score_research_proposals: Research proposal helpers (requires LLM).

Availability notes:

  • Set WORKSPACE_DIR to enable compile_latex, read_pdf, download_pdf_paper, and review_pdf_paper.
  • Set OPENROUTER_API_KEY to enable LLM tools (document_qa, review_pdf_paper, and bitflip tools).
  • Set one or more of EXA_API_KEY, BRAVE_API_KEY, TAVILY_API_KEY to enable web_search and provider tools.

Environment variables

Set as needed, depending on which tools you use:

  • OPENROUTER_API_KEY: required for LLM-related tools.
  • BASE_URL: override OpenRouter base URL.
  • DOCUMENT_QA_MODEL_NAME: override default model for document_qa.
  • BITFLIP_MODEL_NAME: override default model for bitflip tools.
  • TAVILY_API_KEY: enables Tavily in web_search.
  • EXA_API_KEY: enables Exa in web_search and visit_webpage.
  • BRAVE_API_KEY: enables Brave in web_search.
  • WORKSPACE_DIR: directory for generated files (PDFs, temp artifacts).
  • PORT: HTTP port (default 5056).

You can put these in a .env file in the project root.

Docker

Build the image:

docker build -t academia_mcp .

Run the server (HTTP):

docker run --rm -p 5056:5056 \
  -e PORT=5056 \
  -e OPENROUTER_API_KEY=your_key_here \
  -e WORKSPACE_DIR=/workspace \
  -v "$PWD/workdir:/workspace" \
  academia_mcp

Or use existing image: phoenix120/academia_mcp

Examples

Makefile targets

  • make install: install the package in editable mode with uv
  • make validate: run black, flake8, and mypy (strict)
  • make test: run the test suite with pytest
  • make publish: build and publish using uv

LaTeX/PDF requirements

Only needed for LaTeX/PDF tools. Ensure a LaTeX distribution is installed and pdflatex is on PATH, as well as latexmk. On Debian/Ubuntu:

sudo apt install texlive-latex-base texlive-fonts-recommended texlive-latex-extra texlive-science latexmk

Related MCP servers

Automate cloud browsers to navigate websites, interact with elements, and extract structured data.…

0
Apache-2.0
View repository →

Generate professional PowerPoint presentations from text, YouTube videos, or structured JSON data.…

Generate polished PowerPoint presentations from text prompts, YouTube videos, or structured outlin…

Convert and compare dates and times across any timezone with flexible, locale-aware formatting. Ad…

0
TypeScript
MIT
View repository →

Discover festivals worldwide by location, date, and genre. Compare options with key details like d…

0
Python
View repository →

Craft quick, personalized greetings by name. Generate ready-to-use greeting prompts for a consiste…

0
Python
View repository →