Web Researcher Mcp

data-ai MCP Server

Give your AI assistant real web search, full-page reading, and multi-source research with citations that are never fabricated.

VerifiedFreshInstall Ready
data-aidata-ai
3 views61 stars9 forksMIT

Why This Matters

Discovered via github-topic:model-context-protocol and last synced 1d ago.

VerifiedFreshInstall Ready
Source
github-topic:model-context-protocol
Stars
61
Last synced
1d ago
Install
Instructions detected

Install

1. Install the package

uvx web-researcher-mcp

2. Add to claude_desktop_config.json

{
  "mcpServers": {
    "web-researcher-mcp": {
      "command": "uvx",
      "args": [
        "web-researcher-mcp"
      ]
    }
  }
}

Config file location: ~/Library/Application Support/Claude/claude_desktop_config.json (macOS) / %APPDATA%\Claude\claude_desktop_config.json (Windows)

80
Tools
0
Resources
0
Prompts
Standard I/O
Transport

Available Tools (80)

web

pages, PDFs, DOCX, PPTX, YouTube transcripts |

web-researcher-mcp

Perplexity

No

No

Partially

No (separate app)

Cost

**Free forever** (open source)

Tool

What it does

web_search

Search the web — optionally restricted to only the sources you trust via lenses

search_and_scrape

Search and then read the best results — with quality scoring to surface the most reliable sources

image_search

Find images by size, type, color, or format

news_search

Search recent news with date controls and source filtering

academic_search

Find real papers with real DOIs — authors, citation counts, open-access links

paper_fulltext

Fetch a paper's full text in one call from its DOI, Semantic Scholar ID, or URL — no need to chain `academic_search` then `scrape_page`

citation_graph

Walk a paper's citation neighborhood — works it cites and works that cite it, with intent/influence signals

patent_search

Search patent offices (US, Europe, international) with classification codes

filing_search

Search SEC EDGAR for US public-company filings (10-K, 10-Q, 8-K, …) — or pull structured XBRL company facts

legal_search

Search US court opinions and dockets via CourtListener — real cases with real citations

company_recon

OSINT company reconnaissance — Certificate Transparency log SANs, Wayback Machine historical URL inventory, derived subdomains, and a web-search company summary. Each phase fails soft and is independently selectable

research_export

Export a research session as a shareable report (markdown or JSON), with full per-step provenance

brand-guidelines

Research a brand and produce use-case-specific creative direction (landing page, email, video brief) — calls `brand_research` and interprets the structured JSON for you

searchapi

`SEARCHAPI_API_KEY`

biomed

Rare-disease and biomedical knowledge-graph sources — ontology portals, gene-disease databases, curated rare-disease registries

tech

Technology industry

econ_search

Look up economic data — World Bank global development indicators, OECD economic indicators, Eurostat European statistics (all keyless), and FRED US macro series (GDP, CPI, unemployment, rates; requires FRED_API_KEY)

verify_citation

Check a citation before you rely on it — does it exist, match a real record, and is it retracted or a dead link? Evidence, not a verdict

format_bibliography

Turn collected sources into a formatted bibliography — APA, MLA, BibTeX, RIS, or CSL-JSON (Zotero/EndNote/Mendeley-ready)

company-recon

Deep OSINT reconnaissance on a company — maps infrastructure, filings, personnel, and public footprint

SearXNG

`searxng`

News

Notes

clinical

Clinical trials, drug safety, evidence-based medicine

legal

Law, cases, statutes

clinical_search

Search ClinicalTrials.gov — clinical-trial registrations with status, phase, sponsor, and whether results are posted (discovery, not medical advice)

audit_bibliography

Audit a whole reference list in one pass — paste a CSL-JSON/RIS/BibTeX file (or a session) and get per-entry + corpus-level flags for retracted, dead-link, and unverifiable citations

research_panel

Ask the same question to a panel of independently configured LLMs and compare answers — consensus, contradictions, and model-unique points, computed deterministically, never smoothed over by an arbiter model

curriculum-research

Research a subject's syllabus coverage, institutional climate, and academic-freedom context — calls `web_search` with the `curriculum` lens

Tavily

`tavily`

Yes

Zero-config (public REST Search API); searches issues/PRs, not the full web

curriculum

Academic curriculum data, institutional free speech climate, and global education statistics

medical

Health, medicine

monarch_search

Query the Monarch Initiative biomedical knowledge graph — rank diseases and genes by phenotype similarity, look up disease/gene/phenotype entities, traverse gene-disease-phenotype associations

verify_recommendation

Audit an AI-generated recommendation list (listicle, product ranking) for self-promotion, author conflicts of interest, domain reputation, and dead links — catches GEO-gamed picks. Evidence, not a verdict

Template

What it guides your AI to do

Provider

Whole-Web

Exa

`exa`

investigative_records

Public records, corporate filings, FOIA

science

Research, papers

awesome_list_search

Search the ecosyste.ms Awesome API for community-curated "awesome-*" lists on a GitHub topic — structured, filterable coverage (stars, curated-entry count, topics) beyond free-text search

archive_source

Capture a fresh Internet Archive (Wayback Machine) snapshot of a URL via Save Page Now so a cited source stays verifiable if the page later changes or disappears — returns snapshot URL + timestamp (write tool)

comprehensive-research

Run a structured, multi-step deep dive on a topic

DuckDuckGo

`duckduckgo`

hackernews

none

Lens

Focus

programming

Code docs, tutorials, Q&A

government

Policy, regulations

local_search

Search for physical places (restaurants, shops, services, points of interest) by local intent query — structured POI details and descriptions. Requires `BRAVE_API_KEY`

sequential_search

Multi-step deep research — your AI remembers what it already found and builds on it

fact-check

Verify a claim against multiple independent sources

google

`GOOGLE_CUSTOM_SEARCH_API_KEY` + `GOOGLE_CUSTOM_SEARCH_ID`

Reddit

`reddit`

docs

Official documentation and API references only

programming-goggle

Developer-first results re-ranked by Brave's Programming Goggle — surfaces docs, repos, and authoritative technical content (requires Brave)

osint

Open-source intelligence — public records, corporate registries, social footprint, infrastructure

brand_research

Research a company's complete brand identity — colors (hex), logos, typography, tone of voice, and social handles — from any domain or company name. Returns structured JSON for AI content generation. No API key required; BrandFetch key optional for richer data

get_research_session

Recover a research session after context loss — picks up right where you left off

literature-review

Systematically review academic literature on a topic

Serper

`serper`

GitHub

`github`

academic-extended

Preprint servers, OA aggregators, and repositories beyond core journal indexes

news

Current events, journalism

Document

Description

competitive-analysis

Size up a company and its market (news, patents, web)

Brave

`brave`

Bluesky

`bluesky`

academic

Preprint servers, repositories, open-access journals

devops

Infrastructure and operations — Kubernetes, Docker, Terraform, cloud, CI/CD

awesome-lists

Community-curated "awesome-*" lists on GitHub — PR-reviewed tool and resource collections across every domain

security

CVEs, advisories, vulnerability research

finance

Markets, filings

scrape_page

Read any URL in full — web pages, PDFs, Word docs, slideshows, YouTube transcripts, Hacker News threads (read natively via the HN API); supports `mode: raw` for verbatim, unsanitized source (e.g. inspecting JSON or HTML)

Xquik

`xquik`

youcom

`YOUDOTCOM_API_KEY`