Search & Data Extraction MCP Servers
195 Model Context Protocol servers in the Search & Data Extraction category.
48 of 195 shown
pgalyen1987/gate402-mcp
github.comPay-per-call agent APIs over x402 (USDC on Base): per-token LLM inference (Llama 3.1/Qwen/Mistral), per-second GPU/CPU compute, on-chain/DeFi + SEC-EDGAR + news data, and clean-Markdown/Cloudflare-stealth web scraping. Signed receipts, no signup — auto-claims a free-tier key. Install via npx -y gate402-mcp.
KnockOutEZ/wigolo
github.comLocal-first, keyless web intelligence in one server: search, fetch, crawl, extract, cache, find-similar, and research. Multi-engine search with local ML reranking and a persistent SQLite cache, renders JS-heavy pages, and keeps everything on your machine. Install via npx wigolo init --non-interactive --agents=<agent>.
newsagentdata/newsagent-mcp
github.comReal-time, ML-enriched news intelligence — urgency scoring, political lean & event clustering across 190+ countries.
qinisolabs/icdwise
github.comVerified ICD-10-CM medical code lookup, validation & reverse search — official descriptions, never guessed.
AIweather-Anurag/ottasia-mcp-server
github.comWhere to watch any movie or TV show across 30 Asian and Middle Eastern streaming markets (Netflix, Disney+ Hotstar, Wavve, Shahid, Hoichoi, ZEE5, JioCinema, and 17 more). npx -y @ottasia/mcp-server
Alisammour/storyflo-mcp
github.comCurated audio-news with a market-aware news signal. Search articles, fetch narrated audio, subscribe topic feeds, surface stories matched to actively traded Kalshi event contracts (CFTC-regulated; qualitative signal tags + link-out to Kalshi, never raw market data). 8 tools (7 free + 1 x402-paid over USDC on Base). Install via npx -y storyflo-mcp.
mrslbt/rippr
github.comYouTube transcript extraction for AI agents. Clean text, timestamps, or structured JSON from any video. No API keys required. Install via npx rippr-mcp.
scavio-ai/scavio-mcp
github.comUnified real-time search API for AI agents. Google, YouTube, Amazon, Walmart, Reddit, and TikTok through one endpoint. 21 tools for web search, e-commerce, product data, social media, and video platforms. Free tier included.
0xdaef0f/job-searchoor
github.comAn MCP server for searching job listings with filters for date, keywords, remote work options, and more.
hanselhansel/aeo-cli
github.comAudit URLs for AI crawler readiness — checks robots.txt, llms.txt, JSON-LD schema, and content density with 0-100 AEO scoring.
Aas-ee/open-webSearch
github.comWeb search using free multi-engine search (NO API KEYS REQUIRED) — Supports Bing, Baidu, DuckDuckGo, Brave, Exa, and CSDN.
AceDataCloud/MCPSerp
github.comGoogle SERP search including web, images, news, maps, places, videos, and knowledge graph results via Ace Data Cloud API.
AIMLPM/markcrawl
github.comCrawl websites into clean Markdown, search pages, and extract structured data with LLMs. Built-in MCP server for web research and RAG pipelines.
ac3xx/mcp-servers-kagi
github.comKagi search API integration
adawalli/nexus
github.comAI-powered web search server using Perplexity Sonar models with source citations. Zero-install setup via NPX.
adjacentai/necl-hn-mcp
github.comHacker News tools for AI agents: top stories by time window, category feeds (top/new/best/ask/show/job), comment threads, and full-text search via Algolia. No API key required.
ananddtyagi/webpage-screenshot-mcp
github.comA MCP server for taking screenshots of webpages to use as feedback during UI developement.
andybrandt/mcp-simple-arxiv
github.com🐍 ☁️ MCP for LLM to search and read papers from arXiv
andybrandt/mcp-simple-pubmed
github.com🐍 ☁️ MCP to search and read medical / life sciences papers from PubMed.
andyliszewski/webcrawl-mcp
github.comLocal-first web scraping, search, and crawling. Static pages extracted locally via trafilatura; optional Firecrawl fallback only when JS rendering is needed. Four tools: scrape, search (DuckDuckGo), map, crawl.
angheljf/nyt
github.comSearch articles using the NYTimes API
apify/mcp-server-rag-web-browser
github.comAn MCP server for Apify's open-source RAG Web Browser Actor to perform web searches, scrape URLs, and return content in Markdown.
atlasprzetargow/mcp-server
github.comSearch 800 000+ Polish public tenders (BZP + TED). Profiles of procuring entities and contractors by NIP, market statistics by CPV/province, 90+ term procurement glossary.
AutomateLab-tech/citation-intelligence
github.comWhat LLMs cite, for agents. Check which URLs Perplexity, Claude, ChatGPT, Gemini, Bing, and Google AI Overviews cite for any query. Self-hosted, BYO API key. Install via npx @automatelab/citation-intelligence.
Khamel83/argus
github.comMulti-provider search broker with automatic fallback, RRF ranking, content extraction, and budget enforcement.
idapixl/idapixl-web-research-mcp
github.comPay-per-use web research for AI agents on Apify. Search (Brave + DuckDuckGo), fetch pages to clean markdown, and multi-step research with relevance scoring and key fact extraction.
Bigsy/Clojars-MCP-Server
github.comClojars MCP Server for upto date dependency information of Clojure libraries
blazickjp/arxiv-mcp-server
github.comSearch ArXiv research papers
boikot-xyz/boikot
github.comModel Context Protocol Server for looking up company ethics information. Learn about the ethical and unethical actions of major companies.
brave/brave-search-mcp-server
github.comWeb search capabilities using Brave's Search API
cameronrye/activitypub-mcp
github.comA comprehensive MCP server that enables LLMs to explore and interact with the Fediverse through ActivityPub protocol. Features WebFinger discovery, timeline fetching, instance exploration, and cross-platform support for Mastodon, Pleroma, Misskey, and other ActivityPub servers.
cameronrye/gopher-mcp
github.comModern, cross-platform MCP server enabling AI assistants to browse and interact with both Gopher protocol and Gemini protocol resources safely and efficiently. Features dual protocol support, TLS security, and structured content extraction.
capad-xyz/searchts
github.comKeyless web access for AI agents: an escalating open-source unlocker (browser-fingerprint fetch → JS-render relay → stealth browser) reads bot-walled pages as clean Markdown, plus multi-provider web search with rank fusion, subtitles-first video transcripts, and page asset grabbing. No API keys; ships a reproducible benchmark.
einiba/canyougrab-api
github.comConfidence-scored domain availability checking with real-time DNS + WHOIS lookups. Bulk check up to 100 domains per request. Each result includes availability, confidence level, data source, and registration details.
cevatkerim/unsplash-mcp
github.comUnsplash photo search with proper attribution. Returns ready-to-use attribution text and HTML for each photo, making it easy for LLMs to build content pages with properly credited images. Includes search, random photos, and download tracking.
chanmeng/google-news-mcp-server
github.comGoogle News integration with automatic topic categorization, multi-language support, and comprehensive search capabilities including headlines, stories, and related topics through SerpAPI.
chasesaurabh/mcp-page-capture
github.comMCP server that captures webpage screenshots, with viewport or full-page options and base64 PNG output.
comparedge/mcp-server-comparedge
github.comVerified SaaS, AI, and LLM pricing for 490+ tools: plans, hidden costs, alternatives, and comparisons. Free, no API key.
CKBrennan/overtone-news-mcp
github.comReal-time news with tone analysis, brand safety, and narrative shift signals for AI agents.
ConechoAI/openai-websearch-mcp
github.comThis is a Python-based MCP server that provides OpenAI web_search built-in tool.
Crawleo/Crawleo-MCP
github.comCrawleo Search & Crawl API
Crawlora-org/crawlora-mcp
github.comHosted MCP for structured public web data — 319 tools across search, maps, commerce, social, and finance, each returning clean JSON. Free 2,000 credits/mo.
czottmann/kagi-ken-mcp
github.comWork with Kagi without API access (you'll need to be a customer, tho). Searches and summarizes. Uses Kagi session token for easy authentication.
DappierAI/dappier-mcp
github.comEnable fast, free real-time web search and access premium data from trusted media brands—news, financial markets, sports, entertainment, weather, and more. Build powerful AI agents with Dappier.
deadletterq/mcp-opennutrition
github.comLocal MCP server for searching 300,000+ foods, nutrition facts, and barcodes from the OpenNutrition database.
dealx/mcp-server
github.comMCP Server for DealX platform
deficlow/HyperStore-MCP
github.comSearch 6,500+ curated AI applications from the HyperStore directory. 8 tools (keyword + semantic search, full details, browsing), 3 resources, 3 prompts. Install via uvx hyperstore-mcp or use the hosted endpoint at https://mcp.store.hypergpt.ai/mcp.
★devflowinc/trieve
github.comCrawl, embed, chunk, search, and retrieve information from datasets through Trieve
Attribution
Data sourced from punkpeye/awesome-mcp-servers (MIT). Synced every 24 hours.