GetWeb
БесплатноНе проверенIntegrates DuckDuckGo, Google Search, Felo AI, and Jina Reader APIs to provide web search, content extraction, and HTML-to-Markdown conversion with caching, use
Описание
Integrates DuckDuckGo, Google Search, Felo AI, and Jina Reader APIs to provide web search, content extraction, and HTML-to-Markdown conversion with caching, user agent rotation, and configurable text filtering for reliable web research and information retrieval.
README
npm version npm downloads License: MIT GitHub issues
A Model Context Protocol (MCP) server that provides web search and content extraction capabilities.
Quick Start
{
"mcpServers": {
"getweb": {
"command": "npx",
"args": [
"mcp-getweb"
],
"type": "stdio",
"env": {
"GOOGLE_API_KEY": "XXXXXXXXX",
"GOOGLE_SEARCH_ENGINE_ID": "XXXXXXXXX",
"JINA_API_KEY": "jina_XXXXXXXXX",
"LINKUP_API_KEY": "XXXXXXXXX",
"EXA_API_KEY": "XXXXXXXXX"
}
}
}
}
Features
1) DuckDuckGo Search (duckduckgo-search)
Search the web using DuckDuckGo with HTML scraping.
Parameters:
query(string, required): The search querypage(integer, optional): Page number (default: 1, min: 1)numResults(integer, optional): Number of results to return (default: 10, min: 1, max: 20)
2) Google Search (google-search)
Search Google and return relevant results using the Programmable Search Engine.
Parameters:
query(string, required): Search query; quotes enable exact matchesnum_results(integer, optional): Total results to return (default: 5, max: 10)site(string, optional): Restrict to a specific site/domain (e.g.,wikipedia.org)language(string, optional): ISO 639-1 language code (e.g.,en,es)dateRestrict(string, optional): Date filter, e.g.,d7,w4,m6,y1exactTerms(string, optional): Exact phrase that must appearresultType(string, optional): Result type:image|images|news|video|videospage(integer, optional): Page number for pagination (default: 1, min: 1)resultsPerPage(integer, optional): Results per page (default: 5, max: 10)sort(string, optional): Sort order,relevance(default) ordate
Note: Requires GOOGLE_API_KEY and GOOGLE_SEARCH_ENGINE_ID to be set.
3) Linkup Search (linkup_search)
Search the web via Linkup API and return relevant results in Markdown.
Parameters:
query(string, required): Natural-language search queryonlySearchTheseDomains(array of strings, optional): Restrict results to specific domainsdateFilter(object, optional): Date range filterfromDate(string, optional): Start date inYYYY-MM-DDtoDate(string, optional): End date inYYYY-MM-DD
maxResults(integer, optional): Maximum number of results to return (default: 5, min: 1)
Note: Requires LINKUP_API_KEY to be set.
4) Exa Search (exa_search)
Search the web via Exa API and return relevant results in Markdown.
Parameters:
query(string, required): Natural-language search querymaxResults(integer, optional): Number of results to return (default: 10, max: 25)publishedDateRange(object, optional): Published date range filterfromDate(string, optional): Start date, RFC3339 (e.g.,2024-02-09T00:00:00.000Z) orYYYY-MM-DDtoDate(string, optional): End date, RFC3339 (e.g.,2024-02-09T00:00:00.000Z) orYYYY-MM-DD
crawlDateRange(object, optional): Crawl date range filterfromDate(string, optional): Start date, RFC3339 orYYYY-MM-DDtoDate(string, optional): End date, RFC3339 orYYYY-MM-DD
userLocation(string, optional): Two-letter ISO country code (e.g.,US)includeText(string, optional): Exact phrase that must appear in the webpage text (max 5 words)excludeText(string, optional): Exact phrase that must not appear in the webpage text (max 5 words)domain(string, optional): Restrict results to a single domain (e.g.,arxiv.org)
Note: Requires EXA_API_KEY to be set.
5) Felo AI Search (felo-search)
AI-powered search with contextual responses for up-to-date technical information (releases, advisories, migrations, benchmarks, community insights).
Parameters:
query(string, required): The search query or promptstream(boolean, optional): Whether to stream the response (default: false)
6) URL Content Fetcher (fetch-url)
Fetch the clean content of a URL and return it as text.
Parameters:
url(string, required): The URL to fetchmaxLength(integer, optional): Maximum content length (default: 30000, min: 1000, max: 500000)extractMainContent(boolean, optional): Attempt to extract main content when HTML (default: true)
7) URL Metadata Extractor (url-metadata)
Extract metadata (title, description, image, favicon) from a URL.
Parameters:
url(string, required): The URL to extract metadata from
8) URL Fetch to Markdown (url-fetch)
Fetch web pages and convert them to Markdown. Handles HTML, plaintext, and JSON (pretty-printed in a fenced block).
Parameters:
url(string, required): The URL to fetch and convert to Markdown
9) Jina Reader (jina-reader)
Retrieve LLM-friendly content from a URL using Jina r.reader with optional summaries and formats.
Parameters:
url(string, required): The URL to fetch and parsemaxLength(integer, optional): Maximum output length (default: 10000, min: 1000, max: 50000)withLinksummary(boolean, optional): Include links summary (default: false)withImagesSummary(boolean, optional): Include images summary (default: false)withGeneratedAlt(boolean, optional): Generate alt text for images (default: false)returnFormat(string, optional):markdown(default) |html|text|screenshot|pageshotnoCache(boolean, optional): Bypass cache (default: false)timeout(integer, optional): Max seconds to wait (default: 10, min: 5, max: 30)
Note: Requires JINA_API_KEY to be set.
Acknowledgments
- Model Context Protocol specification by Anthropic
- DuckDuckGo for providing a privacy-focused web search experience
- Google Programmable Search Engine and Custom Search JSON API
- Linkup API for high-quality web search results
- Exa API for fast neural web search
- Jina AI r.reader API for high-quality content extraction
- Felo AI for up-to-date, developer-focused search insights
- Rust ecosystem and crates that power this server:
- tokio, reqwest, serde, serde_json, tracing, tracing-subscriber, clap
- html2text, chardetng, encoding_rs, scraper, html5ever, markup5ever_rcdom, regex, once_cell, futures, async-stream
- url, uuid, thiserror, tokio-util, rand, urlencoding
- The broader MCP community for guidance, examples, and discussions
Support
If you encounter any issues or have questions, please open an issue on GitHub.
Установить GetWeb в Claude Desktop, Claude Code, Cursor
unyly install getwebСтавит в Claude Desktop, Claude Code, Cursor и VS Code — сам разбирается с npx, uvx и сборкой из исходников.
Впервые? Поставь CLI: curl -fsSL https://unyly.org/install | sh
Или настроить вручную
Выполни в терминале:
claude mcp add getweb -- npx -y mcp-getwebПошаговые гайды: как установить GetWeb
FAQ
GetWeb MCP бесплатный?
Да, GetWeb MCP бесплатный — установка в пару кликов через Unyly без оплаты.
Нужен ли API-ключ для GetWeb?
Нет, GetWeb работает без API-ключей и переменных окружения.
GetWeb — hosted или self-hosted?
Self-hosted: сервер запускается локально на твоей машине командой из раздела установки.
Как установить GetWeb в Claude Desktop, Claude Code или Cursor?
Открой GetWeb на unyly.org, выбери вкладку своего клиента (Claude Desktop, Claude Code, Cursor) и нажми Install — конфиг сгенерируется автоматически, без правки JSON.
Похожие MCP
Fetch
Web content fetching and conversion for efficient LLM usage.
AWS KB Retrieval
Retrieval from AWS Knowledge Base using Bedrock Agent Runtime.
автор: modelcontextprotocolSpring AI MCP Server
Provides auto-configuration for setting up an MCP server in Spring Boot applications.
llm-analysis-assistant
A very streamlined mcp client that supports calling and monitoring stdio/sse/streamableHttp, and can also view request responses through the /logs page. It also
автор: xuzexin-hzCompare GetWeb with
Не уверен что выбрать?
Найди свой стек за 60 секунд
Автор?
Embed-бейдж для README
Похожее
Все в категории ai
