Web fetch and search MCP Server

Web fetch and search MCP Server is an MCP server that provides search and webpage retrieval tools using an asynchronous OCaml implementation built on the Eio runtime. It connects AI assistants directly to external sources including DuckDuckGo and Wikipedia, while also allowing agents to retrieve and clean HTML content from arbitrary URLs into plain text or structured Markdown. AI application developers and users of desktop LLM clients employ this server to grant models real-time research and web-browsing capabilities. The server implements four primary tools: standard web queries via DuckDuckGo, encyclopedic lookups via Wikipedia, raw webpage content extraction with configurable offsets and byte limits, and Markdown page parsing that integrates with Trafilatura or Jina Reader. Built-in rate limiting restricts search traffic to thirty requests per minute and page fetching to twenty requests per minute to prevent upstream service throttling. Users can run the binary over standard input and output streams for local clients or expose it over HTTP on a dedicated network port for remote integrations.

Category: Browser & Web Automation

Tags: ocaml, web scraping, web search, wikipedia

Visit Web fetch and search MCP Server

How to install and configure Web fetch and search MCP Server

  1. Clone the repository and enter the directory: bash cd snf_mcp 2. Install dependencies and compile the binary with Opam and Dune: bash opam install . --deps-only dune build dune install 3. (Optional) Install trafilatura for higher quality Markdown extraction: bash pip install trafilatura 4. Add the server configuration to your MCP client (such as ~/.llm-tools-mcp/mcp.json, LMStudio, or Jan): json { "mcpServers": { "snf_mcp": { "command": "/path/to/snf-mcp", "args": [ "--stdio" ] } } } 5. Alternatively, start the server in HTTP mode using dune exec snf-mcp -- --serve 3000 to serve requests over a network port.

What you can do with Web fetch and search MCP Server

  • Query DuckDuckGo to obtain relevant search results and summary snippets for up-to-date programming language documentation and technical topics. - Search Wikipedia articles directly from an AI chat interface to quickly retrieve encyclopedic context and factual overviews. - Fetch cleaned text from specific website URLs using configurable byte lengths and starting offsets to ingest targeted web documentation. - Extract parsed Markdown from articles using Trafilatura or Jina Reader to preserve document structure for prompt ingestion. - Run a local or networked MCP endpoint providing asynchronous, rate-limited web browsing tools to desktop LLM interfaces like Jan or LMStudio.

Key facts

  • https://github.com/mseri/snf-mcp
  • Browser & Web Automation, Web Search & Research
  • ocaml, web scraping, web search, wikipedia

Part of MCP Servers

Related MCP servers

  • MCP JSON — MCP JSON is an MCP server collection that bundles tools for file system operations, Google search, browser-based web automation, and…
  • MCP NPX Fetch — MCP NPX Fetch is an MCP server that retrieves online resources and transforms web content into structured formats including HTML,…
  • MCP Naver News — MCP Naver News is an MCP server that connects AI assistants to the Naver News API, enabling automated search and…
  • MCP NIF.PT — MCP NIF.PT is an MCP server that connects LLM clients to the Portuguese NIF.PT public API to retrieve and analyze…
  • MCP Node Fetch — MCP Node Fetch is an MCP server that enables language model assistants to retrieve web content and query remote endpoints…
  • MCP Open Library — MCP Open Library is an MCP server that connects AI assistants to the Open Library catalogue API to retrieve book,…

What tools does Web fetch and search MCP Server provide?

The server provides four tools: search for querying DuckDuckGo, search_wikipedia for locating Wikipedia articles, fetch_content for reading raw or cleaned webpage text, and fetch_markdown for retrieving formatted Markdown from web pages using Trafilatura or Jina Reader.

How do I install and build the server?

The server is written in OCaml. You clone the repository, install required dependencies using opam install . --deps-only, and compile and install the binary with dune build and dune install. The executable can then be run directly from your PATH.

Which MCP clients work with Web fetch and search MCP Server?

Any client supporting the Model Context Protocol can connect to this server. The documentation provides explicit setup instructions for desktop and command-line LLM tools such as LMStudio, Jan, and the llm CLI using the llm-tools-mcp plugin via standard input and output mode. It also supports remote HTTP connections.

Does the server enforce rate limits on requests?

Yes, the server includes built-in rate limits to prevent external services from throttling requests. Web search requests to DuckDuckGo and Wikipedia are restricted to thirty calls per minute, while content fetching is limited to twenty requests per minute.

Why is trafilatura recommended for fetch_markdown?

When parsing webpages into Markdown, the fetch_markdown tool attempts to use the trafilatura Python library if installed on your system because it produces higher quality article extraction. If trafilatura is not installed, the server automatically falls back to Jina Reader.

  • AI Tools
  • Categories
  • Industries
  • CLI Coding Agents
  • MCP Servers
  • MCP Categories