Qdrant Retrieve

Qdrant Retrieve is an MCP server that enables AI assistants to execute semantic searches across vector collections stored in a Qdrant database instance. Developed for machine learning engineers, knowledge managers, and developers building retrieval-augmented generation pipelines, this server connects clients like Claude Desktop to local or remote Qdrant deployments. By integrating local embedding generation through models such as Xenova/all-MiniLM-L6-v2, it automatically converts search queries into vectors and retrieves relevant document chunks without requiring external embedding APIs. The server exposes a dedicated retrieval tool that accepts multiple text queries and searches across multiple collection names simultaneously, returning matching documents accompanied by similarity scores and collection identifiers. This functionality enables language models to locate domain knowledge, ground answers in enterprise reference materials, and retrieve contextual information dynamically during user conversations. It supports both standard input/output and HTTP transports, allowing flexible integration into diverse agentic workflows, local developer environments, and production retrieval architectures.

Category: AI Memory & Context

Tags: embeddings, qdrant, rag, semantic search, vector-database

Visit Qdrant Retrieve

How to install and configure Qdrant Retrieve

  1. Open your Claude Desktop configuration file (claude_desktop_config.json). 2. Add the server under the mcpServers object using npx: json { "mcpServers": { "qdrant": { "command": "npx", "args": ["-y", "@gergelyszerovay/mcp-server-qdrant-retrive"], "env": { "QDRANT_API_KEY": "your_api_key_here" } } } } 3. If connecting to a remote or non-default Qdrant instance, append --qdrantUrl=<url> to args (default is http://localhost:6333). 4. Save the configuration file and restart Claude Desktop. The initial search may take longer while the default embedding model downloads.

What you can do with Qdrant Retrieve

  • Searching multiple Qdrant document collections simultaneously to retrieve relevant reference context for user questions. - Running batch multi-query semantic searches to retrieve and synthesize contextual documents for RAG workflows. - Grounding AI assistant responses in technical documentation stored across distinct vector database collections. - Extracting scored text passages from local Qdrant collections to verify source accuracy without external embedding APIs.

Key facts

  • Open Source
  • https://github.com/gergelyszerovay/mcp-server-qdrant-retrive
  • AI Memory & Context, Databases & Data Stores
  • embeddings, qdrant, rag, semantic search, vector-database

Part of MCP Servers

Related MCP servers

  • MCP Memory Dashboard — MCP Memory Dashboard is an MCP server desktop interface that connects to the MCP Memory Service to provide visual semantic…
  • MCP Kanban Memory — MCP Kanban Memory is an MCP server that provides a kanban-based task management and state retention system for AI-driven workflows.…
  • MCP Memory Keeper — MCP Memory Keeper is an MCP server that provides persistent context management and memory storage for Claude AI coding assistants.…
  • MCP Knowledge Base — MCP Knowledge Base is an MCP server that processes local documents and answers queries based on their contents through similarity…
  • MCP Notes — MCP Notes is an MCP server that provides note-taking and note-management capabilities to AI models using Amazon Web Services DynamoDB…
  • MCP Neo4j Server — MCP Neo4j Server is an MCP server that provides an integration between the Neo4j graph database and Model Context Protocol…

What is Qdrant Retrieve?

Qdrant Retrieve is an open-source MCP server that provides semantic search capabilities across Qdrant vector database collections. It allows compatible AI clients to run vector similarity searches using local embedding models like Xenova/all-MiniLM-L6-v2, returning matched document text, source collection names, and similarity scores.

How do I install Qdrant Retrieve in Claude Desktop?

You can configure Qdrant Retrieve in Claude Desktop by editing your claude_desktop_config.json file. Add an entry under mcpServers running npx with the package @gergelyszerovay/mcp-server-qdrant-retrive. Optionally include your QDRANT_API_KEY in the env block and specify your database location using the --qdrantUrl argument.

What parameters does the qdrant_retrieve tool support?

The qdrant_retrieve tool accepts three primary parameters: collectionNames, which takes an array of Qdrant collection names to query; query, which takes an array of search strings; and topK, an optional number specifying how many top similar documents to return per query, defaulting to three.

Does Qdrant Retrieve require an external embedding API?

No, Qdrant Retrieve downloads and runs a local embedding model, defaulting to Xenova/all-MiniLM-L6-v2. Because the embeddings are generated locally by the server, an external API key for embedding providers is not required, though the initial search may take slightly longer while the model files download.

Is Qdrant Retrieve open source?

Yes, Qdrant Retrieve is an open-source project hosted on GitHub under the repository gergelyszerovay/mcp-server-qdrant-retrive. It can be executed directly using npx or cloned and customized locally to adjust embedding models, transport protocols, and database connection settings.

  • AI Tools
  • Categories
  • Industries
  • CLI Coding Agents
  • MCP Servers
  • MCP Categories