Hub Semantic Search MCP
H

Hub Semantic Search MCP

An unofficial Hugging Face Hub semantic search MCP server that provides model and dataset search functionality based on natural language queries for MCP-compatible clients such as Claude.
2.5 points
6.5K

What is the Hugging Face Hub Semantic Search MCP Server?

This is a server based on the Model Context Protocol (MCP) that allows users to search for, discover, and explore models and datasets on Hugging Face through natural language queries. It provides semantic search functionality to help users find the desired content more accurately.

How to use the Hugging Face Hub Semantic Search MCP Server?

This server can be integrated with MCP-compatible clients such as Claude. Users only need to input natural language queries to perform searches. You can add the server to your client through the configuration file and directly use various search commands.

Use Cases

It is suitable for researchers, developers, and AI enthusiasts who want to quickly find models or datasets that meet specific requirements without delving into technical details.

Main Features

Semantic Search
Perform similarity searches based on AI-generated summaries rather than simple keyword matching to improve the relevance of search results.
Dataset Search
Search for datasets based on natural language descriptions to help users quickly find data suitable for their tasks.
Model Search
Support filtering models by the number of parameters to help users find suitable models.
Similar Content Recommendation
Recommend similar models or datasets based on specified content to help users expand their research scope.
Popular Content Display
Provide a list of currently popular models and datasets to help users stay informed about the latest trends.
Detailed Metadata Retrieval
Retrieve detailed information about models or datasets, including technical specifications and configurations.
Document Download
Download the README cards of models or datasets to obtain detailed instructions and usage methods.
Advantages
Improve search accuracy through semantic search and avoid the limitations of keyword matching.
Support multiple search methods, such as model search, dataset search, and similar content recommendation.
Provide detailed metadata and documentation to help users gain in-depth understanding of models or datasets.
Easy to integrate into existing MCP-compatible clients, such as Claude Desktop.
Limitations
It depends on the Hugging Face API, and adjustments may be required if the API changes.
Some advanced features may require specific configurations or permissions.
The interface and documentation may not be user-friendly for non-Chinese users.

How to Use

Install UV
Ensure that UV (a fast Python package manager) is installed to run the server.
Configure Claude Desktop
Add MCP server information to the configuration file of Claude Desktop so that the client can recognize and connect to the service.
Start the Server
Run the server to start processing search requests.

Usage Examples

Search for climate-related datasets
Users can easily find datasets related to climate change through natural language queries.
Find small language models
Users can quickly locate text generation models with less than 1 billion parameters through parameter filtering.
Find datasets similar to SQuAD
Users can find datasets similar to SQuAD for research on question-answering tasks.

Frequently Asked Questions

Is this MCP server free?
How to ensure the accuracy of search results?
Does it support multilingual search?
What should I do if I encounter problems?

Related Resources

Model Context Protocol GitHub
Official SDK and documentation for the Model Context Protocol.
Hugging Face Hub
Official documentation for the Hugging Face Hub, introducing the management of models and datasets.
Claude Desktop
Claude desktop application, a client that supports the MCP protocol.
UV Package Manager
UV is a fast Python package manager used to install and run the MCP server.

Installation

Copy the following command to your Client for configuration
{
  "mcpServers": {
    "huggingface-hub-search": {
      "command": "uvx",
      "args": [
        "git+https://github.com/davanstrien/hub-semantic-search-mcp.git"
      ],
      "env": {
        "HF_SEARCH_API_URL": "https://davanstrien-huggingface-datasets-search-v2.hf.space"
      }
    }
  }
}

{
  "mcpServers": {
    "huggingface-hub-search": {
      "command": "uv",
      "args": [
        "--directory",
        "/path/to/hub-semantic-search-mcp",
        "run",
        "python",
        "app.py"
      ],
      "env": {
        "HF_SEARCH_API_URL": "https://davanstrien-huggingface-datasets-search-v2.hf.space"
      }
    }
  }
}
Note: Your key is sensitive information, do not share it with anyone.

Alternatives

C
Claude Context
Claude Context is an MCP plugin that provides in - depth context of the entire codebase for AI programming assistants through semantic code search. It supports multiple embedding models and vector databases to achieve efficient code retrieval.
TypeScript
5.3K
5 points
A
Acemcp
Acemcp is an MCP server for codebase indexing and semantic search, supporting automatic incremental indexing, multi-encoding file processing, .gitignore integration, and a Web management interface, helping developers quickly search for and understand code context.
Python
10.2K
5 points
B
Blueprint MCP
Blueprint MCP is a chart generation tool based on the Arcade ecosystem. It uses technologies such as Nano Banana Pro to automatically generate visual charts such as architecture diagrams and flowcharts by analyzing codebases and system architectures, helping developers understand complex systems.
Python
8.4K
4 points
M
MCP Agent Mail
MCP Agent Mail is a mail - based coordination layer designed for AI programming agents, providing identity management, message sending and receiving, file reservation, and search functions, supporting asynchronous collaboration and conflict avoidance among multiple agents.
Python
8.6K
5 points
M
MCP
The Microsoft official MCP server provides search and access functions for the latest Microsoft technical documentation for AI assistants
12.2K
5 points
A
Aderyn
Aderyn is an open - source Solidity smart contract static analysis tool written in Rust, which helps developers and security researchers discover vulnerabilities in Solidity code. It supports Foundry and Hardhat projects, can generate reports in multiple formats, and provides a VSCode extension.
Rust
9.8K
5 points
D
Devtools Debugger MCP
The Node.js Debugger MCP server provides complete debugging capabilities based on the Chrome DevTools protocol, including breakpoint setting, stepping execution, variable inspection, and expression evaluation.
TypeScript
10.0K
4 points
S
Scrapling
Scrapling is an adaptive web scraping library that can automatically learn website changes and re - locate elements. It supports multiple scraping methods and AI integration, providing high - performance parsing and a developer - friendly experience.
Python
10.9K
5 points
M
Markdownify MCP
Markdownify is a multi-functional file conversion service that supports converting multiple formats such as PDFs, images, audio, and web page content into Markdown format.
TypeScript
27.7K
5 points
G
Gitlab MCP Server
Certified
The GitLab MCP server is a project based on the Model Context Protocol that provides a comprehensive toolset for interacting with GitLab accounts, including code review, merge request management, CI/CD configuration, and other functions.
TypeScript
18.7K
4.3 points
N
Notion Api MCP
Certified
A Python-based MCP Server that provides advanced to-do list management and content organization functions through the Notion API, enabling seamless integration between AI models and Notion.
Python
16.6K
4.5 points
D
Duckduckgo MCP Server
Certified
The DuckDuckGo Search MCP Server provides web search and content scraping services for LLMs such as Claude.
Python
55.9K
4.3 points
U
Unity
Certified
UnityMCP is a Unity editor plugin that implements the Model Context Protocol (MCP), providing seamless integration between Unity and AI assistants, including real - time state monitoring, remote command execution, and log functions.
C#
24.6K
5 points
F
Figma Context MCP
Framelink Figma MCP Server is a server that provides access to Figma design data for AI programming tools (such as Cursor). By simplifying the Figma API response, it helps AI more accurately achieve one - click conversion from design to code.
TypeScript
52.8K
4.5 points
G
Gmail MCP Server
A Gmail automatic authentication MCP server designed for Claude Desktop, supporting Gmail management through natural language interaction, including complete functions such as sending emails, label management, and batch operations.
TypeScript
17.4K
4.5 points
C
Context7
Context7 MCP is a service that provides real-time, version-specific documentation and code examples for AI programming assistants. It is directly integrated into prompts through the Model Context Protocol to solve the problem of LLMs using outdated information.
TypeScript
76.3K
4.7 points
AIBase
Zhiqi Future, Your AI Solution Think Tank
© 2025AIBase