I craft unique cereal names, stories, and ridiculously cute Cereal Baby images.

visión del conjunto de datos
Servidor MCP para abrazar el espectador de conjuntos de datos
1
Github Watches
7
Github Forks
12
Github Stars
Dataset Viewer MCP Server
An MCP server for interacting with the Hugging Face Dataset Viewer API, providing capabilities to browse and analyze datasets hosted on the Hugging Face Hub.
Features
Resources
- Uses
dataset://
URI scheme for accessing Hugging Face datasets - Supports dataset configurations and splits
- Provides paginated access to dataset contents
- Handles authentication for private datasets
- Supports searching and filtering dataset contents
- Provides dataset statistics and analysis
Tools
The server provides the following tools:
-
validate
- Check if a dataset exists and is accessible
- Parameters:
-
dataset
: Dataset identifier (e.g. 'stanfordnlp/imdb') -
auth_token
(optional): For private datasets
-
-
get_info
- Get detailed information about a dataset
- Parameters:
-
dataset
: Dataset identifier -
auth_token
(optional): For private datasets
-
-
get_rows
- Get paginated contents of a dataset
- Parameters:
-
dataset
: Dataset identifier -
config
: Configuration name -
split
: Split name -
page
(optional): Page number (0-based) -
auth_token
(optional): For private datasets
-
-
get_first_rows
- Get first rows from a dataset split
- Parameters:
-
dataset
: Dataset identifier -
config
: Configuration name -
split
: Split name -
auth_token
(optional): For private datasets
-
-
get_statistics
- Get statistics about a dataset split
- Parameters:
-
dataset
: Dataset identifier -
config
: Configuration name -
split
: Split name -
auth_token
(optional): For private datasets
-
-
search_dataset
- Search for text within a dataset
- Parameters:
-
dataset
: Dataset identifier -
config
: Configuration name -
split
: Split name -
query
: Text to search for -
auth_token
(optional): For private datasets
-
-
filter
- Filter rows using SQL-like conditions
- Parameters:
-
dataset
: Dataset identifier -
config
: Configuration name -
split
: Split name -
where
: SQL WHERE clause (e.g. "score > 0.5") -
orderby
(optional): SQL ORDER BY clause -
page
(optional): Page number (0-based) -
auth_token
(optional): For private datasets
-
-
get_parquet
- Download entire dataset in Parquet format
- Parameters:
-
dataset
: Dataset identifier -
auth_token
(optional): For private datasets
-
Installation
Prerequisites
- Python 3.12 or higher
- uv - Fast Python package installer and resolver
Setup
- Clone the repository:
git clone https://github.com/privetin/dataset-viewer.git
cd dataset-viewer
- Create a virtual environment and install:
# Create virtual environment
uv venv
# Activate virtual environment
# On Unix:
source .venv/bin/activate
# On Windows:
.venv\Scripts\activate
# Install in development mode
uv add -e .
Configuration
Environment Variables
-
HUGGINGFACE_TOKEN
: Your Hugging Face API token for accessing private datasets
Claude Desktop Integration
Add the following to your Claude Desktop config file:
On Windows: %APPDATA%\Claude\claude_desktop_config.json
On MacOS: ~/Library/Application Support/Claude/claude_desktop_config.json
{
"mcpServers": {
"dataset-viewer": {
"command": "uv",
"args": [
"run",
"dataset-viewer"
]
}
}
}
Usage Examples
- Validate a dataset:
{
"dataset": "stanfordnlp/imdb"
}
- Get dataset information:
{
"dataset": "stanfordnlp/imdb"
}
- Search dataset contents:
{
"dataset": "stanfordnlp/imdb",
"config": "plain_text",
"split": "train",
"query": "great movie"
}
- Filter and sort rows:
{
"dataset": "stanfordnlp/imdb",
"config": "plain_text",
"split": "train",
"where": "label = 'positive'",
"orderby": "text DESC",
"page": 0
}
- Get dataset statistics:
{
"dataset": "stanfordnlp/imdb",
"config": "plain_text",
"split": "train"
}
License
MIT License - see LICENSE for details
相关推荐
I find academic articles and books for research and literature reviews.
Evaluator for marketplace product descriptions, checks for relevancy and keyword stuffing.
Confidential guide on numerology and astrology, based of GG33 Public information
Emulating Dr. Jordan B. Peterson's style in providing life advice and insights.
Your go-to expert in the Rust ecosystem, specializing in precise code interpretation, up-to-date crate version checking, and in-depth source code analysis. I offer accurate, context-aware insights for all your Rust programming questions.
Advanced software engineer GPT that excels through nailing the basics.
Converts Figma frames into front-end code for various mobile frameworks.
This GPT assists in finding a top-rated business CPA - local or virtual. We account for their qualifications, experience, testimonials and reviews. Business operators provide a short description of your business, services wanted, and city or state.
Descubra la colección más completa y actualizada de servidores MCP en el mercado. Este repositorio sirve como un centro centralizado, que ofrece un extenso catálogo de servidores MCP de código abierto y propietarios, completos con características, enlaces de documentación y colaboradores.
La aplicación AI de escritorio todo en uno y Docker con trapo incorporado, agentes de IA, creador de agentes sin código, compatibilidad de MCP y más.
Manipulación basada en Micrypthon I2C del expansor GPIO de la serie MCP, derivada de AdaFruit_MCP230xx
Plataforma de automatización de flujo de trabajo de código justo con capacidades de IA nativas. Combine el edificio visual con código personalizado, auto-anfitrión o nube, más de 400 integraciones.
Espejo dehttps: //github.com/agentience/practices_mcp_server
🧑🚀 全世界最好的 llM 资料总结(数据处理、模型训练、模型部署、 O1 模型、 MCP 、小语言模型、视觉语言模型) | Resumen de los mejores recursos del mundo.
Una lista curada de servidores de protocolo de contexto del modelo (MCP)
Reviews

user_KNZZm1lq
AgenticProductSearching by Gen-AI-Developer is a game-changer! The seamless integration and user-friendly interface are impressive. The product has significantly enhanced my search efficiency and accuracy. I highly recommend it to anyone seeking an effective solution for product searches. The performance is robust and reliable. Check it out: https://mcp.so/server/AgenticProductSearching/Gen-AI-Developer