An MCP server for interacting with the [Hugging Face Dataset Viewer API](https://huggingface.co/docs/dataset-viewer), providing capabilities to
Dataset Viewer MCP server is a hosted integration for AI assistants that speak the Model Context Protocol. An MCP server for interacting with the Hugging Face Dataset Viewer API, providing capabilities to browse and analyze datasets hosted on the Hugging Face Hub.
Once Dataset Viewer is connected, these are the calls the assistant has available:
Parameters — - dataset: Dataset identifier (e.g. 'stanfordnlp/imdb')config — Configuration namesplit — Split namequery — Text to search forwhere — SQL WHERE clause (e.g. "score > 0.5")Resources — The Resources tool exposed by this serverTools — 1. validate - Check if a dataset exists and is accessible - Parameters: - dataset: Dataset identifier (e.g. 'stanfordnlp/imdb') - auth_tokenPrerequisites — The Prerequisites tool exposed by this serverSetup — 1. Clone the repository: bash git clone https://github.com/privetin/dataset-viewer.git cd dataset-viewerBeing a remote server, there is no local install. You register the endpoint with your client, authorise it once, and the tools appear.
You will need one environment variable: HUGGINGFACE_TOKEN. The server will not start without them, which is usually why the tools fail to appear on a first run. Keep credentials in your client's env block or a secrets manager rather than in a file you might commit.
Among the developer tooling options, the useful question is rarely "what can it do" but "what does it cost you to run" — permissions, credentials, and how much of your context its toolset consumes. Dataset Viewer's toolset — Parameters, config, split and 6 more — is a fair guide to whether it matches your workflow. It is maintained by privetin; worth a glance at recent repository activity before you build anything load-bearing on it.
SyncDev reviews every entry in this directory against the project's own documentation before publishing, and revisits them as servers change.
| Tool | What it does |
|---|---|
| Parameters | - dataset: Dataset identifier (e.g. 'stanfordnlp/imdb') |
| config | Configuration name |
| split | Split name |
| query | Text to search for |
| where | SQL WHERE clause (e.g. "score > 0.5") |
| Resources | The Resources tool exposed by this server. |
| Tools | 1. **validate** - Check if a dataset exists and is accessible - Parameters: - dataset: Dataset identifier (e.g. 'stanfordnlp/imdb') - auth_token (optional): For private datasets |
| Prerequisites | The Prerequisites tool exposed by this server. |
| Setup | 1. Clone the repository: bash git clone https://github.com/privetin/dataset-viewer.git cd dataset-viewer |
| Variable | Description | Required |
|---|---|---|
| HUGGINGFACE_TOKEN | Credential the server authenticates with. | Yes |
Kill hallucinated APIs — version-accurate, up-to-date library documentation injected straight into context.
Microsoft's official browser automation server — drive a real browser through the accessibility tree, no screenshots needed.
GitHub's official server — repos, issues, pull requests, Actions and code security, straight from your assistant.
Issue tracking at the speed of conversation — Linear's official hosted server with OAuth and zero install.
Local repository surgery — status, diffs, commits, branches and history for any repo on disk.
Timezone sanity for AI — current time anywhere and correct conversions, without the model doing date math.