Rust MCP server for denoised web search — fetch, clean, and rerank web content for AI agents.
If you already use Webshift, the webshift mcp server is the piece that lets your assistant work with it directly. Rust MCP server for denoised web search — fetch, clean, and rerank web content for AI agents.
WebShift is a Rust library and MCP server that shifts noisy web pages into clean, right-sized text for LLM consumption.
The toolset is worth reading before you wire it up, because it tells you what the integration is really for:
webshift_query — Full search pipeline: search + fetch + clean + rerank + (optional) summarizewebshift_fetch — Single page fetch and cleanwebshift_onboarding — Returns a JSON guide for the agent (budgets, backends, tips)Because this one is hosted, setup is mostly authentication — you point your client at the endpoint and approve access. Nothing runs on your machine, so there is no runtime to keep patched.
Configuration is passed through the environment: WEBSHIFT_SEARXNG_URL, WEBSHIFT_BRAVE_API_KEY, WEBSHIFT_GOOGLE_API_KEY, WEBSHIFT_BING_API_KEY, WEBSHIFT_LLM_BASE_URL. Treat anything key-shaped as a real credential — scope it to the minimum the server needs, and rotate it if it ever lands in a shared config.
Plenty of search and retrieval servers cover similar ground. The differences that matter in practice are scope of access and how much setup stands between you and a working tool call. Webshift's toolset — webshift_query, webshift_fetch, webshift_onboarding — is a fair guide to whether it matches your workflow. It is maintained by annibale-x; worth a glance at recent repository activity before you build anything load-bearing on it.
We check each listing at SyncDev against the project's documentation before it goes live — if something here drifts out of date, it is a bug worth reporting.
| Tool | What it does |
|---|---|
| webshift_query | Full search pipeline: search + fetch + clean + rerank + (optional) summarize |
| webshift_fetch | Single page fetch and clean |
| webshift_onboarding | Returns a JSON guide for the agent (budgets, backends, tips) |
{
"mcpServers": {
"webshift": {
"command": "docker",
"args": ["run", "-i", "--rm", "searxng/searxng"],
"env": {
"WEBSHIFT_SEARXNG_URL": "your-value",
"WEBSHIFT_BRAVE_API_KEY": "your-value",
"WEBSHIFT_GOOGLE_API_KEY": "your-value",
"WEBSHIFT_BING_API_KEY": "your-value",
"WEBSHIFT_LLM_BASE_URL": "your-value"
}
}
}
}Add to claude_desktop_config.json, then restart Claude Desktop.
| Variable | Description | Required |
|---|---|---|
| WEBSHIFT_SEARXNG_URL | Endpoint or connection string the server talks to. | Yes |
| WEBSHIFT_BRAVE_API_KEY | Credential the server authenticates with. | Yes |
| WEBSHIFT_GOOGLE_API_KEY | Credential the server authenticates with. | Yes |
| WEBSHIFT_BING_API_KEY | Credential the server authenticates with. | Yes |
| WEBSHIFT_LLM_BASE_URL | Endpoint or connection string the server talks to. | Yes |
The simplest web tool that matters — fetch any URL and get model-ready markdown back.
Industrial-strength web extraction — render, scrape, crawl and search entire sites into clean markdown.
Puppeteer-powered browser control that drives pages from the accessibility tree instead of pixels.
Search built for AI, not humans — semantic web search that returns model-ready content, plus code context.
Answers, not links — delegate questions to Perplexity's search-grounded models and get cited responses back.
Put 6,000+ pre-built scrapers at your assistant's fingertips through one MCP endpoint.