MCP servers for websites in 2026: 44 tools, sorted by who runs them

MCP servers for websites in 2026: 44 tools, sorted by who runs them

Here’s a question I kept turning over: could any AI agent search a website and ask it questions, the way a visitor would, without the owner building anything special? The Model Context Protocol (MCP) is the obvious way in, so in the first week of October 2026 I went looking for every service, platform and open-source project that puts an MCP server in front of a website’s content, for personal sites, small businesses and enterprises alike. I found more than I expected: 44 tools and two standards, each checked against its own documentation or repository on 7 October 2026. The most useful way to sort them turned out to be one plain question: who runs the server?

In MCP is in its Yahoo era I guessed that every business with a website will have an MCP face. This is how far along that is.

In brief

  1. Who runs the server is what sorts them. A server the site owner publishes shows what the owner chooses. A server the agent’s user runs crawls someone else’s site, and the owner has no part in it.
  2. On some platforms you already have one. Every Wix site, every Shopify store, every published GitBook site and every Mintlify site gets an MCP server from the platform, at a fixed address, with no setup.
  3. Hosted services work on any site, inside their own stack. Cloudflare AI Search wants your domain on your Cloudflare account; CustomGPT.ai, SiteSpeakAI and Kapa.ai add one to a chatbot or assistant built from your site; Azure AI Search turns each knowledge base into a server.
  4. Open source that the owner runs is real but scattered. NLWeb is the project built as a public “ask this site” endpoint. Retrieval-augmented generation (RAG) platforms and search engines can serve a site over MCP once you’ve crawled it into them.
  5. Most content management system (CMS) plugins are editing tools behind a login. They’re built for the site’s own team, not for a visitor’s agent.
  6. Discovery has just taken a step. MCP Server Cards became a final proposal on 6 October 2026, and WebMCP is in a Chrome trial. For now, someone usually still has to give the agent the address.
  7. One layer is still missing: one that works on any platform, is run by the site owner, keeps content in sync, lets the owner choose what’s exposed, shows what agents ask and publishes discovery files. A closed beta pitches it; nothing established does it yet.

Who runs the server

An MCP server for a website can be run by four different parties, and that choice decides what an agent can see and who is in control of it.

  • The platform. Wix, Shopify, GitBook and Mintlify switch one on for every site they host. No setup, but only for sites hosted there.
  • A hosted service. You point a service at your site and it hosts the server for you.
  • The site owner. You run open-source software, a CMS plugin or a docs framework, and you decide what it exposes.
  • The agent’s user. Someone who wants to ask your site questions crawls it into a local index and serves it to their own agent. You aren’t involved.
A website at the centre of four parties that can run its MCP server: the platform, a hosted service, the site owner and the agent's user
Four parties can run a website’s MCP server. Only the last one leaves the owner out.

Here are the 44, grouped that way.

Run for the site by its platform or a serviceRun by the site ownerRun by the agent's user, and parts to build withBuilt into the platformWixShopifyGitBookMintlifyWordPress MCP AdapterWebflowHosted servicesCloudflare AI SearchCustomGPT.aiSiteSpeakAIKapa.aiAzure AI SearchOpen sourceNLWeb, NLWebNetwp-jsonld-exporterRAGFlow, OnyxDify, LangflowAgentsetMeilisearch, ElasticCMS pluginsAI EngineDrupal MCP ServerMO MCP Server, MCPIOStrapi, PayloadUmbracoCrestApps for Orchard CoreDocs and static sitesdocusaurus-plugin-mcpFumapressBlume, TomeSumi DocsGitMCPRun by the agent's usersitemcpdocs-mcp-servermcp-crawl4ai-ragmcpdoc, toMCPMCPDocSearchpagefind-mcpBuilding blocksMCP-BVercel mcp-handler

Counted by group, the three groups the site owner runs hold 24 of the 44 tools, and 14 of those serve documentation or a CMS’s own team.

Tools in each groupTOOLS, OF 44, COUNTED OFF THE SECTIONS BELOWBuilt into the platform6Hosted services5Open source, owner-run10CMS plugins8Docs and static sites6Run by the agent's user7Building blocks2

Built into the platform

If your site lives on one of these platforms, you may already have an MCP server without having done anything. The catch is that each one only covers sites hosted there.

PlatformAddressWhat an agent can do
Wixyoursite.com/_api/mcpSearch the site, get business details, and take actions such as booking an appointment. Any visitor’s agent can use it without authentication.
Shopify/api/mcp and /api/ucp/mcpAsk about store policies on /api/mcp, with no authentication. Product search and cart actions moved to /api/ucp/mcp, and the old tools were retired in June and August 2026; the new endpoint asks for no API key but wants an agent profile with every request.
GitBook/~gitbook/mcpRead the published docs. It’s read-only and follows the site’s access settings.
Mintlify/mcp, advertised at /.well-known/mcpSearch the docs. Mintlify hosts it for every site.
WordPress MCP AdapterAn official pluginUse the “abilities” a site registers, as MCP tools.
WebflowYour own sites, through OAuthWork on your own sites.

WordPress and Webflow are different from the rest. WordPress itself has no MCP server, and the adapter logs in with application passwords; Webflow’s server works on your own sites after you authorize it. Both are tools for site owners, not a public “ask my site” endpoint.

Hosted services that work on any site

If your site isn’t on one of those platforms, a hosted service can crawl it and host the server for you. Each one ties you to its own stack in some way.

ServiceWhat you get
Cloudflare AI SearchIt crawls your site and offers a search tool at a /mcp endpoint. The domain has to be on your Cloudflare account, and the endpoint is open to anyone unless you put Cloudflare Access in front of it.
CustomGPT.aiA chatbot builder where MCP is a setting you turn on. Every agent gets a hosted MCP server at no extra cost.
SiteSpeakAIA chatbot builder that turns a chatbot trained on your site into an MCP endpoint.
Kapa.aiA hosted server you set up in one click at yourname.mcp.kapa.ai, built from a crawl of your website plus other sources. Kapa’s product page says 30 or more; the connector list in its docs has 23, the web crawler included.
Azure AI SearchFor enterprises: each knowledge base is its own MCP server with a retrieve tool, and you authenticate in a request header. It runs on the generally available API version 2026-04-01, which returns grounding data only; answer synthesis and query planning still need the 2026-08-01 preview.

Open source the site owner runs

This is the group closest to the question I started with: software you run yourself, on any platform, that answers agents about your own site.

NLWeb and its add-ons

NLWeb is MIT-licensed and led at Microsoft. Every installation is an MCP server with an ask tool, and it builds on the structured data that sites already publish for Schema.org. Of everything I found, it’s the one built as an open protocol for a public “ask this site” endpoint that the owner publishes.

Two projects build on it. NLWebNet is a .NET implementation of the NLWeb protocol, with a reusable library and a demo app that provides /ask and /mcp endpoints; its own README calls it a proof of concept, not ready for production. wp-jsonld-exporter is a WordPress plugin that exports posts as Schema.org JSON-LD so NLWeb’s data loader can import them.

RAG platforms with an MCP server

A RAG platform ingests content and answers questions over it. Several now offer that knowledge over MCP, so a crawl of your site becomes something any agent can query.

PlatformWhat it offers
RAGFlowUp to version 0.27.2, a separate MCP server runs alongside RAGFlow and offers its knowledge-base search, with one fixed API key or shared by many users. The 1.0 release candidate, still a preview, builds MCP into the server itself.
OnyxThe Community Edition is MIT-licensed and has 40+ connectors, including its own web crawler. Its MCP server applies Onyx’s existing permissions automatically.
DifyAny Dify app can be an MCP endpoint. Since version 1.6.0 that’s a switch on the app’s Access Point tab, off by default; before that it took a community plugin.
LangflowEach project gets its own MCP server, its flows become tools, and each project has its own authentication settings.
AgentsetAn open-source RAG platform with content ingestion, built-in multi-tenancy, hosting on custom domains and an MCP server. Worth a look if you want to serve many sites from the start.

A search engine and a crawler

Meilisearch’s scrapix crawls a website and loads it into Meilisearch, and Meilisearch has had a built-in MCP endpoint since version 1.54, still experimental and switched on with a feature flag. In the same way, Elastic’s open-code Open Crawler indexes a site into Elasticsearch; it reached version 1.0 in May 2026 and is no longer labelled beta. Agent Builder then makes the index available over MCP, as a preview in version 9.2 and generally available since 9.3.

CMS plugins: mostly editing tools behind a login

CMS platforms are getting MCP servers too, but most of these are built for the site’s own team: an agent signs in and edits, rather than a visitor’s agent asking questions.

CMSPluginWhat it does
WordPressAI EngineTurns a site into an MCP server that lets agents browse content, edit posts and manage media.
DrupalMCP ServerA module built on the official MCP PHP SDK.
DrupalMO MCP ServerRespects Drupal’s content access rules.
DrupalMCPIOA new alpha with just two tools: search and execute.
StrapiBuilt inAn MCP endpoint at /mcp, now generally available, with detailed permission controls.
PayloadOfficial MCP pluginFollows Payload’s access control rules, and can be turned off per collection.
UmbracoDeveloper MCP ServerFully open source, working through Umbraco’s Management API. Its Base MCP SDK and server generator are open for others to build their own.
Orchard CoreCrestApps MCP moduleAn MCP module, with add-ons that expose files as MCP resources.

Docs and static sites

Documentation sites are well served. These are run by the owner, next to the docs.

ToolWhat it does
docusaurus-plugin-mcpTakes a snapshot of your docs and OpenAPI specs at build time. Serving it at /mcp takes a small serverless function you add yourself, because a static build can’t answer requests.
Fumapress MCP pluginAdds /mcp with three tools: search, get_page and list_pages.
BlumeAn open-source docs framework that includes an MCP server, so agents can search and read the docs.
TomeAn open-source docs platform with a built-in MCP command.
Sumi DocsShows people an Astro and Starlight site and gives agents read-only MCP access to the same content.
GitMCPApache-2.0 licensed and self-hostable. It turns any GitHub repository or GitHub Pages site into a remote MCP server.

Run by the agent’s user

These turn the question around: the person using the agent points them at someone else’s site. They’re handy, and the site owner isn’t involved at all.

ToolWhat it does
sitemcpCrawls a site into a local index and serves it over MCP. Its README now says the project is archived.
docs-mcp-serverCrawls a site into a local index and serves it over MCP.
mcp-crawl4ai-ragCrawls a site from its sitemap, its llms-full.txt or by following links, stores it in a vector database, and offers search tools over it.
mcpdocFrom LangChain: serves a list of llms.txt files plus a tool to fetch their pages. LangChain archived it on 18 September 2026.
toMCPPut tomcp.org/ in front of a URL and it turns the page into clean Markdown.
MCPDocSearchCrawls a documentation site, indexes it and serves it over MCP.
pagefind-mcpSearches any static site that already has a Pagefind search index.

Building blocks

Two projects help you build your own rather than giving you a server:

  • MCP-B’s libraries recreate WebMCP in browsers that lack it, and can pass a page’s tools to desktop agents such as Claude Desktop or Cursor. The libraries are MIT-licensed; the browser extension isn’t open source.
  • Vercel’s mcp-handler adds an MCP server to apps built with Next.js, Nuxt, SvelteKit or Hono.

Standards to watch

WebMCP

WebMCP is the browser-based version of the idea: the page itself offers tools to an agent working in the browser. It’s now called document.modelContext, and Chrome is testing it through version 156 in an origin trial, which lets sites try a new feature early. Its tools only work while the browser tab is open, whereas a server-side endpoint answers whenever it’s asked, so WebMCP adds to an endpoint rather than replacing it. Shopify has already switched it on: since August 2026 its WebMCP tools are live on every Liquid storefront.

WebMCP adds to an endpoint: WebMCP works while the browser tab is open, a server-side endpoint answers whenever it's asked
WebMCP lives in the open tab; the endpoint is always there.

Discovery: MCP Server Cards

An agent can only use a site’s MCP server if it knows where to find it. MCP Server Cards are the answer the MCP project settled on: a site publishes an AI Catalog at /.well-known/ai-catalog.json that links to or embeds a Server Card for each of its MCP servers, and each card says how to connect. The proposal, written on behalf of the Server Card Working Group, became final on 6 October 2026. Mintlify already advertises its servers through /.well-known/mcp. Until agents and sites catch up, someone usually still has to give the agent the address.

What none of them does yet

Step back and the picture is uneven. The built-in options only cover their own platforms. The hosted ones ask you to put your domain on Cloudflare, take on a chatbot product, or build on a search service such as Azure AI Search; Yext and Algolia also run hosted MCP servers over the data you keep with them, both in beta. The CMS plugins are mostly editing tools that need a login, and the docs tools only cover documentation. The crawl-to-MCP tools are run by whoever is using the agent, not by the site owner. Outside docs sites, NLWeb is still the main open project built as a public “ask this site” endpoint that the owner publishes.

What I didn’t find is a clear leader for a layer that works on any platform and that the site owner controls, especially for WordPress, custom-built and enterprise CMS sites. That layer would need to do four things:

  • keep content in sync,
  • let owners choose what’s exposed,
  • show what agents are asking,
  • publish discovery files.
What none of them does yet: keep content in sync, let owners choose what's exposed, show what agents are asking, publish discovery files
The four jobs a site owner’s own layer would need to do.

The closest I saw is SiteToMCP, which pitches exactly this: one script tag on any stack, WordPress included, an allowlist of what agents can see, audit logs and analytics, and automatic discovery. It’s in closed beta, so I couldn’t judge how well it works. Kapa.ai already shows owners what developers and agents ask through its server, but it’s aimed at docs and knowledge bases. Open-source options remain thin for a version that works on any platform and handles crawling, keeping content current, access rules and discovery.

Method and caveats

  • The search was an AI assistant’s web search in the first week of October 2026, followed by a second, wider pass over open-source projects. I checked every claim against the product’s own documentation, changelog or canonical repository on 7 October 2026, and a second check tried to overturn each claim the first one couldn’t confirm.
  • Where the documentation disagreed with the search (Shopify’s endpoints, the Server Card proposal’s status, Dify, Elastic, Azure AI Search, Drupal’s MCP Server module, docusaurus-plugin-mcp, RAGFlow and Kapa’s connector count), this article follows the documentation. sitemcp and mcpdoc are archived and are listed as such.
  • Two tools from the search are left out: AnyDocs MCP, whose repository no longer exists, and AgentLayer, a demo with no public repository.
  • Where the search cited a fork or a directory listing, the links here go to the canonical repository or the vendor’s documentation.
  • Nothing was installed or run, so this reviews what each tool documents, not how well it works. SiteToMCP and Yext’s and Algolia’s servers are described from their own pages.
  • The count of 44 takes a search engine and its crawler as one tool (Meilisearch with scrapix, Elastic’s Open Crawler with Agent Builder) and each Drupal module on its own. The two standards and the three services named in the last section aren’t counted.

I’m Amir Pournasserian. I build AI and platform systems for a living, maintain FluentCMS and YeSvelte, and write here about what I find along the way.