Blog · · 7 min

MCP server for conversion rate optimization

No MCP tells you where to put a button. What exists splits into three families: measurement (Microsoft Clarity, PostHog, Chrome DevTools), experimentation (GrowthBook) and heuristic audit. Which family to connect, and the sample size below which a test proves nothing.

By

There is no MCP that makes a page convert. There are MCPs that measure a page, MCPs that run an experiment on it, and MCPs that score it against a checklist. Three different jobs, and only the first two produce evidence. We searched the official Model Context Protocol registry on 24 August 2026. conversion returns eight servers, seven of them GroupDocs document format converters and one a PDF-to-JSON tool. upsell, landing page, a/b testing, ui components and component library each return zero. Whatever you connect, you are connecting it from outside that registry.

Two mobile product pages side by side at the same width. Left: a centred generic layout (a purple banner, a title, one photo on black, one price, a paragraph). Right: the tins on an open hand, a gallery strip, the price struck through against the saving, and the cost per cup.
Measurement tells you the left-hand page loses. It does not tell you that the fix is on the right, or why. That is the ceiling every audit family shares.

Family one, measurement

These read what already happened. They are the only tools here that cannot lie to you, because they report rather than judge.

Microsoft Clarity MCP. npx @microsoft/clarity-mcp-server --clarity_api_token=your-token-here. Official Microsoft package, three tools: query the analytics dashboard, list session recordings with filters, query the documentation. Clarity itself is free, which is why it is the default.

Its limits are the part nobody quotes, and they decide whether it is usable in a loop. Microsoft's Data Export API documentation states a maximum of 10 API requests per project per day, retrieval confined to the previous 1 to 3 days, a maximum of three dimensions per request, and responses capped at 1,000 rows with no pagination. An agent that polls Clarity on every iteration will exhaust the day's quota before lunch. Treat it as a once-a-day pull into your own store, not as a live sensor.

PostHog MCP. npx @posthog/wizard mcp add. Runs HogQL queries, ships feature flags, reads error traces. If the merchant already has PostHog, this is the shortest path from a question to a number.

Chrome DevTools MCP. npx chrome-devtools-mcp@latest. Published by Google, Apache-2.0. It drives a real Chrome through puppeteer-core, so an agent can load the actual page and read the actual trace instead of reasoning about the source. This is the one that turns "the page feels slow" into a measurement.

Family two, experimentation

GrowthBook MCP. The @growthbook/mcp package describes itself as a thin MCP for GrowthBook, a skill loader plus an authenticated API passthrough. GrowthBook is open source and self-hostable, which matters when the merchant's traffic is small enough that per-seat experimentation pricing makes no sense.

Intelligems is the Shopify-native price and offer testing tool that comes up constantly in this category. As of 24 August 2026 we found no npm package under that name; if they run a hosted MCP endpoint, connect it as a remote server rather than looking for a CLI.

Family three, heuristic audit and its ceiling

This is where the search results are thickest and the evidence is thinnest. A heuristic audit loads your page, compares it to a list of rules, and returns a score.

Baymard UX-Ray is the serious member of the family. It scans an ecommerce site against Baymard's own UX guideline database and groups findings by severity. Reading its public page on 24 August 2026, it publishes a methodology link but no accuracy figure on that page, and mentions neither an API nor an MCP. There is a free demo scan.

The rest of the category (the landing-page graders, the AI CRO scorers) share one structural problem, and an agent should say it out loud to the merchant rather than pass the number along. A generated score is an exam the tool sets and marks itself. It correlates with the tool's rulebook, not with the merchant's revenue, and none of the ones we could read publish a correlation between their score and a measured conversion change. A score is useful as a checklist. It is not evidence.

A product card that names the format problem (a tub, a scoop, and a kitchen you have to be standing in), then reveals the answer: a case of thirty sticks at €23.90, with the division written out in full underneath.
The division, done on the page: €23.90 ÷ 30 = €0.796666… shown rounded to €0.80. Worth more than any heuristic score out of a hundred.

The number that decides whether any of this matters

Before an agent recommends an A/B test, it should compute whether the merchant can run one at all. This is arithmetic, not opinion.

For a two-proportion test at 95% confidence and 80% power, the sample per variant is roughly n = 2 × (1.96 + 0.8416)² × p̄(1−p̄) / δ².

Run it on a realistic store. Base conversion 2%, and you want to detect a 20% relative lift (2.0% to 2.4%). That needs about 21,000 sessions per variant, so 42,000 sessions for the test. Want to detect a 10% relative lift (2.0% to 2.2%)? About 81,000 sessions per variant, 162,000 in total.

A shop doing 10,000 sessions a month cannot detect a 10% lift inside a quarter. That is not a reason to do nothing. It is a reason to change what you do. Below that threshold, ship changes that are defensible on grounds other than a test: a legal requirement, a documented usability finding, a page that loads. And stop scoring.

What to actually connect

  • If the merchant already has PostHog, run npx @posthog/wizard mcp add and query real funnels.
  • With no analytics at all, install Clarity, then the Clarity MCP, and respect the ten-calls-a-day ceiling.
  • If something is slow or broken and nobody knows where, npx chrome-devtools-mcp@latest.
  • Above roughly 40,000 sessions a month, with a real question to settle, GrowthBook, self-hosted.
  • Below that, no experimentation stack. Fix the mechanics that do not need a test. See where to place an upsell and the bundle selector spec.

The gap, stated plainly

To our knowledge, as of 24 August 2026, no MCP hands an agent a placement rule. Every server in this category returns data about a page that already exists; none tells the agent what to build. That is a real hole, and an agent asked to improve a store today has to fill it from its own priors. That is exactly why the pages it produces resemble each other.

Disclosure. We are building uxgen, an MCP aimed at that hole. It is not installable yet, and nothing above is ours.

FAQ

Is there an MCP server for conversion rate optimization?

None that tells you where to put a button. What exists splits into three families: measurement such as Microsoft Clarity, PostHog and Chrome DevTools, experimentation such as GrowthBook, and heuristic audit. Searching the official registry for conversion returns eight servers, all of them document format converters.

Which one should I connect first?

Measurement, because it is the only family that tells you where the loss actually happens. Experimentation is worth connecting only once you have the traffic to conclude anything from a split.

Do automated CRO scores predict conversion?

No vendor in the category publishes a correlation between their score and a real conversion rate. Treat a generated score as a checklist of omissions, which is useful, and not as a prediction, which it is not.

How much traffic do I need before a test means anything?

For a store converting around 2% chasing a 10% relative gain, roughly forty thousand sessions per variant. Below that, the result is noise pointing in a random direction.

uxgen is a service of UXGen AI, LLC — 131 Continental Dr, Suite 305, Newark, DE 19713, United States.

© 2026 UXGen AI, LLC. All rights reserved.