PART 4

What the commercial tools offer

This part is a plain description of what's on the shelf, with no recommendations attached. The point is to see which capabilities have already been productized, so that you can tell what you'd be rebuilding and what you'd be genuinely inventing. Capabilities are accurate as of August 2026 (with a few September 2026 finds woven in below — LocIn, ContentRX, Gummble, Ditto Specs) and they move fast, and company sizes are approximate.

The two purpose-built tools

Only two products exist whose entire subject is UX content, and both are small teams. Their full mechanics are on their cards — Ditto and Frontitude. What matters at this level:

Emerging beside these two — not replacing them — are agentic review and evidence products such as ContentRX (one editorial standard across chat, editor, PR, and CLI) and Gummble (searchable library of real shipped screens and microcopy for agents). They are commercial layers for checking and researching copy; they are not string systems of record. Be honest about that when you compare them to Ditto or Frontitude, and about what is open versus paid.

The two enterprise governance platforms

Neither is a UX content tool. Both govern all of a company's content — documentation, marketing, support, legal — with UX copy as one content type among many. They matter here because they solved terminology and style enforcement at scale years before the content-design tools existed, and because their customers are the enterprises a content team competes with internally for budget.

The localization platforms, where they overlap

These are translation-first products, but four of them ship style-grounding machinery more rigorous than anything in the content-design category — which is why they keep appearing in this report.

What to borrow MQM — the translation industry's shared error taxonomy — already separates terminology and style from accuracy and fluency. That's a ready-made rubric skeleton for evaluating generated UI copy, and content design has no equivalent of its own. If you build your own eval layer, you don't need to invent the dimensions.

The documentation platforms

These platforms publish documentation rather than product copy. They're here because two of them have shipped, as ordinary features, several things this report otherwise describes as unsolved.

GitBook: a style guide as a product object

GitBook's style guide feature is the most complete productization of this report's own recommendations that exists in any tool, content design included. Five mechanics:

The line to copy verbatim "The style guide is the Agent's only source of rules. If your style guide mentions an upstream guide — like Google's or Microsoft's — as its base, that mention is background for human readers, not an instruction to the Agent. If a convention matters to you, write it down."

That closes the most common hole in a set of grounding documents, which is a guide that defers to another guide the agent can't actually read. Every team whose style guide says "otherwise follow AP style" has this bug.

GitBook also generates an MCP server automatically for every published docs site — "Docs published on GitBook automatically generate an MCP server you can hook up to external tools" — plus an agent that drafts updates from support tickets and product changes, and identifies outdated pages for review.

Mintlify, and one claim that doesn't survive checking

What this means for content design

What's productized, by capability

CapabilityDittoFrontitudeWriterMarkup AILocalization platforms
String source of truthFullFullFull (translation-shaped)
Component-linked stringsYesYes
Guidelines bound to componentsYes
Style guide as structured rulesYesYesYesYesCrowdin, Lingo.dev
Suggestion cites the ruleYesYesYesCrowdin (QA flags)
Terminology / glossary enforcementYesYesYesYesYes (semantic, Lingo.dev)
Generation grounded in your corpusYesYesYesYes (translation)
MCP / agent surfaceYesEarly accessCrowdin, Tolgee, most majors
PR / repo review botYesCI integrations
Figma surfaceYesYesYesYes
Style guide mined from your copyYesCrowdin (from translations)
Quality scoring with a rubricAnalytics onlyDashboardsYes (MQM, Lingo.dev)
Localization built inYesYesNative

Bold marks the strongest implementation of that capability in the commercial layer. "—" means not found in public documentation as of August 2026, not necessarily absent.

Terminology and hierarchy: the two extremes

These two capabilities sit at opposite ends of what the commercial tools offer. One is the most thoroughly productized thing in this entire report. The other barely exists as a product at all.

Terminology is mature, buyable, and mostly built by another industry

If you want terminology enforcement, you are not on the frontier — you are shopping. The localization and technical-writing industries have been building this for years, and 2026 added AI extraction on top:

Four mechanics from that industry transfer regardless of what you buy:

  1. Hard versus soft enforcement per term. Criticality decides whether a violation blocks or warns — not one global severity for the whole glossary.
  2. The termbase exposed via API, so authoring, review, and QA tools all query the same source rather than keeping copies.
  3. Termbase checks in the CI pipeline, so violations surface before strings reach anyone downstream.
  4. Retrieval grounding on the approved termbase before generation — TAG's whole premise, and the same idea as terminology.
What none of them do Every product above manages and enforces a termbase that already exists. None of them helps you decide what the term should be — weighing product consistency against industry convention against what your audience actually says, and recording why. That research-and-decision layer is what the Wix terminology skill does, and it has no commercial equivalent. The split is clean: enforcement is a bought capability; the decision is yours.

Information hierarchy is reviewed by two other disciplines, and by nobody as content

Hierarchy does get checked in 2026 — just never as a content question. Two adjacent forms exist:

Nobody asks the content-design version of the question, which has three parts:

Every commercial content tool in this part operates on strings — one at a time, in isolation, with no model of the screen they sit on. The two closest things that exist are both agent skills, not products: the Wix whole-screen review, which treats structural critique as a permitted output, and better-interface on the Figma shelf, which puts writing rules alongside layout, typography and accessibility and insists on one consolidated verdict rather than separate audits.

There is nothing to buy here.

Three things the commercial layer settled

  1. The system of record is the product; the AI is a feature.

    Every one of these tools sells a place where approved content lives.

    Between 2025 and 2026 the AI layer changed completely — new models, new surfaces, MCP — while the libraries, component linking, and review workflows stayed put.

    That stable part is where the years of work are. It is also the part an internal build usually already has in some form: a translation-key system, a design-system doc, a spreadsheet.

  2. Agent surfaces arrived everywhere in 2026, and nobody treats them as optional.

    Ditto shipped an MCP server, a PR bot, a CLI, Specs, and an agent-skills package. Frontitude entered early access. The localization platforms shipped MCP servers across the board. Newer commercial layers such as ContentRX and Gummble sell the agent doorway for review and evidence rather than for owning the library.

    The bet is uniform: copy will increasingly be generated inside coding agents, so whoever's rules the agent consults becomes the source of truth. Writer's absence from this pattern is still the exception to watch — as of the August 2026 commercial sweep, no MCP or agent surface was disclosed; this draft does not invent later Writer features without a verified source.

  3. "Cite the rule" became table stakes.

    Ditto, Writer, Markup AI, and Crowdin all name the rule behind a flag rather than issuing an unexplained rewrite. GitBook goes one step further and cites a stable rule ID.

    This converged because of the trust math for review tools. Once a meaningful share of suggestions are wrong, users stop trusting all of them — and a suggestion that names its rule can be argued with, which is what keeps it trustworthy.

What nothing commercial covers

Verified absences as of August 2026 — checked against product documentation, not inferred. These are the openings an internal build has to itself. (September 2026 additions such as ContentRX and Gummble narrow the review and evidence gaps a little — same-standard multi-surface checking, and searchable microcopy from shipped products — but they still do not own a string system of record, and they do not close whole-screen structural review.)

Where the blocker actually is

Three findings from the August 2026 sweeps locate it fairly precisely, and none of them is about capability.

First: a Google team used the exact mechanic that could govern copy to fence copy out.

The site-kit-wp repo ships a Gemini review style guide with a section headed "User-facing copy comes from design." It instructs the reviewer: "Treat user-facing strings and i18n copy as authored against design (Figma) and product decisions. Do not suggest wording, tone, or capitalization changes to display strings."

The agent could review copy. It is told not to, because copy authority lives in Figma — outside the repo the agent can see.

Second: of the big-tech design systems checked, almost none publishes its product-copy guidance in a form an agent can read.

Third, and most telling: the companies building these tools do not publicly govern their own product copy with them. A sweep of the AI labs and AI-native products found one public artifact governing a company's own user-facing strings — Gemini CLI's string-reviewer, and it is dogfood-only. Four companies govern their own long-form documentation well: GitHub, Anthropic, OpenAI, Lovable. One does both — Vercel, with a voice test that fails a build. For product interface copy the record is near-silent.

The sharpest data point: anthropics/skills/brand-guidelines publishes Anthropic's own brand as a machine-readable skill, and it contains colours and typefaces only — no voice, no tone, no terminology. The same asymmetry as every design system in this report, at the company that ships the agent runtime.

Put together, this means the tooling to govern copy exists and the guidance exists, but the two are kept apart from each other. Copy is owned in design tools, while the agents work in repos.

That is an organisational boundary, not a technical one — which is why the builds that solved it all did the same thing first. They put the rules where the agent already works.

Read against Part 1, the shape is consistent: