The four files
Written to be executed, not read. Hand the first four to Claude, ChatGPT, Cursor, or your developer's agent, start with ai.md, and load the vertical pack that matches the business. Then run measure.md yourself.
| File | What it does |
|---|---|
| ai.md | The execution spec. Preconditions, then seven ordered tasks — entity graph, ai.txt and agents.md, a reasoned robots.txt audit, honest sitemap lastmod, IndexNow, entity consistency, page-structure retrofit — each with a definition of done. |
| verify.md | The self-check. Forty-odd checks the agent runs against the live site, ending in a dated report you keep. This is what turns advice into a deliverable. |
| measure.md | For you, not an agent. The by-hand protocol for finding out whether any of it changed what the engines actually say: the vantage rules, how to pick questions, what to record, eight ways to fool yourself, and the thirty-day re-run. |
| profile-law-firms.md | Law firms and solo attorneys: schema types, registries, question shapes, and the state-bar boundary. |
| profile-med-spas.md | Med spas and aesthetics: schema types, licence registries, question shapes, and the medical-board and FDA boundaries. |
Why we published it
The honest reason first: the file was never the valuable part. A technical retrofit is a known quantity. Schema, crawler access, a sitemap that tells the truth, pages structured so a passage can be lifted — none of that is a secret, and charging five thousand dollars for it depends on the buyer not knowing that.
The second reason is a conflict we did not want to be in. If we sell you the build and then grade the build, we are marking our own homework. When the spec is public and free, grading against it is clean. You can check our work against the same file we handed you, and so can anyone else.
And it answers the question we previously had no good answer to. Couldn't I just do this myself? Yes. Here is the file. The part you cannot do yourself is knowing whether it worked.
Others publish playbooks too. Most are lead magnets pointing at a build. This one is the boundary of our product: everything on the near side is yours for free, and we are explicit about what sits on the far side.
What it deliberately leaves out
llms.txt is not in the task list. Ahrefs studied 137,210 domains and found 28% publish one — and that 97% of those files received no requests at all in the month measured. We publish one on this site and treat it as hygiene rather than a lever. Add one if you like; it takes ten minutes. It is out of the ordered tasks because this spec will not ask an agent to spend effort on something the evidence says is not being fetched, and because a named deliverable that does nothing is how buyers get sold packages by the item.
The hard stops, and why they are in a technical spec
Both starting verticals are regulated, and an agent left to write freely produces exactly the sentences that draw complaints. So the refusals are written into the file rather than left to judgement.
An agent running this spec will not generate clinical outcome claims, efficacy claims, comparative superiority claims, or case results. It will not touch before-and-after images. It will not publish under a named person's byline without that person's written approval. It stops and escalates on any credential, price, or guarantee. And it will not write, draft, edit, or solicit a review, a testimonial, or a third-party endorsement — not on your site, not on a directory, not on a forum, not as a draft for you to post.
We don't write anonymous third-party endorsements for regulated practices. Some of the competition ships Reddit and Quora answer drafts as a named deliverable. When a state bar or a medical board holds our client responsible for every word published on their behalf, we would rather be the vendor that says no, in writing, on a public page.
The honest limit
This spec makes a business readable. Whether it gets recommended depends on things no retrofit controls: what else exists in the market, what third parties say, and which sources each engine happens to select this week.
Two numbers we publish because they cut against selling you a build. In our own cross-panel measurement — 383 first-position transition pairs across six engines — an unmanaged #1 AI recommendation had an expected lifetime of roughly one to two days. And SISTRIX, across 82,619 prompts over 17 weeks, found ChatGPT replacing up to 74% of its cited sources every week.
So run the file. The site becomes legible. Whether that legibility turned into a recommendation, and whether it held past the week you bought it, is a measurement question.
Version 1.0 of this spec stopped there, and stopping there was the thing we criticise the rest of the category for. Installation is not an outcome, and a readability audit that gets read as a visibility result is worse than no report at all. So measure.md now closes it: the vantage rules, the question list you write before you look, the four events to record, eight ways to fool yourself, and the thirty-day re-run that is the only part that tells you whether you have a position or a coincidence. It is free, it is by hand, and running it will teach you more about your own position than most people learn from a five-thousand-dollar build.
Versioning
This is a live document about a moving target. Crawler names change, schema vocabulary evolves, engine behaviour shifts. Every version is dated and every change is listed, because the version history is itself the evidence that someone is still watching.
| Version | Date | Change |
|---|---|---|
| 1.1 | 27 August 2026 | Added measure.md — the by-hand protocol for measuring whether the build changed what the engines say. v1.0 stopped at "readable", which was the same gap we criticise the category for. ai.md §10 and verify.md now hand off to it. |
| 1.0 | 27 August 2026 | First public release. ai.md, verify.md, and vertical packs for law firms and med spas. |