Why isn't an AI visibility dashboard enough?
Measuring is the cheapest part of the job and the part that changes nothing on its own. What to ask a tool before you buy it.
Published 26 August 2026
At a glance
- A dashboard measures the last move of a ten-move job and sells it as the whole job.
- The daily number moves on its own, mostly noise rather than signal, which is why we re-run a frozen set quarterly instead of watching it jitter daily.
- Who writes the questions decides the answer: ours are generated blind to the brand for the first seven, so the reading can't flatter you.
What does a visibility dashboard actually do?
You give it prompts. It asks the assistants those prompts on a schedule, counts how often your brand appears, and draws that as a percentage over time. Some add which competitors showed up alongside you.
That is a real capability and it was worth building. The problem is what happens next. You have a figure. It went down this week. Nobody in the room can say why, and nobody can say which of the forty things you could do would move it back.
Why does the number move on its own?
Ask an engine the same question twice and the wording moves. Ask it on two consecutive mornings and the brands it names can change without a single thing altering on your site or anyone else's. Model updates, retrieval variance and plain sampling noise all sit underneath that line.
So a daily figure is mostly measuring the model, not you. Watched daily it invites exactly the wrong behaviour: reacting to noise, and mistaking movement for feedback. The method re-runs the same frozen question set quarterly for that reason, and says so on the pricing page rather than selling the extra runs.
Who writes the questions, and why does it matter so much?
This is the part that gets least attention and decides most of the answer. If the tool asks you to supply the prompts, you are setting your own exam. Not dishonestly: you know your category, so you reach for the phrasing where you have a case. Every brand does it, and the reading comes back flattering.
A buyer who has not heard of you cannot phrase a question that way, because they do not know what you sell. Our first seven questions are generated without the product in front of the writer at all, and the brand name is forbidden in them. That is enforced in the generator and re-checked afterwards. Only the last three name you, because by then the buyer has found you and is working out whether you fit.
Isn't a keyword with a question mark on it close enough?
No, and this is testable in about a minute. Paste best crm for small business into an assistant, then paste two honest sentences describing your actual situation. You get different answers, naming different brands, at different lengths.
People do not type queries into chat windows. They type a paragraph with an admission in it. So that is the shape we generate: two sentences, first person, twenty to forty-five words, including what they are unsure about, asking for a framework rather than a fact. Those constraints are validated after generation, not suggested.
"best rfp software 2026"
versus the way people actually type it:
"There's no one here whose actual job this is, so it lands on whoever has capacity and it's usually the people we can least afford to lose for a week. How do other businesses handle this without a dedicated team?"
What is the alternative, concretely?
Ten moves in a fixed order, the same order the product tracks your progress through. The measurement is still there. It arrives six moves in rather than on day one, and you re-run it against a frozen set, quarterly on our advice, rather than watching it jitter daily.
Know who buys
- Review your territory: up to three named audiences read from your own site, then the questions those people really type. You edit every one. You leave holding: a question bank you signed off.
Read your site
- Where are you? Can an assistant reach your site, read it as served, and work out what you sell. You leave holding: your starting position, in plain English.
- Where's your current content? Every page you already have, in one inventory. Most brands can name less of it than they expect. You leave holding: the full list, nothing hiding.
- Audit: every page marked out of 100, with a named reason behind every point. You leave holding: a score per page, and the reason for every point.
- Technical: the thirty points that are not about writing: crawler access, markup, findability. You leave holding: a PDF your developer can act on.
- Improve your current content: fix what exists before commissioning anything new; the audit already said what, and where. You leave holding: rewritten pages, in your own voice.
Ask the engines
- How you placed: every question, every engine, and we keep the words that came back. You leave holding: the full answer, with who got named instead.
- White space: the questions where nobody in your category has been claimed yet. Open ground beats a contested win. You leave holding: topics nobody owns.
Create a snapshot
- Bring your questions across: choose the set that matters and lock it. Same questions, same wording, every time from here. You leave holding: a frozen set.
Close the gaps
- Close the gaps: white space arrives as topics, checked against what you have. Drafts come back in your voice, built to be quoted. You leave holding: the missing answer, written.
Now the snapshot means something. Your frozen set, re-run when you choose. We suggest quarterly. Same questions, same wording: movement you can trust, because everything above happened first.
How the two approaches actually compare, move by move:
| Monitor-led tools | UpliftAEO |
|---|---|
| The asking starts on day one, from your SEO keyword list with question marks added. The six moves that would change the answer are left to you and whoever you can hire. | The asking starts six moves in, after you have fixed what you already had, learned who buys, and signed off the questions. It measures work, not weather. |
What does "actionable" have to mean?
A score is only useful if losing points tells you what to change. So every page you own comes back marked out of 100, seventy for the writing and thirty for the technical side, with every point tied to a specific, nameable reason rather than one opaque number.
The writing half rewards the things that decide whether an answer gets lifted whole into a reply: opening with a direct answer instead of throat-clearing, section headings phrased as the actual questions a reader would type, a takeaway a reader (or a model) can find without reading the whole page, real specifics rather than vague claims, and products or brands named plainly rather than gestured at. The technical thirty covers whether a crawler can reach the page and understand what it says in the first place, and arrives as a PDF your developer can act on directly.
You do not lose eight points and wonder what happened. You lose them because your section titles are not questions, and now you know both the fix and roughly what it is worth. That is the difference between a number about a page and a number about a brand: one of them can be acted on this afternoon.
When is a dashboard the right buy?
Genuinely: if you already have the rest. A large brand with an in-house content team, an SEO lead who understands extractability, and a developer who will act on a crawler audit does not need to be walked through the work. It needs the instrument, and should buy the best instrument it can.
The mismatch is everyone else. Most marketing teams do not have that bench, and a percentage that fell four points this week is not a brief anybody can act on.
Measuring is not the work. It is how you find out whether the work landed.
Which is also why we cannot promise a result. Nobody controls what a model says. What a method can promise is that the same questions, asked the same way, either moved or did not, and that when they did not, you know which of ten moves to look at.
Frequently asked questions
Should I cancel my visibility dashboard?
Not necessarily. Measurement is genuinely useful; the argument here is against treating it as the whole job. A dashboard that tells you the number moved is fine as long as something else tells you which of the ten moves to look at when it does.
How often should the question set actually be re-run?
Quarterly, on a frozen set, is the advice. Not because a scheduler enforces it, but because daily or weekly reads are mostly measuring model noise rather than your own progress. Re-running more often just gives you a jitterier version of the same number.
Who should write the questions used to test visibility?
Not the brand being tested, at least not for all of them. The first seven should be generated blind to the product, with the brand name forbidden, because a brand writing its own exam reaches for the phrasing where it already wins. Only the last few, aimed at someone who has already found you, should name you directly.

Written by Olly Barrett
Read next
AEO vs SEO: what carries over and what doesn't
Search optimisation makes a page worth clicking. Answer optimisation makes it worth quoting. Here is which of your existing habits still help, and which now cost you.
How AI decides which brands to name
The five patterns shared by brands that keep getting named in AI answers: being legible, answering the real question, existing in more than one place, and two more.
The get-found checklist: nine things to check this week
Nine things to check on your own website this week, most in under an hour, before you spend anything on a tool. Start by asking the assistants about yourself.
Ready to find out where you stand?
Give us the address. We’ll read the site, work out your audiences and show you the questions before anything runs. Credits only get spent when you say so.
