What AI engines say about your brand, and who they name instead.

A diagnostic Menlo & Oak runs for your company. We ask AI engines questions written the way your customers write them, and show you, with counts and verbatim quotes, whether they name your brand, whether they recommend it and whether they cite your site.

Measure before you change anything

In “Your Brand Is Not in the Answer” we argued that AI search no longer returns a list of links: it writes an answer and names a handful of brands. That paper's ninety-day plan starts with a baseline: knowing where you stand before you touch the site.

This diagnostic is that baseline, done by us. We run it with Olson, a tool Sebastián Gebhardt built at Yáneken to measure exactly this, and which Menlo & Oak offers to its clients. It isn't software we install for you, or a subscription: it's work we do for you, and it ends in a report.

In the same paper we said most tools in this space are mediocre, and that checking a fixed set of questions by hand every quarter is more honest than most dashboards. We built the diagnostic on that idea: fixed questions, counts in plain view, and the verbatim sentence behind every judgement.

Three figures, never a score

We don't calculate a visibility score. We measure three different things, separately, for each brand:

Named

Does the answer mention your brand?

Recommended

Does the answer advise choosing it? A model makes this call, and it has to quote the sentence in the answer that shows it, so you can read it yourself.

Cited

Does the answer cite your own site as a source?

A brand can be named without anyone recommending it, or have its site cited without the answer naming it. A single number would blend the three and hide which one is missing.

Every rate comes with its counts (for example, “7 of 40”, not just “18%”), so you know how much evidence sits behind each figure. And the same three figures are measured for your competitors, on the same answers.

How we measure

We ask the engines, through their APIs.

We query OpenAI (with web search), Gemini (with Google Search) and Perplexity.

With questions written the way your customers ask.

Some state a need without naming any brand, along the lines of “where's the best place to buy X?”, and show whether you appear when nobody is looking for you by name. Others name your brand and show how you're described.

Each question goes as written.

We add no instructions and no facts about your company, so we measure what the engine says on its own. The questions are drafted without looking at your site, so they come from how your customers ask, not from how you describe yourself.

Your competitors, on the same answers.

If we add a competitor later, it's also counted on the answers we've already collected.

The sources behind the answers.

Which sites the engines cite, and which of them sit behind answers that name your competitors and never you.

Your own site.

What our reader, which doesn't run JavaScript, receives, and which AI crawlers your robots.txt lets in. If your site refuses our reader, we report it as inconclusive, not as a verdict on the page.

What you get

A PDF report

In Spanish or English, with a section for each brand.

A diagnosis

Where you're named and where you aren't, who appears in your place, and which sources the answers cite.

A focus plan

Which pages to publish, where it pays to appear, and what to check on your site.

How it works

  1. 1

    Scope

    We agree with you which brands to measure, against which competitors and on which engines.

  2. 2

    Questions

    We draft the question bank the way your customers would ask, without looking at your site. You review it before the first measurement; from then on it stays fixed until the second, because changing it would break the comparison.

  3. 3

    First measurement

    We run the questions and deliver the report.

  4. 4

    Changes

    Your team works through the focus plan.

  5. 5

    Second measurement

    We run the same bank again, on the same engines and under the same conditions, and compare question by question. With a bank of at least ten unbranded questions, and if your brand appears in at least three of them, the comparison comes with a range showing whether the difference goes beyond what the choice of questions alone would produce.

What this diagnostic doesn't do

Do you give me a visibility score?

No. There's no score, index or grade, by design. There are three separate figures, each with its counts.

Is this the same as asking ChatGPT on my phone?

No. We query the engines through their APIs. What an app shows each person depends on their location, their history and their memory, so we don't present these figures as what any particular customer sees.

Do you guarantee AI will recommend me?

No. Nobody controls what an engine answers. The second measurement records what changed between one run and the next; it doesn't prove the change is down to the work done.

How often do you measure?

When you decide. Each measurement is run on request, because each one has a cost. It isn't ongoing tracking.

What if my site blocks your reader?

We report it as inconclusive. The site reading shows what our reader received, not what any particular AI crawler sees.

Start with the baseline.

Tell us which brands you want measured and against which competitors. We'll reply to agree the scope.