Skip to content

How to improve your AI search visibility

A working guide to being named when a buyer asks an AI assistant for a shortlist. Everything here comes from 860 scored answers across 85 B2B software companies, and from the category sweeps we have published since. Where something does not work, we say so.

of its own category's answers name the median company
Median company rate: 2 of 10 answers, across 85 companies. Interval not reported in the published source.
of companies were named in none of theirs
Observed company proportion: 30 of 85 companies. Interval not reported in the published source.
of the top-cited sources were written by vendors
Observed citation share: 1,454 of 1,879 classified citations among the top 100 domains. Interval not reported in the published source.

Almost every company that sees its own number for the first time is surprised by it, in one direction or the other. Guessing wastes months, so start by measuring.

Write ten questions a buyer would actually type. Not keywords, questions. Cover five shapes, because the shape changes the outcome more than anything you publish: the broad category question, alternatives to a named competitor, two named competitors compared directly, a use case question, and pricing.

Put each question to more than one engine. Claude, ChatGPT, Perplexity and Google AI Overviews disagree with each other more than they agree. A single-engine number is a reading, not a position.

Score two separate things per answer. Whether your company was named, and whether your domain was cited as a source. These diverge more than people expect. In our data 29.4% of companies were cited more often than they were named, and nine were named in exactly zero answers while their own site kept appearing in the citation list.Observed company proportion: 25 of 85 companies. Interval not reported in the published source.

One trap to avoid, because we fell into it ourselves. If four of your ten questions contain a competitor's name, that competitor will appear in those answers almost every time. Counting that as a mention measures your own question, not the engine. Score name-seeded questions separately from open ones. In our category measurement, Rampiq was named 14 times and Zupo 13 times across all 200 answers; neither was named in the 120 name-free answers.

The cited Volume I and Volume II research is published open access under CC BY 4.0 and has not been peer reviewed. The research data, rather than this commercial guide, is covered by that license.

Before publishing anything new, check what the engines already think is true about your company. This is the cheapest work available and it is usually the most wrong.

Three things go stale and keep circulating. Ownership, when a company has been acquired, carved out or renamed and third-party profiles still carry the old parent. Category, when a directory files you under something adjacent to what you actually sell. And name, when your product has several variants in circulation and the engines treat them as different entities.

These are data problems on source pages. Corrections should be made at the source, and their effect checked in later observations.

On your own domain, three things do the same job. One canonical company name used consistently. Organization structured data with a sameAs array listing every property you own, so an engine can tell your LinkedIn page, your GitHub organisation and your research all belong to one company. And the same identifier referenced across every page, rather than each page declaring an unlinked company of the same name.

This is the single most common gap we find, and it has a distinctive signature.

Question shape changes naming rates sharply. Best-of questions named a company 41% of the time in our data. Direct comparison questions named one 20% of the time. And the pattern shifts with how visible a company already is: companies named in none of ten answers mostly lose broad category questions, while companies named in seven to ten lose head-to-head comparisons 68.8% of the time instead.Observed question-shape rate within the 860-answer study; exact per-shape counts are not reported in the published tables. Interval not reported in the published source. The high-visibility absence figure is 11 of 16 classified absences from 8 companies; its interval is not reported.

Data table: Question shape, Named a company
Question shapeNamed a company
Best-of questions41%
Direct comparison questions20%
Observed question-shape rate within the 860-answer study; exact per-shape counts are not reported in the published tables. Interval not reported in the published source.

The reason is usually visible on the company's own site. A page will name three competitors and then say "see how we compare, request a demo". That hands the engines the rival names and withholds the answer, so the comparison gets written from rival and third-party pages, without you in it.

Write the comparison properly instead. Name both sides. Concede rows honestly rather than sweeping every column, because a page that admits a limitation is the most citable thing in most categories and almost nobody publishes one. Include the cases where the other product is the better choice.

An answer engine asked which software to consider in a category needs names to return. A page written to rank for a keyword often contains none.

In one of the prospect sweeps we ran in July 2026, a company whose main guide page was fifteen hundred words on how to choose software in its category and did not name a single company, including its own competitors. It said at one point that everyone knows the biggest players, and then declined to say who they were. That page was cited zero times in a hundred and eighty citations, while six competitor pages that answered the same query with an actual list of platforms were cited repeatedly.

The opposite failure exists too, and it is subtler. In another July 2026 sweep, a company published a category guide that ranked itself first, with the best score in every row of its own comparison table. The engines cited that page three times and named every competitor in the table except the company that wrote it. A vendor ranking itself reads as a claim. The competitor names on the same page are the part that is not self-serving, so that is the part that gets used as evidence.

The lesson from both: name companies, and do not make yourself the answer to every row.

Across the hundred most-cited domains in our study, 77.4% of citations went to vendor-published material. Review platforms took 10.2%, media 6.5%, analyst firms 5.9%, and community sites including Reddit 0.0%. That last number surprises people, so it is worth being precise: this measures what the engines cited when answering B2B software questions in this dataset, not whether Reddit matters generally.Observed citation shares among the top 100 domains: 1,454 vendor, 192 review, 123 media, 110 analyst and 0 community citations, out of 1,879 classified citations. Interval not reported in the published source.

Data table: Source type, Share of citations
Source typeShare of citations
Vendor-published material77.4%
Review platforms10.2%
Media6.5%
Analyst firms5.9%
Community sites including Reddit0.0%
Observed citation shares among the top 100 domains: 1,454 vendor, 192 review, 123 media, 110 analyst and 0 community citations, out of 1,879 classified citations. Interval not reported in the published source.

The practical reading is that vendor-written content dominates the citation layer, and most of it was written by your competitors. The authority is also thin. 56% of cited domains appeared exactly once, and the ten most-cited domains together accounted for only 12% of all citations. There is no settled reference layer in most categories, which is the whole reason the position is still available.Observed citation shares: top 10 domains, 609 of 5,160 citations; top 100, 1,879 of 5,160. The singleton-domain rate uses 1,753 domains; its exact numerator is not reported in the published table. Interval not reported in the published source.

So the work is to be one of the vendors whose material gets read. Answer the questions your category leaves unowned. In our data, use case questions returned no vendor at all 40% of the time, which means nobody has claimed them. Publish original data if you have any, because it is the most citable asset a company can own and almost nobody publishes it. And correct your entry on the third-party pages that already rank, because those are being read whether you like them or not.Observed question-shape rate within the 860-answer study; exact per-shape counts are not reported in the published tables. Interval not reported in the published source.

Keyword-shaped pages with no entities in them. If a page never names a company, there is nothing in it for an engine to lift.
Publishing volume without answering questions. Only 47% of the answers that named a company also cited that company's own site. Most of the time you are being described from someone else's page, so more of your own pages does not automatically help.Observed conditional naming/citation rate within the 860-answer study; the joint naming/citation count is not reported in the published tables. Interval not reported in the published source.
Chasing a single number. AI answers vary between runs and between engines. Treat a stable figure as a position and a moving one as a reading, and never report one blended score.
Waiting for an attribution model. Nobody has a clean one yet. Presence is measurable today and it is the necessary condition for anything downstream. Measure what you can and be honest that the rest is not yet measurable.

Re-run the same questions on the same engines. Not similar questions, the same ones, or you are measuring the change in your question set.

Correct inaccurate entity information and publish useful source material. Then repeat the same questions with the same method to check whether the observed answers changed. There is no fixed timetable for that change.

Report the range, not a point. If you tell a board that you went from three mentions to eight, the first question should be how much that number moves on its own. If you have not measured that, you do not know.

Report naming and citation separately. More citations mean the domain appeared in more of the recorded source lists. That does not establish a corresponding increase in vendor naming, engine visits, leads or revenue.

Start with a measured baseline.

See which buyer questions leave you out, the sources visible in those answers, and the three changes worth testing first. The Category Audit covers ten questions; the Diagnostic extends to 35 questions and 525 scheduled observed answers, adaptive runs extra. AI answers vary, so repeated observations and disagreements are reported.

Start with the $490 Category Audit

Credits in full against the Diagnostic within 30 days.

Get the Diagnostic, $990

Credits in full against the program.

Every figure comes from the published method, v1.1.