TL;DR
When an AI engine answers a B2B software buying question, it does not usually quote your website. It quotes a handful of other domains that it trusts to describe your category. Getting recommended is mostly a question of appearing on those domains.
- Your own site is rarely the source that gets you named. It is necessary, and it is almost never sufficient.
- For B2B SaaS specifically, review platforms and community threads do the heaviest lifting. G2, Capterra, TrustRadius and Reddit come up again and again in category questions.
- The list is different for every category. There is no universal top 50 that helps you. The useful list is the one you generate from your own buying questions, and you can build it in an afternoon.
- Most of the work is unglamorous. Getting added to an existing roundup beats publishing a new page almost every time.
This guide is the practical version: how to find the domains that decide your category, how to get onto them in priority order, and what to ignore.
Why the citing domain matters more than your ranking
Classic SEO trained everyone to think about position. AI search does not really have positions, it has sources. The engine runs a search, retrieves a set of pages, and writes an answer from what it retrieved. Whichever brands are described favourably inside those retrieved pages get named.
That changes where the leverage sits. If your category's buying question consistently retrieves a G2 grid, two comparison articles and a Reddit thread, then your visibility is decided inside those four documents. You can rewrite your homepage every quarter and it will not move, because your homepage was never in the retrieval set.
The uncomfortable implication
Most AI visibility budgets are spent on the one property that has the least influence over the outcome: the company's own website. The site still matters, because it is where the engine confirms what you claim to be and pulls specifics like pricing and integrations. But it is the confirmation layer, not the discovery layer.
The discovery layer is other people's domains. That is the part most teams have no plan for.
Why this is good news
The set of domains that matter for a given category is small, stable and finite. It is usually somewhere between fifteen and forty pages. That is a work queue, not a strategy deck. You can look at it, decide which ones you can realistically get onto, and start.
Compare that to classic SEO, where the target is an abstraction called authority. Here the target is a list of URLs.
How to build your own list in an afternoon
Do not use somebody else's top 50. The domains that decide enterprise HR software are not the domains that decide developer tooling. Here is the method we use at the start of every engagement, and you can run it yourself for free.
Step one: write the buying questions, not the keywords
Write 30 to 40 questions the way a buyer actually types them, with constraints included. Not "best CRM". Something like "best CRM for a 40 person B2B sales team that lives in Outlook and needs HubSpot migration".
The constrained questions matter more than the head terms. They are where a smaller brand can actually be named, because the engine is looking for a specific fit rather than the biggest name in the category.
Step two: run them and record the sources, not the brands
Run each question in ChatGPT, Google AI Mode, Perplexity and Copilot. Most people record which competitors got named and stop there. That is the score. The useful data is the citation list underneath.
For each answer, write down every domain cited. Run each question at least three times, because answers vary between runs and one pass will mislead you.
Step three: count and rank
Tally the domains by how often they appear across your whole prompt set. You will get a steep curve: a few domains appearing constantly, a long tail appearing once. The top ten to fifteen are the ones that decide your category.
Step four: mark the gaps
For each frequently cited domain, answer three questions. Does it mention a competitor? Does it mention you? Could you realistically get onto it?
The rows where a competitor appears, you do not, and the page is reachable are your work queue. In most B2B categories that list is twenty to forty items, and a meaningful share of it is achievable within a quarter.
If you would rather see a baseline before doing the manual work, the free ChatGPT visibility checker runs five buyer prompts and shows you who gets named instead of you, along with the sources the answer used.
The nine domain types that decide B2B SaaS visibility
Ranked by leverage for B2B software specifically. Consumer categories look different, and if you sell to consumers this order is wrong for you.
1. Review platforms, the highest leverage domains in B2B
G2, Capterra, TrustRadius, Software Advice and GetApp. If you sell B2B software and you only fix one thing, fix these.
Why they dominate
Answer engines are trying to give a defensible answer to a subjective question. A review platform is the closest thing to an objective source that exists in software: structured category taxonomies, star ratings, review counts, and comparison pages that map exactly onto the questions buyers ask. When a buyer asks which vendor is best for a use case, these pages are almost purpose-built to be the retrieved source.
How to actually get leverage here
- Claim and complete every profile. An unclaimed profile with a thin description is an entity signal that says you are not a serious participant in the category. Completing it is free.
- Get into the right categories, and the narrow ones. Being ranked in a niche subcategory you can win beats being invisible in the broad one. The subcategory names frequently match the constrained buyer questions.
- Run a review drive, and keep it running. Review recency matters. A cluster of reviews from eighteen months ago reads as a company that has stopped growing.
- Write the comparison pages they host. Most platforms let you complete vendor comparison content. Those head-to-head pages are among the most frequently retrieved documents in the whole category.
What most teams get wrong
Treating reviews as a quarterly campaign rather than an always-on process, and ignoring the free profile fields. The description you write on your G2 profile is a sentence an answer engine may quote verbatim. Write it like it will be quoted, because it will be.
2. Community domains, where the honest opinions live
Reddit above all, plus Hacker News, Stack Overflow and category-specific forums.
Why they punch above their weight
Answer engines are heavily biased toward sources that read like real human experience rather than marketing. A thread where five practitioners argue about which tool actually works is exactly the kind of document that gets retrieved for a subjective buying question. It is also the kind of document your marketing team has no direct control over, which is precisely why the engines trust it.
How to get leverage without getting banned
- Participate as a person, not a brand. Reddit communities are extremely good at detecting marketing, and a removed post is worse than no post.
- Answer questions in your area of genuine expertise. The mention that helps you is the one where someone else recommends you, or where you demonstrate competence and disclose who you work for.
- Find the existing threads, do not start new ones. Search for the buying question you want to win and look for threads that already rank. Adding a genuinely useful comment to an established thread is far more effective than creating a new post nobody sees.
- Disclose affiliation every time. It costs you nothing in credibility if the contribution is useful, and it is the difference between a durable mention and a ban.
We wrote the full playbook for this, including the moderation rules that trip most companies up, in our Reddit guide for SaaS.
3. Comparison and roundup publishers
The independent blogs, niche media sites and agency publications that write "best X for Y" articles in your category.
Why they matter
These pages exist because they match buyer intent exactly, which is also why engines retrieve them. Unlike review platforms, they are editorial, which means a human decides who is on the list. That makes them addressable in a way an algorithmic ranking is not.
How to get onto them
- Find the ones already being cited. Your citation tally from the method above is the target list. Do not guess at publications, use the ones the engines actually quote.
- Lead with a reason to update, not a request for a favour. "Your 2025 roundup lists a tool that shut down and is missing two current options" is a genuinely useful email. "Please add us" is not.
- Offer the thing only you have. Verified pricing, a feature comparison you have actually tested, a customer example. Editors update pages when the update is easy for them.
- Accept partial wins. Being added as an honourable mention still puts your brand name inside a document the engine retrieves.
4. Knowledge infrastructure: Wikipedia and Wikidata
The reference layer that answer engines use to decide what an entity is.
What it actually does for you
Wikipedia rarely decides who wins a "best tool for X" question. What it does is establish that your company exists, what category it belongs to, and what it is associated with. That is entity grounding, and without it the engine is less confident about you, which quietly costs you inclusion in answers you would otherwise be eligible for.
The honest advice
Do not try to create a Wikipedia article for a company that does not meet notability requirements. It will be deleted, and the attempt can leave a public record that does you no favours. If you genuinely qualify, meaning independent significant coverage exists, engage properly and disclose conflicts of interest.
Wikidata is more accessible and underused. A correct, complete Wikidata entry is a legitimate and low-friction way to strengthen the machine-readable picture of your company.
5. Professional networks, mainly LinkedIn
LinkedIn shows up constantly in B2B answer citations, and most companies treat it purely as a distribution channel rather than a source.
How to make it work as a citation source
- Your company page description is entity data. Make it identical in substance to your homepage description. Inconsistency here is a common and invisible own goal.
- Long-form posts from named experts outperform brand posts. Engines weight individual expertise, and a founder or head of product writing substantively is a stronger signal than a company account.
- Write posts that answer the buying question. A post titled with the exact question your buyers ask has a real chance of being retrieved, because it matches the query semantically and sits on a very high authority domain.
6. Video, mainly YouTube
Video is consistently among the most cited sources across AI answers, and it is the category most B2B SaaS teams skip entirely.
Why it works
A product walkthrough or comparison video has a transcript, and that transcript is text the engine can retrieve. It also carries an implicit credibility that a landing page does not, because someone actually demonstrated the thing.
The minimum viable version
You do not need a content studio. Three assets cover most of the ground: a product walkthrough, an honest comparison against your main competitor, and a use-case demo for your strongest vertical. Write real descriptions and let the transcripts be indexed. The bar in most B2B categories is low.
7. Developer and technical documentation
If you have an API, your docs are a citation asset. Stack Overflow answers, GitHub repositories and your own documentation get retrieved for technical questions in a way marketing pages never do.
The pattern is straightforward. Documentation is specific, current, and unambiguous, which is exactly what a retrieval system wants. If your docs are behind a login or rendered in a way crawlers struggle with, you have removed yourself from every technical question in your category.
8. Trade press and industry publications
Sector publications carry disproportionate weight for regulated and enterprise categories, where the engine is looking for a source that will not get it in trouble.
This is slow work and it is real digital PR: original data, a genuine point of view, or a customer story with numbers. It is the most expensive item on this list per mention, which is why it should not be first.
9. Your own website, last on purpose
Your site is not where you get discovered. It is where the engine confirms the specifics after something else has named you: pricing, integrations, who the product is for, what it does not do.
What actually matters on your own domain
- Be indexed in Bing, not just Google. Copilot is Bing grounded. Plenty of B2B sites are strong in Google and thin in Bing because nobody checked. It is free to fix.
- Answer the buying question directly on the page. A clear, quotable paragraph that answers the question outright is more retrievable than a beautifully written page that circles it.
- Publish the specifics competitors hide. Real pricing, real limitations, real integration lists. Specificity is retrievable. Vagueness is not.
- Keep the entity description identical everywhere. Homepage, about page, LinkedIn, review profiles. This is the cheapest fix on the whole list.
The order to actually do this in
The nine categories above are ranked by leverage, not by sequence. Leverage and effort are different things, and doing this in the wrong order is how teams spend two quarters and move nothing.
Weeks one to two: the free entity work
Claim every review profile. Make the company description identical across your homepage, about page, LinkedIn and every review platform. Verify your key commercial pages are indexed in Bing.
None of this costs money and all of it raises the floor. An engine that is confident about what you are will include you in more answers without you publishing a single new page.
Weeks three to six: the roundup queue
Work the list of comparison articles and roundups that your citation tally showed are actually being retrieved. Being added to an existing, already-cited page is the highest return per hour available in this entire discipline, because the page has already proven it gets retrieved.
Expect a low hit rate and work the volume. Twenty well-researched, genuinely useful emails will typically produce a handful of updates, and each one is permanent.
Weeks six to twelve: reviews and community
Run a real review drive. Start participating properly in the two or three communities where your buyers actually are. Both are slow-compounding and neither can be sprinted, which is why they start early and never stop.
Quarter two onward: the assets
Video, original data, trade press. These are the expensive items. They are worth doing, and they are worth doing after the cheap items are done, because they take months to pay back and the cheap items take weeks.
What does not work
An honest list, including things we have watched teams spend real money on.
Publishing more blog posts on your own domain
This is the default response and it is mostly wrong. If the engine is retrieving G2, two roundups and a Reddit thread for your buying question, a fourteenth blog post on your own site does not enter that retrieval set. It might help the confirmation layer. It will not get you named.
Chasing a universal top 50 list
Every AI citation study produces a list topped by the same few giant domains. It is interesting and it is not actionable, because you are not getting your B2B software into a general encyclopaedia entry or a mass-market news site this quarter. The list that matters is the one generated from your own category's questions.
Paid placements in low-quality roundups
There is a cottage industry selling listicle placements. Most of those pages are not retrieved by anything, because they are thin, unlinked and written for nobody. Before paying, check whether the page actually appears in your citation tally. Usually it does not.
Stuffing schema and hoping
Structured data helps machines parse what is already there. It does not create authority. Schema on a page nobody cites is a well-labelled document nobody reads.
Optimising for one engine
Teams sometimes build a whole programme around ChatGPT because it is the one they use personally. The retrieval sets differ between engines, and in B2B, Copilot behaves differently again because it is Bing grounded. Sample across engines before concluding anything. Our roundup of Microsoft Copilot rank trackers covers which tools actually let you do that, and which claim to and do not.
How this differs by SaaS category
The nine categories above are the general B2B shape. The weighting shifts meaningfully by vertical, and getting this wrong wastes the first quarter.
Developer tools and infrastructure
Documentation, GitHub and Stack Overflow move to the top, and review platforms drop. Developers do not read G2 and the engines have learned that. Your docs being crawlable and specific is the single highest leverage item.
HR, finance and other operations software
Review platforms dominate more heavily than in any other category, because the buyer is a non-technical evaluator who relies on peer signals. G2 and Capterra category placement is close to the whole game.
Regulated and enterprise categories
Trade press, analyst coverage and government or standards-body references carry disproportionate weight, because the engine is more conservative when the question has compliance implications. This is the one vertical where slow, expensive digital PR genuinely belongs early in the sequence.
Marketing and sales tools
The most crowded and most contested set of sources, with an enormous volume of comparison content and an unusually active community layer. Differentiation on a narrow use case matters more here than in any other vertical, because the head terms are unwinnable.
How to measure whether any of this is working
The thing to avoid is measuring your own website, which is what most analytics setups do by default.
Track the shortlist rate, not the mention rate
Of the buying questions your customers actually ask, in what percentage are you named at all. It is honest, it is unforgiving, and unlike share of voice it maps onto something the business already understands.
Exclude branded prompts from it. Asking an engine about your own company by name produces a flattering number that measures nothing.
Track source coverage as the leading indicator
Shortlist rate is a lagging indicator. It moves months after the work. The leading indicator is source coverage: of the fifteen domains that decide your category, on how many do you now appear.
That number moves within weeks, which matters, because it lets you show progress before the outcome metric catches up. It is also directly controllable, which the outcome metric is not.
Re-run the citation tally quarterly
The retrieval set changes. Publications update, threads die, new comparison pages get written. A list built twelve months ago is describing a category that no longer exists. Quarterly is enough.
For benchmarks on what movement looks like across categories, our State of AI Search Visibility study has the measured version.
The entity audit: twelve things to check before you spend anything
This is the free work, and it is the work most teams skip on the way to buying something. Run through it in an afternoon.
- Is your one-line company description identical on your homepage, about page, LinkedIn, G2 and Capterra? Not similar. Identical in substance.
- Does that description name your category in the words buyers use? If you describe yourself as a platform for revenue intelligence and your buyers search for sales forecasting software, the engine has to make a leap it may not make.
- Are your key commercial pages indexed in Bing? Check, do not assume. This is the most common silent failure in B2B.
- Is your pricing page public and specific? "Contact us" is not retrievable. A page with numbers on it is.
- Do you have a claimed profile on every relevant review platform? Including the ones you think are irrelevant.
- Are you in the narrow subcategories, not just the broad one? The subcategory names often match buyer questions word for word.
- Is your most recent review from the last 90 days? Recency is a quality signal.
- Does your site say plainly who the product is not for? Counterintuitively, stated limitations increase retrievability, because they are specific and unusual.
- Do you have a Wikidata entry, and is it correct? Low effort, genuinely useful, almost universally ignored.
- Are your integrations listed as text, not just logos? A wall of logos is invisible to a language model. A list of names is not.
- Does a named human appear as the author on your substantive content? Expertise attaches to people more readily than to brands.
- Is anything important locked behind a form or a login? Gated content cannot be cited, which means it cannot help you get recommended.
Most companies fail four to six of these. Fixing them costs nothing but attention, and it raises the ceiling on everything else you do afterwards.
The outreach that actually gets you added to a roundup
Getting added to an existing, already-cited comparison page is the single highest return activity in this discipline. It is also where most outreach fails, because it is written as a favour request.
Pick the target properly
Only pitch pages that appeared in your citation tally. A roundup that no engine retrieves is not worth a single email, no matter how good the domain metrics look. This one filter eliminates most of what a link building agency would put in front of you.
Lead with the update, not the ask
The editor's job is keeping the page accurate and current. Give them a reason that serves that job.
Things that work: a tool on the list has shut down or been acquired, a price on the list is wrong, the list is missing a category of option entirely, a claimed feature no longer exists. Things that do not work: "we would love to be featured".
Do the editor's work for them
Send the entry written, in their format, at their length, with the specifics they would otherwise have to research. Verified pricing, engine or feature coverage, who it is best for, and an honest limitation. That last part matters more than people expect. An entry that includes a genuine downside reads as credible and gets used.
Expect a low hit rate and work volume
A well-targeted campaign might convert one in five. That is a good outcome, because each win is a permanent placement inside a document that the engines have already demonstrated they retrieve.
Offer reciprocity honestly
If you publish comparison content yourself, a straightforward link exchange between genuinely relevant pages is a reasonable trade. Keep it relevant and keep it modest. Bulk swapping across unrelated sites is a different activity with a different risk profile.
How to run a review drive without annoying your customers
Reviews are the highest leverage asset in most B2B categories and the request is universally dreaded. A few things make it work.
Ask at the moment of demonstrated value
Not at renewal, when the customer is thinking about cost. Ask after a support interaction they rated well, after an onboarding milestone, or after they tell you something worked. The trigger matters more than the wording.
Ask individuals, not accounts
A named person asking a named person converts several times better than a broadcast email from a no-reply address. This is unglamorous and it is the whole difference.
Make it specific
"Would you leave us a review" produces generic five-star reviews that say nothing and are worth little to an answer engine. "Would you mention how the migration went, since that is what people ask us about most" produces a review that contains the exact language buyers search with.
That is the real objective. You are not collecting stars, you are collecting sentences that describe your product in the words your buyers use, on a domain the engines trust.
Never incentivise in ways that violate the platform
Every major review platform has rules about incentives, and violating them can get reviews removed retroactively. Read the specific policy. A generic gift card campaign is a good way to lose a year of accumulated reviews.
Keep it always on
A burst of reviews followed by silence is a worse signal than a steady trickle. Build the ask into a recurring process rather than a quarterly campaign.
What the first 90 days usually produces
Setting expectations honestly, because this is where most programmes get abandoned.
What moves in weeks
Source coverage. Profile completeness, entity consistency, Bing indexation and a handful of roundup additions are all achievable inside a quarter, and they are all directly controllable.
What moves in months
Shortlist rate. The engines need to re-crawl the pages you have changed, and the retrieval sets need to shift. Expecting the outcome metric to move in six weeks is the single most common cause of a programme being cancelled just before it would have worked.
What does not move at all
Head terms in a category dominated by incumbents with a decade of accumulated references. If you are a challenger, the constrained questions are where you win, and the honest version of this work involves accepting that rather than fighting the head term for two years.
The realistic first win
For most B2B SaaS companies the first visible change is being named in constrained, use-case-specific questions where the field is thinner. That is not a consolation prize. Those questions convert better than the head term, because the buyer has already told the engine exactly what they need.
A starter checklist of domains to check first
Your real list comes from your own citation tally. But if you want somewhere to start this afternoon, these are the domain types worth checking your presence on before you have run anything. Check each one for whether a competitor appears and you do not.
- Review platforms: G2, Capterra, TrustRadius, Software Advice, GetApp, and any vertical-specific review site your category has.
- Community: the two or three subreddits where your buyers actually post, plus Hacker News or Stack Overflow if you sell to technical audiences.
- Comparison publishers: whichever independent blogs currently rank for "best [your category]" and "[your main competitor] alternatives".
- Reference: Wikidata, and Wikipedia only if you genuinely meet notability.
- Professional: your LinkedIn company page and the personal profiles of your two or three most credible subject matter experts.
- Video: YouTube, including the comparison videos other people have made about your category.
- Technical: your own documentation, GitHub, and any developer community relevant to your product.
- Trade: the two or three publications your buyers actually read, not the ones with the best domain metrics.
Eight buckets, maybe twenty five domains. That is a tractable audit, and it will tell you more about why AI does not recommend you than any dashboard will.
Where an agency fits, and where it does not
Most of this list is work you can do yourself, and you should be suspicious of anyone who tells you otherwise. The entity audit is free. Claiming review profiles is free. Running the citation tally costs an afternoon.
What tends to justify outside help is volume and consistency: working a forty item roundup queue every month, running an always-on review programme, participating credibly in communities week after week, and producing the original data or assets that earn trade coverage. Those fail on capacity, not on knowledge.
That is the work Arobis AI does as a done-with-you programme for B2B SaaS, and our pricing is public so you can judge whether it is worth outsourcing rather than hiring. If you want to understand the discipline before deciding, answer engine optimization and generative engine optimization explain the two halves of it, and how to choose an AI SEO agency covers what to ask any provider, us included.
Before any of that, get a baseline. The AI visibility checker scores how ready your site is to be cited, free and without a signup.
Frequently asked questions
Which domains are most influential in AI search?
For B2B software, review platforms and community threads carry the most weight, followed by independent comparison publishers. Reference sites like Wikipedia matter for establishing what your company is rather than for winning category questions. The universal lists topped by giant general domains are real but not actionable for a B2B SaaS company.
How do I find which domains influence my own category?
Run 30 to 40 constrained buyer questions through ChatGPT, Google AI Mode, Perplexity and Copilot, three times each, and tally the cited domains rather than the named brands. The top ten to fifteen by frequency are the domains that decide your category.
Why does my own website not get cited?
Because answer engines retrieve sources that describe a category from the outside. Vendor sites are treated as claims rather than evidence. Your site confirms specifics like pricing and integrations once something else has named you, which makes it necessary but not sufficient.
Is Reddit really that important for B2B?
Yes, disproportionately, because answer engines favour sources that read like genuine practitioner experience. The caveat is that Reddit punishes marketing behaviour severely, so the approach has to be real participation with disclosed affiliation rather than campaign activity.
Do backlinks still matter for AI search?
Indirectly. A link is useful mainly because it usually means you are mentioned on a page that the engine may retrieve. The mention inside a retrievable document is doing the work, not the link attribute. That is why an unlinked mention in a frequently cited roundup can be worth more than a followed link nobody retrieves.
How long does this take to show results?
Source coverage moves within weeks. Shortlist rate, meaning how often you are named at all, typically takes months, because pages must be re-crawled and retrieval sets have to shift. Anyone promising movement in the outcome metric inside six weeks is overselling.
Should I create a Wikipedia page for my company?
Only if you genuinely meet notability requirements, meaning substantial independent coverage already exists. Attempts that do not qualify get deleted and can create a public record you would rather not have. Wikidata is the more accessible and more commonly neglected option.
Does this differ between ChatGPT, Gemini, Perplexity and Copilot?
Yes. Retrieval sets differ between engines, and Copilot differs most because it is grounded in Bing rather than Google. If your pages are thin in Bing you can be invisible in Copilot while performing acceptably elsewhere.
Are AI search rankings the same as Google rankings?
No. There is no ordered list to rank in. You are either named in a written answer or you are not, and the answer is assembled from retrieved sources. Ranking well in Google helps because it increases the chance your pages are retrieved, but the two are not the same measurement.
What is the single highest leverage thing to do first?
For most B2B SaaS companies, claim and complete every review platform profile and make your company description identical everywhere. It is free, it takes an afternoon, and it raises the ceiling on everything else you do afterwards.
How often should I redo the citation tally?
Quarterly. Publications update their roundups, threads die, and new comparison pages appear. A source list built a year ago describes a category that has already moved.
Can I just buy placements on the domains that matter?
Sometimes, and it is usually a poor trade. Paid listicle placements are frequently on pages that no engine retrieves. Before paying for any placement, confirm the page actually appeared in your own citation tally. Most do not.
Related reading
- How ChatGPT, Perplexity, Claude and Gemini choose which brands to mention, the mechanism behind everything on this page.
- The best free AI visibility tools, if you want to automate the citation tally without a budget.
- The best Profound alternatives, for the paid tools that track sources rather than only scores.
- The full AI visibility tools roundup, for the wider category.



