Skip to main content
Auto Alpha AdvisoryAuto Alpha Advisory
blog

832 AI Answers: What Actually Gets a Business Cited

Matt Owen · 26 July 2026 · 9 min read

832 AI Answers: What Actually Gets a Business Cited

When a buyer asks an AI assistant which company to use, they get a shortlist of three or four names. There is no page two. Either the assistant names you or the buyer never learns you exist. Most businesses have never checked which side of that line they are on, and the ones who do check usually check wrong.

I'm a CA(SA) who builds AI systems and measures AI visibility for a living, so rather than theorise about it I ran the measurement. In June 2026 I put real buyer questions to the four assistants people actually use — ChatGPT, Claude, Perplexity and Gemini — across five South African industries, and read back every source each one cited. Not what the assistants said. What they read to say it.

That produced 832 grounded answers and 5,459 individual citations. Five findings came out of it. Some of them contradict what the market is selling.

What I measured, and what this is not

Five verticals: payments, residential solar, advisory, insurance and online education. Four engines. Prompts written the way buyers actually phrase them, all specifying South Africa or naming SA cities. Data collected 5 June 2026.

Two honest caveats before the findings. First, this is a five-vertical snapshot of 832 answers — the directions are consistent and strong, but the exact percentages will drift, and I re-measure quarterly. Second, this is a different study from the 600-answer analysis of property, banking and crypto I published in June. That one asked which brands get named. This one asks which sources get read. Different samples, different questions, different dates — read them as siblings, not as one number restated.

The underlying idea has an academic foundation. The term GEO comes from a Princeton-led paper presented at KDD 2024, where Pranjal Aggarwal and colleagues showed "that GEO can boost visibility by up to 40% in generative engine responses." What follows is what that looks like in a real market.

Do all AI assistants cite sources equally?

No, and this is the finding that changes how you measure yourself.

Engine Answers citing any source Avg sources per cited answer
Gemini 208 9.8
Perplexity 208 7.4
Claude 153 10.8
ChatGPT 50 4.9

ChatGPT grounded on roughly a quarter as many answers as Gemini or Perplexity, and pulled about half as many sources when it did. It leans on trained memory. The others lean on live retrieval.

The practical consequence is uncomfortable: almost everyone checks their AI visibility by typing their brand into ChatGPT. That samples the least grounded engine, and it systematically understates what your real position is. One surface lies. Four surfaces measured tell you the truth.

It also means "optimising for AI" is two different jobs. ChatGPT visibility is won slowly, by being written about widely enough to enter the model's memory. The retrieval engines are won now, by being crawlable and citable today. Any proposal that doesn't say which of those it's buying you is selling a blur.

Why do South African questions get answered by foreign websites?

Because a third of the sources aren't local. Across all 5,459 citations — on prompts that all named South Africa or an SA city — the split was:

  • .co.za: 34%
  • .com / .org / .net: 57%
  • everything else: 9%

US authority sites turned up repeatedly as sources for South African buyer questions. Nerdwallet, Bankrate, the US Chamber of Commerce and several American online-schooling sites all appeared, deciding answers for SA buyers.

Ranking on the local web is necessary and not sufficient. On your own local queries you are competing against international content the engines already trust. The counter is authoritative, genuinely SA-specific content — and because most local competitors haven't built it, the gap is winnable rather than hopeless.

Which single website gets cited most often?

YouTube. Not a competitor, not a directory, not a publication.

youtube.com was cited in 66 answers — more than any brand, aggregator or publication, and it appeared across all five unrelated verticals. Video is the most underweighted surface in AI visibility right now. Almost no audit tooling treats it as a citation channel at all.

For most businesses, a small library of videos that directly answer buyer questions — properly transcribed and indexed — is disproportionately valuable for citation, and almost nobody local is doing it.

Comparison sites decide the shortlist before your brand does

Every vertical had a mediator sitting above the brands in citation volume:

Vertical Dominant cited mediator Brand-side reality
Insurance hippo.co.za (41 answers) brands' own sites cited far less
Solar energybee.co.za, solar.co.za comparison pages set the shortlist
Moving / logistics wisemove.co.za (29)
Payments eezipay.com (27) international content fills the gap
Education US online-school sites lead local providers trail
Every vertical youtube.com (66), reddit.com (21) universal

This is why the "just build authority" advice underdelivers. The engine often isn't reading brand sites at all for the shortlist question — it's reading the comparison page that ranks the brands. There are two moves: get well represented on the mediators the engines already read, or become one by publishing the comparison content yourself.

What kind of page actually gets cited?

This is the finding I'd act on first, because it's the most specific.

Own-site citation tracked share-of-voice almost exactly. The leaders in each category had 27–48 citations to their own domain. The laggards had between one and nine — and in the worst case, that single citation was the homepage.

But the mechanism is precise, and it is not "publish more content". Reading the actual cited URLs, homepages get cited a little for everyone. That's table stakes. What separates the winners is a deep page that directly answers one specific buyer question, in one of two shapes:

  • Service pages named for the thing itself. Sable International earns citations on pages like /south-african-tax/tax-emigration-from-south-africa — the page is named for the question.
  • Articles shaped exactly like the question. GoSolr earns them on /enlighten/articles/subscription-vs-ownership and /solar-system-prices.

Ask an engine about subscription versus ownership solar and it cites the article called subscription-vs-ownership. It is question-matching, not content volume. Generic company news posts were cited essentially never.

The best example in the whole dataset: the University of Cape Town's strongest single page is a blog post reviewing twelve leading online high schools, cited eight times. They wrote the category comparison, so the engines cite them as the neutral authority on their own competitive set. That's the "become a mediator" move, caught in the wild.

So is schema markup the answer?

No — and this is where the market's advice and the evidence part company. SE Ranking's study of 129,000 ChatGPT citations found "pages with FAQ schema average 3.6 citations, while those without reach 4.2. This means that schema alone does not significantly increase ChatGPT citation likelihood." My own data agrees: the brand with the most structured data in one vertical sat mid-table.

Schema is hygiene. Do it because it's cheap and correct, not because it's the lever.

What does correlate, in the same study, is freshness — content updated within three months averaged 6 citations against 3.6 for stale pages. And SE Ranking's separate analysis of 100,013 keywords tracks how quickly the AI Overview surface itself keeps shifting.

Why being cited is worth the work

Two independent numbers make the business case.

First, the traffic that used to arrive by clicking is going away. Searches that trigger an AI Overview now show an 83% zero-click rate, against roughly 60% for searches without one.

Second, citation is what recovers some of it. Seer Interactive's study of 3,119 search terms across 42 organisations found brands cited in an AI Overview get 35% higher organic click-through than when they aren't. Seer is careful about what that proves: "We cannot definitively prove that citation causes higher CTRs, it's equally possible that brands with stronger authority and higher baseline CTRs are simply more likely to be cited by Google's AI." Either way, the correlation is strong enough to act on.

And ranking is not a substitute. Semrush's study of 5,000 queries and over 150,000 citations found only 35% of the URLs cited in Google's AI Mode also appear in Google's own organic top 10 for the same query. You can hold position one and still be absent from the answer above it.

What to do with this

Five findings, five actions, in the order I'd take them:

  1. Measure all four engines, not ChatGPT. Your ChatGPT result is not your visibility.
  2. Check you're readable. Google's own documentation is explicit: "To be eligible to be shown as a supporting link in AI Overviews or AI Mode, a page must be indexed and eligible to be shown in Google Search with a snippet, fulfilling the Search technical requirements." If a crawler can't read the page, nothing else on this list matters. Why your website might be invisible to AI search covers the mechanics.
  3. Build one question-matched page per real buyer question. Named for the question, answering it in the first sentence. Not a blog.
  4. Get onto the mediators, or publish the comparison yourself.
  5. Put your answers on video while the surface is still uncontested.

None of this is exotic. It's the ordinary work of being the most useful, most readable answer to a question your buyer is actually asking — which is roughly what the technical GEO breakdown has been saying, now with the receipts.

If you want to know where your business currently sits, that's exactly what my nine-domain audit measures. Get my visibility audit — R1,490, one-time, with the weak results shown rather than smoothed.

Sources · 7
  1. GEO: Generative Engine Optimization — Aggarwal et al., KDD 2024
  2. AI features and your website — Google Search Central
  3. How Google's AI Mode compares to traditional search — Semrush, 5,000 queries
  4. AIO Impact on Google CTR, September 2025 update — Seer Interactive
  5. How to optimize for ChatGPT — SE Ranking, 129,000 citations
  6. Google AI Overviews research — SE Ranking, 100,013 keywords
  7. What is zero-click search — LLMrefs

Study data collected 5 June 2026 across 832 grounded answers and 5,459 citations in five South African verticals. Figures are reproducible from the run archive; methodology available on request. Percentages will drift — this is re-measured quarterly.

See where the machine reader stands on your site.

The free read runs nine domains, including GEO and content structure — the same discipline this writing is about.

← All writing