GEO Guide

How to improve GEO and AI Overview visibility

A practical guide to getting cited by Google AI Overviews, ChatGPT, Gemini and Perplexity — for marketers running their own site or an agency running many.

Jordy De Rosa · Founder, Bi-Rank · September 2026 · 12 min read
Google AI Overview ChatGPT Perplexity Gemini

For twenty years the job was simple to describe: rank on page one, get the click. That contract is breaking. Google now answers a growing share of queries directly in an AI Overview, and a growing share of people never open Google at all: they ask ChatGPT, Gemini or Perplexity and take the answer they get.

The answer is still built from web pages. Somebody's page. The question is whether it is yours.

That is the whole of Generative Engine Optimization (GEO): making your content the thing an AI engine picks up, quotes and links when it answers a question in your field. Being ranked used to be the goal. Being cited is the new ranking.

This guide covers what actually moves the needle, what does not, and how to measure it, so you can run it for your own site or, if you are an agency, across every client you manage.

What "GEO" means, and how it differs from SEO

SEO gets a page into an index and up a list of ten blue links. GEO gets a page into the answer.

The two overlap heavily. Google AI Overviews are built on the same index and largely the same ranking signals as classic search. ChatGPT's browsing uses Bing's index. Perplexity runs its own crawler but leans on conventional search rankings to pick candidates. If you are not crawlable, indexable and reasonably ranked, you are not in the running for any of them.

But ranking is only the entry ticket. Once an engine has a set of candidate pages, it does something search never did: it reads them, extracts the passages that answer the question, and decides which ones to trust and quote. That step rewards a different set of qualities:

Ranking gets you considered. Those four get you cited.

First, understand the three ways your brand can show up

Before you optimize anything, be precise about what you are optimizing for. AI engines can surface a brand in three distinct ways, and they are not worth the same.

Mention

Your brand name appears in the answer text — "Popular options include Acme, Beta and Gamma." No link, no traffic, but you are on the shortlist.

Citation

One of your pages appears among the answer's sources: a link in the AI Overview's carousel, a footnote in Perplexity, a source chip in ChatGPT. This is the valuable one — it sends visits and tells the user the engine trusts you.

Recommendation

The engine names you as the answer to a commercial question ("best X for Y"). Rare, high value, and mostly a function of how consistently third-party sources describe you.

When you track results, track them separately. A brand with high mentions and zero citations has a reputation problem to fix on its own pages. A brand with citations but no mentions has content that is useful but no entity identity around it. The fix is different in each case.

The value gap between the three is not hypothetical. Seer Interactive's 2026 analysis of 53 brands across 5.47 million tracked queries (January 2025–February 2026) found that being cited in an AI Overview lifts organic click-through by about 120% per impression compared with the same query when the brand goes uncited: 2.07% average CTR when cited versus 0.94% when not. A mention with no citation does not move that number.


Part 1 — Make your content extractable

This is where most of the leverage is, and most of it is within your control.

Lead with the answer

Every page that targets a question should answer it in the first 40–60 words, in plain declarative sentences, before any context or caveats. Not "In this article we will explore…" but the actual answer.

Not citable

Choosing the right CRM for a small business can be a daunting task, and there are many factors to consider.

Citable

A small business typically needs a CRM with contact management, a shared pipeline, email integration and pricing under about €30 per user per month. The three most common choices in that bracket are [A], [B] and [C].

The second version can be lifted verbatim into an answer. The first cannot. Models are trained to prefer passages that stand on their own; the "citability threshold" is, in practice, whether a paragraph makes sense with the rest of the page removed.

One question, one heading, one answer

Structure the body as a series of questions people actually ask, each as an H2 or H3, each followed immediately by a direct answer, then supporting detail. This is not a gimmick: it mirrors how retrieval works. Engines chunk pages by heading and score chunks against the query. A heading that is the query, followed by a paragraph that is the answer, is the easiest possible match.

Use the exact phrasing people use. "How much does X cost?" beats "Pricing considerations." "Is X safe for Y?" beats "Safety profile."

Use lists, tables and definitions where the content has that shape

Steps, comparisons, specifications, prices, and pros/cons are extracted far more reliably from a list or a table than from prose. AI Overviews in particular lean on ordered and unordered lists for "how to" and "best of" queries.

Do not force it. A numbered list of one-word bullets is noise. A list where each item is a complete, specific statement ("Enable two-factor authentication on the admin account") is an asset.

Define your terms explicitly

If your page is about a concept, include a sentence of the form "X is [definition]." Engines look for definitional patterns when answering "what is" queries, and a clean definition is one of the most frequently quoted structures in AI Overviews.

Keep paragraphs short and self-contained

Three to five sentences. One idea each. Avoid pronouns that point back to a previous paragraph ("this approach", "it") because a chunk that starts with "This approach works because…" is meaningless in isolation and will be skipped.

Cut the filler

Introductions that restate the title, "in today's fast-paced world," paragraphs that promise what the next section will say: all of this dilutes the density of extractable statements per page. Engines do not reward length. They reward the ratio of useful, specific, verifiable statements to total words.


Part 2 — Give engines a reason to trust you

Extractable content gets you considered. Trust gets you chosen over the other extractable pages on the same topic.

Publish something nobody else has

The single most reliable way to be cited is to be the original source of a fact. Engines, and the humans who write the pages engines learn from, cite the origin.

Practical ways to do that at agency or SMB scale, without a research department:

Put the numbers in the text, not only in an image or a PDF. Models read text.

Show who wrote it and why they are qualified

E-E-A-T (experience, expertise, authoritativeness, trust) is a Google quality framework, not a ranking factor you can toggle, but every generative engine is trying to solve the same problem: "can I safely repeat what this page says?" Make it easy:

Anonymous content by "Admin" on a page with no date is exactly what an engine is trained to be cautious about.

Be consistent about who you are, everywhere

Engines build a picture of an entity from every place it appears: your site, your LinkedIn page, directories, reviews, press, Wikipedia if you have it, partners' sites. If those sources disagree on what you do, where you are, or what you are called, the entity is fuzzy and fuzzy entities do not get recommended.

Audit for consistency: same brand name (including capitalization and hyphenation), same one-line description, same address and category, on every profile you control. Then ask which third-party pages the engines are already citing for your topic (see Part 4) and get listed there, accurately.

Earn mentions, not just links

For classic SEO the link was the unit of authority. For generative engines, the mention in context matters at least as much: a directory that lists you as a provider in your category, a comparison article that includes you, a forum thread where someone recommends you by name. ChatGPT and Perplexity in particular frequently synthesize "best X" answers from exactly those pages.

This is digital PR, but with a narrower brief: get named in the specific pages that already show up as sources for your target questions.


Part 3 — Get the technical foundations right

None of this works if the engines cannot read you. The checklist is short, but each item is a hard gate.

Let the AI crawlers in

Each engine uses its own user agent, and a blanket block will silently remove you from everything:

Crawler user agents by engine
Engine / useUser agentNotes
Google AI Overviews and AI Mode Googlebot Same crawler as search. Blocking Google-Extended does not remove you from AI Overviews; it only opts out of Gemini model training.
ChatGPT browsing / search OAI-SearchBot Separate from GPTBot (training). Block GPTBot if you want; keep OAI-SearchBot allowed if you want to be cited.
Perplexity PerplexityBot Perplexity's own index.
Claude ClaudeBot / Claude-User Anthropic's crawlers for training and user-initiated fetches.
Bing (feeds ChatGPT, Copilot) Bingbot Do not neglect Bing: it is the index behind ChatGPT search.

Check your robots.txt, your CDN's bot rules and your WAF. Many sites block "AI bots" wholesale in 2024 and forgot. If your goal is to be cited, that decision is now working against you.

Publish an llms.txt

llms.txt is a plain-text file at your site root that tells language models what the site is about and which pages matter most. It is an emerging convention, not a standard every engine has committed to, and it does not directly influence AI Overviews. But it costs ten minutes, it is a clean summary of your entity for any crawler that reads it, and it forces you to decide which pages are your canonical answers.

Serve content as HTML, not behind JavaScript

Googlebot renders JavaScript. Most AI crawlers do not, or do so inconsistently. If your key content only appears after client-side rendering, assume the engines other than Google see an empty page. Server-side render or pre-render anything you want cited.

Use schema.org markup, for the right reasons

Structured data does not make an engine cite you. It makes the engine less likely to misread you. That is still worth doing:

Keep it truthful. Marking up an FAQ that is not on the page is a fast way to lose the trust you are trying to build.

Speed and stability

AI crawlers have tight fetch budgets. A page that takes six seconds or intermittently returns 5xx will be sampled less often and refreshed less often. Nothing new here: fix Core Web Vitals, fix server errors.


Part 4 — Optimize per engine, because they are not the same

"AI visibility" is four different games with overlapping rules. Treat them separately.

Google AI Overviews and AI Mode · Googlebot

ChatGPT (with search) · OAI-SearchBot

Perplexity · PerplexityBot

Gemini · grounded on Search


Part 5 — Measure it, or you are guessing

Most GEO advice stops at "do these things." The difficult part is knowing whether they worked, because AI answers are non-deterministic, personalized, and change week to week.

A workable measurement process:

  1. Define a prompt set. Twenty to fifty questions that a prospective customer would actually ask an AI about your category: informational ("how to choose a…"), comparative ("X vs Y"), and commercial ("best … in Milan"). Include the questions where competitors currently win.
  2. Run every prompt on every engine, repeatedly. One run tells you nothing. Several runs per prompt, spread over days, give you a share of voice: the percentage of answers where the brand is mentioned or cited.
  3. Separate mentions from citations. Track both. Citations are the number to grow.
  4. Record who is cited. The domains that keep appearing as sources for your prompts are your real competition, and often your best outreach targets: a directory or comparison page you can get listed on.
  5. Score the pages you control on citability, before and after changes, with a fixed set of checks (answer-first opening, heading structure, definitions, author and date, schema, crawler access). A deterministic score is the only way to see whether a rewrite actually improved the page and not just your opinion of it.
  6. Re-run monthly, at minimum. Engines change their behavior faster than search ever did.

Doing this by hand for one brand is a few hours a month. Doing it for fifteen clients on four engines is a job, and it is the reason tools exist.

How Bi-Rank fits in

Bi-Rank is built for exactly step 5 and steps 1–4 above, for agencies that manage several clients.

  • GEO Score (0–100). A deterministic, repeatable score per page and per site, based on a fixed set of extractability, authority and technical checks. It tells you where a page sits against a citability threshold and which issues would move it, so you can prioritize and show a client a before/after that is not a matter of opinion.
  • Multi-engine tracking. You define the prompt set; Bi-Rank runs it on ChatGPT, Gemini, Perplexity and Google AI Overviews, repeatedly, and reports share of voice, mentions versus citations, sentiment, and which competitor domains are being cited instead of you.
  • Technical foundations check. Whether AI crawlers can read the domain, whether llms.txt and schema.org are in place, re-checked weekly.
  • White-label. The reports carry your agency's brand and domain, with read-only client access.

It does not write your survey or earn your mentions. It tells you, with numbers, whether what you did worked.

Start a free GEO scan

Conclusion

Generative engines have not replaced search; they have added a second filter on top of it. Ranking still gets you into the room. Being clear, specific, verifiable and consistently described is what gets you quoted.

Start with the pages that already rank for questions in your category, rewrite their openings so they answer the question outright, add an author and a date, publish one piece of data that is genuinely yours, and check that the AI crawlers are not blocked. Then measure: same prompts, every engine, repeated, tracked over time.

If you want to see where a site stands today, run it through Bi-Rank and look at the score and the prompts before changing anything. The baseline is the part everyone skips.