For twenty years the job was simple to describe: rank on page one, get the click. That contract is breaking. Google now answers a growing share of queries directly in an AI Overview, and a growing share of people never open Google at all: they ask ChatGPT, Gemini or Perplexity and take the answer they get.
The answer is still built from web pages. Somebody's page. The question is whether it is yours.
That is the whole of Generative Engine Optimization (GEO): making your content the thing an AI engine picks up, quotes and links when it answers a question in your field. Being ranked used to be the goal. Being cited is the new ranking.
This guide covers what actually moves the needle, what does not, and how to measure it, so you can run it for your own site or, if you are an agency, across every client you manage.
What "GEO" means, and how it differs from SEO
SEO gets a page into an index and up a list of ten blue links. GEO gets a page into the answer.
The two overlap heavily. Google AI Overviews are built on the same index and largely the same ranking signals as classic search. ChatGPT's browsing uses Bing's index. Perplexity runs its own crawler but leans on conventional search rankings to pick candidates. If you are not crawlable, indexable and reasonably ranked, you are not in the running for any of them.
But ranking is only the entry ticket. Once an engine has a set of candidate pages, it does something search never did: it reads them, extracts the passages that answer the question, and decides which ones to trust and quote. That step rewards a different set of qualities:
- Extractability. Can a model lift a clean, self-contained answer out of your page without rewriting it?
- Verifiability. Does the passage carry evidence (numbers, named sources, dates, authorship) that makes it safe to cite?
- Entity clarity. Does the engine know who you are, what you do, and that you are a legitimate source on this topic?
- Consistency. Is the same fact stated the same way across your site and across third-party sources the engine also reads?
Ranking gets you considered. Those four get you cited.
First, understand the three ways your brand can show up
Before you optimize anything, be precise about what you are optimizing for. AI engines can surface a brand in three distinct ways, and they are not worth the same.
Your brand name appears in the answer text — "Popular options include Acme, Beta and Gamma." No link, no traffic, but you are on the shortlist.
One of your pages appears among the answer's sources: a link in the AI Overview's carousel, a footnote in Perplexity, a source chip in ChatGPT. This is the valuable one — it sends visits and tells the user the engine trusts you.
The engine names you as the answer to a commercial question ("best X for Y"). Rare, high value, and mostly a function of how consistently third-party sources describe you.
When you track results, track them separately. A brand with high mentions and zero citations has a reputation problem to fix on its own pages. A brand with citations but no mentions has content that is useful but no entity identity around it. The fix is different in each case.
The value gap between the three is not hypothetical. Seer Interactive's 2026 analysis of 53 brands across 5.47 million tracked queries (January 2025–February 2026) found that being cited in an AI Overview lifts organic click-through by about 120% per impression compared with the same query when the brand goes uncited: 2.07% average CTR when cited versus 0.94% when not. A mention with no citation does not move that number.
Part 1 — Make your content extractable
This is where most of the leverage is, and most of it is within your control.
Lead with the answer
Every page that targets a question should answer it in the first 40–60 words, in plain declarative sentences, before any context or caveats. Not "In this article we will explore…" but the actual answer.
Choosing the right CRM for a small business can be a daunting task, and there are many factors to consider.
A small business typically needs a CRM with contact management, a shared pipeline, email integration and pricing under about €30 per user per month. The three most common choices in that bracket are [A], [B] and [C].
The second version can be lifted verbatim into an answer. The first cannot. Models are trained to prefer passages that stand on their own; the "citability threshold" is, in practice, whether a paragraph makes sense with the rest of the page removed.
One question, one heading, one answer
Structure the body as a series of questions people actually ask, each as an H2 or H3, each followed immediately by a direct answer, then supporting detail. This is not a gimmick: it mirrors how retrieval works. Engines chunk pages by heading and score chunks against the query. A heading that is the query, followed by a paragraph that is the answer, is the easiest possible match.
Use the exact phrasing people use. "How much does X cost?" beats "Pricing considerations." "Is X safe for Y?" beats "Safety profile."
Use lists, tables and definitions where the content has that shape
Steps, comparisons, specifications, prices, and pros/cons are extracted far more reliably from a list or a table than from prose. AI Overviews in particular lean on ordered and unordered lists for "how to" and "best of" queries.
Do not force it. A numbered list of one-word bullets is noise. A list where each item is a complete, specific statement ("Enable two-factor authentication on the admin account") is an asset.
Define your terms explicitly
If your page is about a concept, include a sentence of the form "X is [definition]." Engines look for definitional patterns when answering "what is" queries, and a clean definition is one of the most frequently quoted structures in AI Overviews.
Keep paragraphs short and self-contained
Three to five sentences. One idea each. Avoid pronouns that point back to a previous paragraph ("this approach", "it") because a chunk that starts with "This approach works because…" is meaningless in isolation and will be skipped.
Cut the filler
Introductions that restate the title, "in today's fast-paced world," paragraphs that promise what the next section will say: all of this dilutes the density of extractable statements per page. Engines do not reward length. They reward the ratio of useful, specific, verifiable statements to total words.
Part 2 — Give engines a reason to trust you
Extractable content gets you considered. Trust gets you chosen over the other extractable pages on the same topic.
Publish something nobody else has
The single most reliable way to be cited is to be the original source of a fact. Engines, and the humans who write the pages engines learn from, cite the origin.
Practical ways to do that at agency or SMB scale, without a research department:
- A small survey of your own customers or audience, even 100–200 responses, with the methodology stated, becomes a citable fact the moment you publish real numbers with a sample size and a date. Ahrefs did exactly this in September 2026: an analysis of the domains most cited across more than 3 million US queries found YouTube leading with a 22.9% mention share, ahead of Reddit (18.5%) and Facebook (10.1%), and even ahead of Google's own properties (8.8%). That single proprietary dataset is now something other sites quote and link to. You do not need Ahrefs' scale to do the same thing at your level; you need the number, the sample, and the date.
- Aggregated data from your own operations. Average ticket resolution time, typical project cost ranges, the most common issues you find in audits. Anonymize and publish.
- A benchmark or comparison you actually ran. Test five tools, publish the numbers and the setup.
- A yearly "state of" report for your niche. Even a modest one becomes the reference other pages link to, and links are still how engines discover authority.
Put the numbers in the text, not only in an image or a PDF. Models read text.
Show who wrote it and why they are qualified
E-E-A-T (experience, expertise, authoritativeness, trust) is a Google quality framework, not a ranking factor you can toggle, but every generative engine is trying to solve the same problem: "can I safely repeat what this page says?" Make it easy:
- A named author with a short bio and a link to a profile (LinkedIn, an author page, a professional register).
- A visible "last updated" date, and actually update it when you revise.
- First-person, specific experience where relevant: "In the 40 migrations we ran last year, the most common failure was…"
- Sources for any claim that is not yours, linked inline.
Anonymous content by "Admin" on a page with no date is exactly what an engine is trained to be cautious about.
Be consistent about who you are, everywhere
Engines build a picture of an entity from every place it appears: your site, your LinkedIn page, directories, reviews, press, Wikipedia if you have it, partners' sites. If those sources disagree on what you do, where you are, or what you are called, the entity is fuzzy and fuzzy entities do not get recommended.
Audit for consistency: same brand name (including capitalization and hyphenation), same one-line description, same address and category, on every profile you control. Then ask which third-party pages the engines are already citing for your topic (see Part 4) and get listed there, accurately.
Earn mentions, not just links
For classic SEO the link was the unit of authority. For generative engines, the mention in context matters at least as much: a directory that lists you as a provider in your category, a comparison article that includes you, a forum thread where someone recommends you by name. ChatGPT and Perplexity in particular frequently synthesize "best X" answers from exactly those pages.
This is digital PR, but with a narrower brief: get named in the specific pages that already show up as sources for your target questions.
Part 3 — Get the technical foundations right
None of this works if the engines cannot read you. The checklist is short, but each item is a hard gate.
Let the AI crawlers in
Each engine uses its own user agent, and a blanket block will silently remove you from everything:
| Engine / use | User agent | Notes |
|---|---|---|
| Google AI Overviews and AI Mode | Googlebot | Same crawler as search. Blocking Google-Extended does not remove you from AI Overviews; it only opts out of Gemini model training. |
| ChatGPT browsing / search | OAI-SearchBot | Separate from GPTBot (training). Block GPTBot if you want; keep OAI-SearchBot allowed if you want to be cited. |
| Perplexity | PerplexityBot | Perplexity's own index. |
| Claude | ClaudeBot / Claude-User | Anthropic's crawlers for training and user-initiated fetches. |
| Bing (feeds ChatGPT, Copilot) | Bingbot | Do not neglect Bing: it is the index behind ChatGPT search. |
Check your robots.txt, your CDN's bot rules and your WAF. Many sites block "AI bots" wholesale in 2024 and forgot. If your goal is to be cited, that decision is now working against you.
Publish an llms.txt
llms.txt is a plain-text file at your site root that tells language models what the site is about and which pages matter most. It is an emerging convention, not a standard every engine has committed to, and it does not directly influence AI Overviews. But it costs ten minutes, it is a clean summary of your entity for any crawler that reads it, and it forces you to decide which pages are your canonical answers.
Serve content as HTML, not behind JavaScript
Googlebot renders JavaScript. Most AI crawlers do not, or do so inconsistently. If your key content only appears after client-side rendering, assume the engines other than Google see an empty page. Server-side render or pre-render anything you want cited.
Use schema.org markup, for the right reasons
Structured data does not make an engine cite you. It makes the engine less likely to misread you. That is still worth doing:
- Organization (with sameAs pointing to your social and directory profiles) on the homepage, to anchor the entity.
- Article with author, datePublished, dateModified on editorial content.
- FAQPage where you genuinely have Q&A content; HowTo for step-by-step guides.
- Product with real offers and aggregateRating for product pages.
Keep it truthful. Marking up an FAQ that is not on the page is a fast way to lose the trust you are trying to build.
Speed and stability
AI crawlers have tight fetch budgets. A page that takes six seconds or intermittently returns 5xx will be sampled less often and refreshed less often. Nothing new here: fix Core Web Vitals, fix server errors.
Part 4 — Optimize per engine, because they are not the same
"AI visibility" is four different games with overlapping rules. Treat them separately.
Google AI Overviews and AI Mode · Googlebot
- Candidates come from ranking. Pages cited in an AI Overview are overwhelmingly pages that already rank in the top results for the query or a closely related one. Classic SEO is the prerequisite.
- Frequency varies enormously by query type. In the same Seer Interactive study, AI Overviews appeared on 36% of informational searches but only 5% of transactional ones, and on 95.4% of head-to-head "X vs Y" comparison searches. Decide where GEO effort pays off before spending it everywhere.
- Query fan-out. AI Mode and increasingly AI Overviews break a question into several sub-queries and pull sources for each. A page that answers a specific sub-question thoroughly can be cited even if it does not rank for the broad head term. Write for the sub-questions.
- Format matters. Lists and short definitional paragraphs are the most common quoted structures. Comparison tables for "vs" queries.
- Freshness. For anything time-sensitive, a visible and accurate dateModified and genuinely updated content is regularly rewarded.
- Measure with Search Console, cautiously. As of this writing Search Console does not separate AI Overview impressions from classic ones; a drop in clicks with stable impressions is your main signal that an AI Overview is absorbing the query. Direct measurement means running the queries and looking.
ChatGPT (with search) · OAI-SearchBot
- Bing is upstream. Make sure you are indexed in Bing Webmaster Tools, and use IndexNow to push updates. A site invisible in Bing is invisible to ChatGPT search.
- Entity and brand descriptions carry a lot of weight. ChatGPT often answers "best X" questions from a blend of its training knowledge and a handful of listicles and directories. Being present, consistently described, in those third-party pages moves the needle more than on-page tweaks.
- Answers vary run to run. Same prompt, different day, different sources. Any measurement needs multiple runs per prompt before you conclude anything.
Perplexity · PerplexityBot
- Most citation-transparent of the four. Every answer shows numbered sources, so you can see exactly which pages win and reverse-engineer what they have in common.
- Rewards recency and specificity. Perplexity leans on fresh, factual, well-structured pages, and is more willing than Google to cite a smaller site if it has the precise answer.
- Direct answers and clear headings matter even more here because Perplexity's extraction is aggressive: it quotes short passages.
Gemini · grounded on Search
- Grounded on Google Search, so the AI Overview advice applies. Also inherits Google's knowledge graph, which makes entity consistency (Organization schema, Google Business Profile, consistent descriptions) particularly important.
- The API and the consumer app can disagree. If you test via the Gemini API and see nothing, do not assume the same for the app a user opens; grounding behavior differs. Test the surface your audience actually uses.
Part 5 — Measure it, or you are guessing
Most GEO advice stops at "do these things." The difficult part is knowing whether they worked, because AI answers are non-deterministic, personalized, and change week to week.
A workable measurement process:
- Define a prompt set. Twenty to fifty questions that a prospective customer would actually ask an AI about your category: informational ("how to choose a…"), comparative ("X vs Y"), and commercial ("best … in Milan"). Include the questions where competitors currently win.
- Run every prompt on every engine, repeatedly. One run tells you nothing. Several runs per prompt, spread over days, give you a share of voice: the percentage of answers where the brand is mentioned or cited.
- Separate mentions from citations. Track both. Citations are the number to grow.
- Record who is cited. The domains that keep appearing as sources for your prompts are your real competition, and often your best outreach targets: a directory or comparison page you can get listed on.
- Score the pages you control on citability, before and after changes, with a fixed set of checks (answer-first opening, heading structure, definitions, author and date, schema, crawler access). A deterministic score is the only way to see whether a rewrite actually improved the page and not just your opinion of it.
- Re-run monthly, at minimum. Engines change their behavior faster than search ever did.
Doing this by hand for one brand is a few hours a month. Doing it for fifteen clients on four engines is a job, and it is the reason tools exist.
How Bi-Rank fits in
Bi-Rank is built for exactly step 5 and steps 1–4 above, for agencies that manage several clients.
- GEO Score (0–100). A deterministic, repeatable score per page and per site, based on a fixed set of extractability, authority and technical checks. It tells you where a page sits against a citability threshold and which issues would move it, so you can prioritize and show a client a before/after that is not a matter of opinion.
- Multi-engine tracking. You define the prompt set; Bi-Rank runs it on ChatGPT, Gemini, Perplexity and Google AI Overviews, repeatedly, and reports share of voice, mentions versus citations, sentiment, and which competitor domains are being cited instead of you.
- Technical foundations check. Whether AI crawlers can read the domain, whether llms.txt and schema.org are in place, re-checked weekly.
- White-label. The reports carry your agency's brand and domain, with read-only client access.
It does not write your survey or earn your mentions. It tells you, with numbers, whether what you did worked.
Start a free GEO scanConclusion
Generative engines have not replaced search; they have added a second filter on top of it. Ranking still gets you into the room. Being clear, specific, verifiable and consistently described is what gets you quoted.
Start with the pages that already rank for questions in your category, rewrite their openings so they answer the question outright, add an author and a date, publish one piece of data that is genuinely yours, and check that the AI crawlers are not blocked. Then measure: same prompts, every engine, repeated, tracked over time.
If you want to see where a site stands today, run it through Bi-Rank and look at the score and the prompts before changing anything. The baseline is the part everyone skips.