// guide
Does GEO actually work, or is it hype?

Written by: Dmitry Filippov, Founder, GET-GEO.AI
Published: 2026-09-09 · Updated: 2026-09-17
// short answer
GEO works at the level it can be measured. A page written for a specific buyer prompt, on a site AI crawlers can read, gets cited on that prompt within weeks to a quarter, depending on the access work needed. Brand-level lift on generic prompts in crowded categories is not proven, and anyone claiming it skipped the measurement.
- GEO works where it can be measured: a page for a specific buyer prompt on a crawlable site gets cited within weeks to a quarter; generic-prompt brand lift is unproven.
- Seven pieces of evidence, one of them ours, support GEO for narrow prompts, from the 2023 GEO paper to OpenAI’s own description of ChatGPT search; none proves “best provider” recommendations in crowded categories.
- In our September 2026 study of 569 links cited by four assistants, 48% pointed to vendors’ own sites; a vendor’s page appeared in 96% of Perplexity answers, 67% of ChatGPT’s.
- On our own site, ChatGPT named get-geo.ai in five of five clean sessions on August 13, 2026, roughly two months after launch, on a narrow multilingual prompt.
- The $0 test: ten buyer prompts run three times each, one access fix, one new page, then the same ten prompts again three weeks later.
Does GEO actually work?
The evidence supports GEO for narrow questions; results for generic questions remain unproven. Three kinds of evidence are available: a research paper using a synthetic benchmark, large studies of what assistants cite published by SEO tool vendors alongside the platforms’ own documentation, and small experiments on sites such as ours. All three support work on pages that answer specific prompts. None establishes what buyers often want most: a recommendation for “the best” provider in a category full of established names.
Two terms before the evidence. Citation rate is the share of a fixed set of buyer prompts where an assistant links to one of your pages. Share of voice is your share of brand mentions on that same set, against the competitors named beside you. Generative engine optimization (GEO) is the work of moving those two numbers.
What does the evidence say?
The evidence supports changes to content and crawler access for specific prompts. It does not yet establish broader gains in brand visibility. The table assesses how useful each source is when deciding whether to invest.
| Claim | Evidence | Source and sample | How much weight it carries |
|---|---|---|---|
| Content changes move visibility inside generated answers | Adding statistics, quotations and citations of sources raised visibility on the paper’s benchmark; keyword stuffing did not | Aggarwal et al., arXiv, November 2023; synthetic benchmark (GEO-bench) | Directionally strong, numerically weak: a research setup, not live assistants; the percentages should not be quoted as expectations |
| Being cited does not require ranking in Google | 28.3% of ChatGPT’s 1,000 most-cited pages rank for zero organic keywords | Ahrefs, October 2025; 1,000 pages | Shows that citation and Google rankings differ, but does not explain how those pages earned citations |
| Assistants search and re-search before citing | ChatGPT typically rewrites the question into one or more targeted searches and can send further queries after the first results; placement is not guaranteed | OpenAI Help Center, ChatGPT search page; checked September 2026 | Documentation rather than a study: it establishes that the access layer is checked at answer time, so an unreadable page fails before content is judged; it says nothing about how often |
| Comparison lists get cited on “best X” prompts | “Best X” blog lists were 43.8% of all page types among the source links ChatGPT cited on “best X” prompts | Ahrefs, December 2025; 750 prompts | Useful for choosing where to seek independent coverage; also helps explain why broader brand visibility takes time |
| Fresh pages are cited more | Of 564 most-cited pages with a detectable date, 76.4% updated within the previous month | Ahrefs, October 2025; 564 of 1,000 pages | Correlation only; more than half the subset is Wikipedia and the timestamps are unreliable |
| A vendor’s own page is in most answers, alongside pages it does not control | Of 569 links cited across 120 answers, 48% pointed to vendors’ and agencies’ own sites, 16% to directories and comparison sites, 10% to forums, 4% to media; a vendor’s own page appeared in 96% of Perplexity answers with sources, 67% of ChatGPT’s and 62% of AI Mode’s; ChatGPT cited no forums and 25% platform documentation, AI Mode 22% forums | GET-GEO.AI, September 2026; 569 links from 120 answers, four assistants (ChatGPT, Perplexity, Google AI Mode, Gemini app), 14–15 September 2026 | Own research, not independent: two days, 23 prompts in our own categories, one classifier; shows which page types are in play per assistant, not what caused any citation |
| No special GEO work is needed for Google’s AI features | Google: no additional requirements or optimizations for AI Overviews and AI Mode; no AI files or schema needed | Google Search Central, updated December 2025 | Strong, and it is the skeptic’s best exhibit: for Google’s surfaces, GEO is largely SEO fundamentals done properly |
What does the skeptic’s case get right?
Three things, and a buyer should hold on to all of them. First, Google’s position is documented: its AI Overviews and AI Mode draw on the same index as classic search, require nothing extra, and no AI-specific file or schema changes eligibility. If a vendor’s GEO package for Google surfaces is llms.txt and “AI schema”, the skeptic is right and the package is housekeeping. Our own guide on llms.txt says the same.
Second, most large-sample studies are published by companies that sell SEO or GEO services, ours included. That does not make them wrong, and Ahrefs in particular publishes its caveats, but it means there is no independent index of GEO results yet. Third, generated answers vary between runs. In our August 2026 runs of one prompt family in ChatGPT, twelve different competitors appeared beside us across five clean sessions. A single screenshot proves nothing, in either direction.
Where the skeptic overreaches is in concluding that nothing can be measured. A fixed prompt set, run several times in clean sessions every few weeks, produces citation rate and share of voice with enough repetition to see movement. The variance is a reason to sample more, and the sampling is what a measurement protocol is for.
What happened when we applied GEO to our own site?
get-geo.ai went live in June 2026 with the technical layer mostly in place: server-rendered text, language tags (hreflang) across eight languages at launch, structured data, llms.txt. On August 4 an audit of our own site found a redirect loop between the bare domain and www, the kind of thing Google’s crawler tolerates and AI fetchers, which often give up after one or two requests, frequently do not. We fixed it that day.
On August 13 we asked ChatGPT in five clean sessions to recommend a GEO agency for a company working in several languages, varying the phrasing. It named get-geo.ai in five of five. On August 17 a sixth run, in English and Spanish, did the same. From launch to the first citations we could reproduce on those prompts: roughly two months.
The limits, in our own words. Six runs across two days is evidence of a first result, and the prompts fit our positioning unusually well, so this is the narrow-prompt case working and nothing more; it says nothing about the generic “best GEO agency” prompt, where established names still dominate. The rest of the limitations are in the case study and in our guide on how long GEO takes.
When does GEO not work?
These four situations account for most of the cases where prospective clients tell us GEO produced no results.
- The site cannot be read. A single-page application that renders text in the browser, a cookie wall in front of the content, a redirect loop, or robots.txt blocking OAI-SearchBot or PerplexityBot, the crawlers ChatGPT and Perplexity send to fetch pages: the verification step in the table above fails before content quality is ever considered. Fix this first or nothing else counts.
- The category is crowded and the prompt is generic. “Best CRM” or “top law firm in London” is answered from comparison lists and years of independent mentions. A new entrant can move on the narrow prompts (“CRM for dental clinics with two locations”) within a quarter; the generic one moves on the slowest layer, described in our guide on how long GEO takes.
- Nobody else describes you. Assistants weigh corroboration: other sites saying what your site says. In our page-types study, a classification of 569 links cited across 120 answers by ChatGPT, Perplexity, Google AI Mode and the Gemini app on 14–15 September 2026, 48% pointed to vendors’ and agencies’ own sites, 16% to directories and comparison sites, 10% to forums and 4% to media, and the mix differed by assistant: ChatGPT’s 143 links included no forum threads and 25% platform documentation, where Google AI Mode sent 22% of its links to forums. A company invisible outside its own domain gets cited on its own facts (prices, features, hours) and not much else until the independent mentions exist.
- The goal was traffic, not answers. Pew Research found that Google users click through less often when an AI summary appears, so a citation may bring fewer visits than a classic ranking did. GEO works when the goal is being the source of the answer; if the plan needs the old click volume, the arithmetic has to be redone first.
How do you test GEO yourself for $0 in 90 days?
You do not need to buy anything to find out. Pick ten prompts your buyers would actually type, in one language, and run each three times in a clean session (logged out, or a temporary chat with memory off) in one assistant. Record who is named, in what position, and which pages are cited. That is your baseline, and it takes an afternoon. Our guide on how we measure GEO results describes the full version.
Then fix access: fetch three important pages with an AI crawler’s user-agent, or simply view the page source, and make sure the text is there without JavaScript, on one host, with the crawlers allowed. Write one page for the prompt where you are furthest from being cited: the prompt as the headline, a 40–60 word direct answer, then the detail a model would verify. Wait three weeks, run the same ten prompts the same way, and compare.
If the new page is cited on its prompt, GEO works for you at the level the evidence supports, and the question becomes how many prompts are worth a page. That is also the only honest way to answer “is GEO worth it”: GEO ROI is the referral traffic and share of voice on prompts you chose, against the cost of the pages, rather than a percentage from a vendor deck. If it is not, the first suspect is still access, and the second is that the page answers a question nobody asked. Either way you now hold a baseline, which is more than most proposals will offer you. Our free AI visibility audit is this test, done for you, on ten prompts.
Related questions
Is GEO just SEO with a new name?
For Google’s AI features, largely yes: Google documents that no extra work is needed beyond SEO fundamentals. For ChatGPT, Perplexity and Claude, the overlap is real but incomplete: a large share of ChatGPT’s most-cited pages rank for nothing in Google, and assistants verify sources with their own background searches. Same foundations, different objective, different measurement.
Is GEO worth it for a small company?
It is worth it where narrow prompts exist: a clinic, a niche SaaS or a regional service can be cited on “X for Y in Z” questions within a quarter, from one page each. It is not worth paying for brand-level promises in a crowded category. Run the $0 test above before signing anything.
Why do some GEO studies show huge gains and others nothing?
Because they measure different things. The GEO paper measured visibility inside a synthetic benchmark, where content changes moved the numbers a lot. Vendors and agencies measure what live assistants cite across thousands of prompts, which reflects corroboration and freshness more than any single page. A vendor quoting the paper’s percentages as an expectation for your site is mixing the two.
Does GEO bring traffic, or only mentions?
Both, but less traffic per appearance than a classic ranking. Pew found Google users click less when an AI summary is shown, and an assistant answer that names you may satisfy the question without a visit. The value is being the source the buyer hears first; measure referral traffic from assistants in your analytics as the final arbiter.
Related guides
Sources
- 01Aggarwal et al. — GEO: Generative Engine Optimization (arXiv, November 2023)
- 02Ahrefs — ChatGPT’s most cited pages (October 2025)
- 03Ahrefs — What ChatGPT cites on “best X” prompts (December 2025)
- 04OpenAI Help Center — Searching the web with ChatGPT (query rewriting; “placement is not guaranteed”)
- 05Google Search Central — AI features and your website
- 06Pew Research — Google users click less when an AI summary appears (July 2025)
- 07Our research: what kinds of pages AI assistants cite — 569 links, four assistants, September 2026
- 08Our guide: how long does GEO take to show results?
- 09Our guide: what is llms.txt and does your site need one?
- 10Our guide: how we measure GEO results
- 11Our case study: we asked ChatGPT to recommend a GEO agency
// share
How to cite this page
You may quote and reuse this content with attribution and a link to this page.
“Does GEO actually work, or is it hype?” — GET-GEO.AI, 2026-09-17. https://get-geo.ai/en/guides/does-geo-work