// guide
Do GEO agencies offer a guarantee, and what should one look like?

Written by: Dmitry Filippov, Founder, GET-GEO.AI
Published: 2026-09-22 · Updated: 2026-09-25
// short answer
No GEO agency can guarantee a citation in ChatGPT, Perplexity or Gemini: OpenAI states that placement is not guaranteed, and answers vary between runs. What can be guaranteed is the process. Ours is a formula: 30–50 prompts per language, clean sessions, remeasurement every 2–4 weeks, and at day 90 a review with three options and data that stays yours.
- OpenAI's help page for ChatGPT search, checked 16 September 2026, says results are ranked on multiple factors and placement is not guaranteed. No agency sits above that rule.
- In our ChatGPT run of 14 September 2026, the guarantee prompt returned 11 different agency domains across 3 repeats; only one domain appeared in all three.
- A process guarantee has five checkable parts: 30–50 prompts per language, clean sessions, a day-one baseline, remeasurement every 2–4 weeks and a written day-90 review.
- Our policy in one line: if citation rate and share of voice stay inside the day-one baseline's spread at day 90, we say so first and you choose one of 3 options.
- In our September 2026 check of 60 AI-cited GEO agencies, 8 (13%) promised a result or refund, 19 (32%) said results are not guaranteed and 30 (50%) said nothing.
Why can no agency guarantee a citation?
No agency can guarantee a citation because the answer is generated by a system the agency does not control, and the platform says so itself. OpenAI's help article on searching the web with ChatGPT, checked on 16 September 2026, states that results are ranked on multiple factors and that placement is not guaranteed. It also says ChatGPT can turn a question into targeted search queries. Those search queries may differ from the wording your buyer used; that does not mean the assistant answers a different question.
The second reason is variation between runs. Ask the same question three times in clean sessions (logged out, or in a temporary chat with memory off) and the list of cited domains changes. In our August 2026 runs of one prompt family in ChatGPT, we were named in five of five clean sessions, and twelve different competitors appeared beside us. A single screenshot can neither prove a guarantee was met nor that it was broken.
Different layers also have different timescales. OpenAI's crawler documentation says robots.txt changes are picked up within about 24 hours. This concerns access rules, not a promised crawler visit or citation. Brand mass, meaning how often independent sources name your company, can take quarters to change. Our results timeline guide explains these horizons.
What can be guaranteed instead: the process?
What an agency can guarantee is its own conduct: what it measures, how often, against what, and what it does when the numbers do not move. The client can observe and check these actions; the contract should state the obligations and consequences clearly. Two terms carry the structure. Citation rate is the share of prompts in a fixed set where an assistant cites one of your pages. Share of voice is your share of all brand mentions in the answers to that same set, against the competitor list recorded on day one. Both are counted per assistant and per language; our measurement guide defines them in full. A process guarantee has five parts, each of which can be written into a contract and checked on a calendar.
- A fixed prompt set: 30–50 buyer prompts per language, agreed before work starts, with a copy in your hands. If the list can change mid-engagement, no later chart can be compared with day one.
- A session protocol: repeat every prompt in comparable sessions. OpenAI now offers Personalized and Unpersonalized temporary chats. Choose Unpersonalized to exclude existing memories, custom instructions and plugins; record the setting. Temporary chats do not create or update memories, but the personalized mode can use existing ones.
- A day-one baseline: raw answers and screenshots for the whole set, per assistant, before anything is changed.
- A remeasurement cadence: the same set, the same protocol, every 2–4 weeks during a pilot, logs attached.
- A written day-90 review: what is checked, who reads it first, the named options if the check fails, and who keeps the data.
What does our 90-day policy commit to, in numbers?
Our 90-day policy commits to a measurement, a threshold, a trigger and three outcomes, and it fits in one formula. It is part of the 90-day pilot; the pricing guide lists what the pilot includes and what it costs.
The threshold is the baseline's own spread. The day-one measurement is repeated several times, and those repeats disagree because generated answers vary. If the day-90 rates fall inside that spread, nothing has moved that we can distinguish from noise; if at least one rate, in at least one assistant, moves beyond it, we discuss what the change means. Our guide on what happens if GEO doesn't work after 90 days walks through the day-90 conversation itself.
What the formula leaves out matters as much: no promised position, no promised assistant, no promised date for a first citation, and no refund clause: our policy instead provides the three options above.
- Inputs: the same 30–50 prompts per language as on day one; each run several times in clean sessions; on ChatGPT, Perplexity and Gemini at minimum; remeasured every 2–4 weeks; two metrics per assistant, citation rate and share of voice.
- Condition: at day 90, both rates remain inside the spread the day-one baseline showed between its own repeat runs, in every assistant measured.
- Trigger: we tell you first, in writing, with the day-90 table next to the baseline and our analysis of where progress stalled and why.
- Outcome, your choice of three: a free diagnostic month (baseline re-run, one untested hypothesis, a second measurement, a written recommendation that can be "stop"); a change of scope, rewritten around what the data showed; or a stop.
- Ownership, in every case: the prompt set, the day-one baseline with raw answers and screenshots, every interim measurement, the log extracts, the pages and the change log stay with you.
How do you read a guarantee clause: four questions?
You read a guarantee clause by asking four questions of it, and a clause that cannot answer all four is not a guarantee, whatever the heading says. The questions are the same for a refund clause, a "results in 30 days" line and our own policy. Our guide to questions to ask a GEO agency covers the whole engagement; these four apply to the promise itself.
Measured on what? A guarantee needs a prompt set with a count, a language, a list of assistants and a session protocol. "Mentions in ChatGPT" without those four can be satisfied by one screenshot from one logged-in account.
Compared with what? A guarantee needs a day-one baseline and a rule for the baseline's own variation. A rise from one citation to four is a 300% increase, but the absolute counts alone do not establish a lasting change.
Checked when, and by whom? A guarantee needs a cadence and a named reader. If you can re-run any prompt yourself, logged out, and compare it with the logs, you are not relying on a narrative; our guide on how to check your AI visibility for free is the protocol.
What happens if the check fails, and who keeps the data? A guarantee needs named consequences and a data clause: a refund, a free month, a scope change or a stop, written down, with the prompt set, the baseline and the pages yours whichever applies.
Which guarantee wordings should you accept, negotiate or walk away from?
In our September 2026 check of the 60 GEO agencies that ChatGPT, Perplexity and Google AI Mode cited in our runs of 14–15 September, 8 (13%) promised a result or a refund on their websites, 3 (5%) described what happens if the work does not pay off without promising citations, 19 (32%) stated that results are not guaranteed, and 30 (50%) said nothing about it. Half of this sample did not address the issue publicly, leaving it to be clarified in the proposal; this is not an estimate for the whole market.
Most guarantee wordings fall into seven patterns, and each gets a verdict once you apply the four questions. The verdicts are structural: they depend on what the clause fixes in writing, not on who wrote it. Our red-flags guide covers nine proposal phrases in the same spirit, and our guide on what happens if GEO doesn't work has a companion table with three kinds of GEO promise. This one goes down to the wording, so you can mark up a proposal line by line.
| Promise wording | What you can check on day 90 | Verdict |
|---|---|---|
| "Guaranteed #1 / top-3 mention in ChatGPT" | Nothing, unless the clause fixes a prompt set, a session protocol and a repeat count; a position in one generated answer is not a position | Walk away. The platform itself says placement is not guaranteed |
| "N guaranteed citations within 30 days" | A count and a date without a denominator: on which prompts, out of how many runs, in which assistant | Walk away, unless the prompt set and protocol are attached; then read it as the refund pattern below |
| "Refund if you are not cited on N of M prompts on K platforms by week W" | A testable formula, if you hold the M prompts, sessions are clean, and "cited" means a link to your page | Negotiate. Ask who chose the M prompts, how many repeats count, and what happens to the baseline and pages if the refund is paid |
| "Results in 30 days" | A date, and a metric only if the clause names one; the access layer (robots.txt, redirects, rendering) can be presented as the whole result | Negotiate. Ask which layer, which metric, which prompt set; do not pay quarterly prices for a one-week access fix |
| "Our proprietary visibility score will rise by X%" | A number whose definition is not public and whose raw runs are not shared cannot be checked against anything you hold | Walk away, unless the definition, the prompt set and the raw runs are in the contract |
| "No guarantees; results vary" and nothing else | Honest about the outcome, but nothing to check: no baseline, no cadence, no day-90 procedure | Negotiate. Keep the honesty, add the process: prompt set, baseline, remeasurement, day-90 review, data clause |
| A written process: fixed prompt set, clean-session protocol, day-one baseline, remeasurement cadence, day-90 review with named options, data stays with the client | All of it: you hold the set and the baseline, you can re-run any prompt, and the day-90 outcome is one of the options named in advance | Acceptable. This is the pattern our own policy follows; check that all six parts are in the contract, not the pitch |
What did our September 2026 run show about guarantee promises?
Our September 2026 run showed that the guarantee question is answered from agency websites, that the set of agencies changes from one answer to the next, and that a published policy is what gets cited. On 14 September 2026 we ran twelve prompts in ChatGPT, signed out, web search on, three repeats each: 36 answers. Two prompts were about promises: what happens if GEO does not work after 90 days, and whether GEO agencies offer any guarantee and what one should look like.
Across the six answers to those two prompts, ChatGPT cited 18 different agency domains. On the guarantee prompt alone, three answers produced 11 different domains; one domain appeared in all three repeats. get-geo.ai was cited on both prompts, through our agency page and our measurement guide: the pages that state the 90-day policy. We do not name the other domains and draw no conclusion about them: a domain seen once has been sampled once.
On 15 September 2026 we ran the same prompts twice each in Perplexity (logged in, personalisation off) and Google AI Mode (private window). On the day-90 prompt both cited the what-if guide published the day before, AI Mode in both repeats. Across three assistants, every citation of get-geo.ai came on the two prompts about the 90-day policy and guarantees. That is first reproducible data, not a stable position, as our case study also says. The next remeasurement is scheduled for 26 October 2026.
What should a guarantee never contain?
A guarantee should never contain a promised position, a promised assistant, a promised date for a first citation, or a metric you cannot recompute from data you hold. Each of those turns the clause into something that can only be argued about, and an argument at day 90 is the outcome a guarantee exists to prevent. Three further items look like guarantees and are not.
The first is a paid placement presented as a guaranteed mention. A slot in a list article that assistants already cite can produce a citation, and our guide on whether to buy listicle placements covers when that is worth paying for; it guarantees nothing, because the list can be re-crawled, re-ranked or dropped. The second is a guarantee that covers only the cheapest layer: crawler access can be fixed in days, so a clause that guarantees access and calls it results has guaranteed the part that was never in doubt. The third is a lock-in dressed as a guarantee: a refund or free extension conditional on renewing, or one that keeps the prompt set and the baseline if you leave.
If you are comparing proposals now, the fastest test is the one we recommend for any vendor, including us: ask for ten prompts where you are invisible today, the current answers, and the specific changes that would put you in them. That is our free AI visibility audit, one language, in writing. Email hello@get-geo.ai with your site and the language you sell in, and read our answer against the four questions above.
Related questions
Is a money-back guarantee better than a process guarantee?
Not on its own. A refund returns the fee; ownership of the work and data is a separate contractual question. A useful clause should preserve your prompt set, baseline, interim measurements and pages regardless of the remedy. A refund can sit inside that written process rather than replace it.
Can an agency guarantee a mention by buying a listicle placement?
It can buy the placement; it cannot guarantee the mention. A list article that assistants cite today can be re-crawled, re-ranked or dropped, and answers vary between runs anyway. A placement is a corroboration tactic with its own price and checks; it says nothing about whether your own pages get cited.
What counts as "moved" under your 90-day policy?
Movement is a day-90 citation rate or share of voice, in at least one assistant, outside the spread the day-one baseline showed between its own repeat runs. Inside that spread we cannot separate change from noise, so the policy applies. There is no fixed percentage; the yardstick is your own baseline's variation.
Does the 90-day policy apply to the deep audit or the retainer?
The policy is part of the 90-day pilot, because that is where the baseline, the remeasurement cadence and the day-90 report all exist. The deep audit ends with a baseline and a plan; the retainer runs month to month on monthly measurement. Our pricing guide lists what each tier includes.
Can I test a guarantee before I sign?
Yes. Run the agency's proposed prompt set yourself, logged out, three times per prompt, and record which domains are cited. That gives you a baseline before day one and shows whether the promised metric can be recomputed from data you hold. Our free-check guide is the protocol, and it takes an afternoon.
What if an agency guarantees results in one assistant only?
Ask why that one. ChatGPT, Perplexity and Gemini draw on different source pools, so a promise limited to one assistant may be the one where the agency's usual tactic works. A useful clause measures at least the three, per language, and reports them separately, so you see where movement happened.
Why is the review at day 90 and not at day 30?
Day 90 is the pilot's review date, not a promised citation deadline. Technical changes can be checked early; fresh crawls and independent mentions may take longer. OpenAI's 24-hour guidance concerns robots.txt rules, not page retrieval. We measure every 2–4 weeks and use the day-90 review to decide what to do next.
Related guides
Sources
- 01GET-GEO.AI research: what 60 cited GEO agencies publish — prices, guarantees, method, languages (checked 22 September 2026)
- 02Our guide: what happens if AI visibility hasn't grown after 90 days, with the three-kinds-of-promise table
- 03Our measurement guide: baseline, metrics, monthly deliverables and the 90-day policy
- 04Our guide: red flags in GEO agency proposals
- 05Our guide: what to ask a GEO agency before you hire one
- 06Our pricing guide: tiers, what the 90-day pilot includes, and the policy it carries
- 07Our guide: how long GEO takes, by layer
- 08Our guide: how to check what AI assistants say about your company, for free
- 09Our guide: should you buy listicle placements
- 10Our case study: the clean-session protocol applied to ourselves
- 11OpenAI Help Center — Searching the web with ChatGPT (ranking on multiple factors; placement is not guaranteed)
- 12OpenAI — Overview of OpenAI crawlers (robots.txt changes picked up within about 24 hours)
- 13OpenAI Help Center — temporary chats: Personalized and Unpersonalized settings (checked 25 September 2026)
// share
How to cite this page
You may quote and reuse this content with attribution and a link to this page.
“Do GEO agencies offer a guarantee, and what should one look like?” — GET-GEO.AI, 2026-09-25. https://get-geo.ai/en/guides/geo-guarantee-explained