Back to Research
citedby Research · Preliminary Case Study

ThriveStack AI Visibility Case Studies: How 20 Brands Grew Citations and Revenue in 90 Days

An AI visibility case study tracks what happens after a brand fixes technical SEO, prompts and citations across ChatGPT, Perplexity and Google AI Mode.

This preliminary case study follows 20 citedby customers, from Agorapulse to FullStory, through a 90-day pilot. Average AI visibility score rose 26.0%, AI crawl bot traffic rose 55%, and AI-referred revenue rose 30%. The table below shows every customer, the method, and the disclaimer that comes with early data.

ChatGPTChatGPTPerplexityPerplexityGeminiGeminiGoogle AI ModeGoogle AI Mode
+30%
average AI-referred revenue lift across 20 customers
+55%
average AI crawl bot traffic increase in 90 days
+26.0%
average AI visibility score increase
Sep 2026 · 12 min read · Preliminary data · 20 customers, 90-day window

This AI visibility case study is preliminary. It covers one 90-day pilot across 20 citedby customers, over a single quarter. Read it as a first signal that still needs more time to confirm.

20
citedby customers in this preliminary cohort
90
days from audit to remeasurement
+26.0%
average AI visibility score increase
+30%
average AI-referred revenue lift
01 · The results

What happened to AI visibility score, traffic and revenue across 20 brands?

Twenty citedby customers ran a 90-day pilot. Average AI visibility score rose 26.0%, AI crawl bot traffic rose 55%, human traffic rose 6.0%, and AI-referred revenue rose 30%.

This is a preliminary case study. The sample is small, the window is short, and every brand started from a different baseline. Treat the numbers below as an early read that still needs more time to confirm.

Two things skew those averages, and we would rather show them than hide them. ThriveStack's own site ran every fix the day it shipped. Its score moved from 6 to 14, a 133% jump. Take that one account out, and the typical score gain across the rest of the cohort was 20.4%.

Four smaller accounts saw revenue barely move: data-mania.com, oyubotanica.com, ordercup.com and tractionresearchlabs.com. Each was up just 1% in 90 days, even though their visibility and crawl traffic moved like everyone else's. The other sixteen averaged a 38% revenue lift. Visibility can move before revenue catches up, especially for a smaller site in a short window.

The table lists all 20 customers, their industry, and the before and after numbers we tracked. Traffic and revenue are shown as an index, with the pre-audit period set to 100. An index shows the size of the move without publishing each company's exact traffic or revenue figures. The color next to each percentage is a red, amber and green read of that single metric on its own.

DomainCompanyIndustryAI Visibility ScoreAI Crawl Bot Traffic IndexHuman Traffic IndexAI-Referred Revenue Index
Agorapulseagorapulse.comAgorapulseSocial media management software7 → 9 +28.6%100 → 166 +66%100 → 105 +5.2%100 → 140 +40%
Hotel Investor Appshotelinvestorapps.comHotel Investor AppsHospitality investment software12 → 14 +16.7%100 → 152 +52%100 → 106 +5.5%100 → 131 +31%
AskPorteraskporter.comAskPorterAI property operations (PropTech)12 → 14 +16.7%100 → 162 +62%100 → 105 +5.2%100 → 137 +37%
FiveXfivex.comFiveXB2B growth and RevOps software16 → 21 +31.2%100 → 170 +70%100 → 107 +6.6%100 → 132 +32%
Kenken.soKenAI knowledge management8 → 10 +25.0%100 → 147 +47%100 → 105 +5.2%100 → 148 +48%
PeopleKitpeoplekit.comPeopleKitHR and people analytics11 → 13 +18.2%100 → 151 +51%100 → 106 +6.3%100 → 145 +45%
ClickPostclickpost.aiClickPostLogistics and delivery tracking10 → 12 +20.0%100 → 157 +57%100 → 106 +6.5%100 → 141 +41%
Timelines.aitimelines.aiTimelines.aiWhatsApp business communication11 → 13 +18.2%100 → 141 +41%100 → 106 +6.2%100 → 134 +34%
RVsharervshare.comRVshareRV rental marketplace12 → 15 +25.0%100 → 148 +48%100 → 106 +6.3%100 → 134 +34%
Data-Maniadata-mania.comData-ManiaData science education and consulting11 → 13 +18.2%100 → 157 +57%100 → 106 +5.8%100 → 101.0 +1.0%
LeadTaggerleadtagger.comLeadTaggerSales intelligence software11 → 14 +27.3%100 → 166 +66%100 → 106 +6.1%100 → 144 +44%
Traction Research Labstractionresearchlabs.comTraction Research LabsGTM market research3 → 4 +33.3%100 → 155 +55%100 → 106 +6.2%100 → 101.0 +1.0%
Awery Aviation Softwareawery.aeroAwery Aviation SoftwareAviation ERP software6 → 7 +16.7%100 → 169 +69%100 → 106 +6.3%100 → 133 +33%
OrderCupordercup.comOrderCupE-commerce shipping software7 → 8 +14.3%100 → 168 +68%100 → 107 +6.6%100 → 101.0 +1.0%
Oyu Botanicaoyubotanica.comOyu BotanicaBotanical skincare (DTC beauty)8 → 9 +12.5%100 → 148 +48%100 → 106 +5.8%100 → 101.0 +1.0%
BetterPicbetterpic.ioBetterPicAI headshot generation7 → 8 +14.3%100 → 166 +66%100 → 105 +5.4%100 → 136 +36%
AvaSureavasure.comAvaSureVirtual patient monitoring (health tech)10 → 12 +20.0%100 → 151 +51%100 → 106 +6.2%100 → 147 +47%
FullStoryfullstory.comFullStoryDigital experience analytics7 → 8 +14.3%100 → 141 +41%100 → 107 +6.6%100 → 136 +36%
HR Cloudhrcloud.comHR CloudHR information systems12 → 14 +16.7%100 → 146 +46%100 → 106 +6.3%100 → 135 +35%
ThriveStackthrivestack.aiThriveStackRevenue intelligence platform6 → 14 +133.3%100 → 144 +44%100 → 106 +6.5%100 → 131 +31%

Green: strong lift for that metricAmber: moderate liftRed: minimal or not yet moved

Domain, company, industry and the small site icon next to each domain are public information pulled from each company's own website. Score, traffic and revenue columns are preliminary pilot data collected by ThriveStack and have not been independently audited.

AI-referred revenue moved more than the visibility score itselfAverage percent lift by metric across the 20-customer pilot, 90 daysAI visibility score+26.0%AI crawl bot traffic+55%Human traffic+6.0%AI-referred revenue+30%citedby customer pilot cohort, n=20, preliminary, Sep 2026.
Source: citedby customer pilot cohort, n=20, preliminary, Sep 2026.
02 · The mechanism

What AI search visibility factors move a brand in and out of an answer?

Eight variables move an AI answer at once, even when the prompt is identical. Search indexation, prompt quality, run frequency, engine choice and the API versus interface gap are yours to control. Persona, memory and model version are not.

Run the same prompt twice and you rarely get the same answer. In the largest public study of the question, fewer than 1 in 100 repeat runs returned the same set of brands, and fewer than 1 in 1,000 returned them in the same order.

ThriveStack's own research puts it plainly: "Rank is noise. Frequency is signal. Most dashboards sell the noise." Presence held steadier than rank in that same research. Leading headphone brands appeared in 55% to 77% of runs. One cancer care brand appeared in 97% of ChatGPT answers about its category, yet ranked first in only a quarter of them.

This is why citedby's audit starts with the checklist below. The fixes aim at appearing in more runs, since rank position resets on its own anyway.

VariableWhat it changesObserved effectControl
1. Search indexation
(GSC / IndexNow)
Whether content enters the RAG indexUnindexed pages are 100% invisible to live search-augmented AI enginesYes
2. Prompt qualityWhich brands surfaceReal prompts on one intent averaged 0.081 semantic similarityYes
3. Run frequencyBrand set and orderingUnder 1% of repeat runs return the same set of brandsYes
4. Region & localeWhich brands are knownMention rates range roughly 18% to 50% depending on countryPartly
5. AI engineWhich sources get cited11% source overlap between ChatGPT and Perplexity on identical promptsYes
6. API vs. UILength, brands, citationsAround 24% brand overlap between the two surfaces; 406 vs. 743 wordsYes
7. Persona / memoryPersonalization of the answerAccount state, history and session memory shift what is returnedNo
8. Model versionThe entire boardA rollout reshuffles rankings with no changelog you can readNo
Source: ThriveStack, "AI Visibility Tracking: Why Prompts Return Different Answers," https://www.thrivestack.ai/research/ai-visibility-tracking-non-determinism
03 · Tactic 1

What does an AI brand visibility checklist fix on your site?

citedby runs a 73-point audit across four categories: technical infrastructure, content structure, entity signals and earned media. Every fix carries a before and after check, run automatically against the homepage and every key page.

  • Technical infrastructure (16 fixes). Robots.txt rules that allow AI crawlers, an llms.txt file, sitemap coverage and page speed.
  • Content structure (18 fixes). A direct answer in the first 40 to 60 words, question-style headings, FAQ blocks and comparison tables.
  • Entity signals (19 fixes). Organization schema, FAQPage schema, a Wikidata entry and a verified Knowledge Panel.
  • Earned media (20 fixes). Press coverage, guest posts, podcast mentions and third-party reviews.

Most customers in this cohort started with fewer than half the technical infrastructure items in place. An unindexed page cannot be cited, no matter how well it reads, so that category comes first.

CategoryFixesExamples
Technical infrastructure16 fixesRobots.txt rules, an llms.txt file, sitemaps, page speed
Content structure18 fixesA direct answer up top, question headings, FAQ blocks
Entity signals19 fixesOrganization schema, FAQPage schema, a Wikidata entry
Earned media20 fixesPress coverage, guest posts, podcast mentions, reviews
Source: ThriveStack, AI Brand Visibility Checklist.
04 · Tactic 2

What is the F.A.C.T. diagnosis framework for AI visibility?

F.A.C.T. stands for Findable, Agent Accessible, Citable and Trustable. Every customer gets scored on all four before any fix work starts, so the team knows which category to fix first.

Findable maps to indexation: is the page in the sitemap and the llms.txt file, and can a crawler discover it at all. Agent Accessible maps to the technical infrastructure fixes: server-side rendering, clean metadata and SEO-ready markup that an AI agent can parse without running JavaScript. Citable maps to content structure: does the page name the right entities and answer the question in the first few sentences. Trustable maps to earned media: do other sites, reviews and publications independently mention the brand.

A page can read well and still never surface in an AI answer if a crawler cannot reach it or an agent cannot parse it. Fix findable and agent-accessible issues first, every time, then move to citable and trustable fixes. See our AEO Indexing Case Study on how fixing agent crawl errors unlocked massive indexation leaps.
F.A.C.T. categoryWhat it checksFix category
FindableCan AI crawlers and search indexes discover the page at allIndexation, sitemap and llms.txt fixes
Agent AccessibleServer rendered pages with clean metadata, SEO readySSR, metadata and SEO-ready fixes
CitableEntities named and an answer-first structure up topEntity markup and answer-first content fixes
TrustableDo third parties independently mention the brandEarned media and citation fixes
Source: ThriveStack citedby audit framework, 2026.
05 · Tactic 3

How do you create AI visibility prompts that target your market?

Start with buyer-intent prompts, the kind a real prospect would actually type into ChatGPT. citedby builds a starter set of 10 or more prompts across the funnel, then ranks each one by revenue potential.

Comparative prompts such as best [category] for [ICP] and competitive prompts such as [brand] vs [competitor] tend to carry the most buying intent. The panel is not fixed. It is reviewed weekly, and new prompts get added as competitors ship features or a category shifts.

Setup takes 5 to 10 minutes. After that, the weekly loop, adding or editing prompts, scanning the report, and routing gaps to the right team, takes about 30 minutes per brand. For a tactical step-by-step breakdown, explore our guide on How To Increase AI Visibility within Days.

Prompts span the whole buyer funnel from first look to renewalA starter panel of 10 or more prompts, ranked by revenue potentialAwarenessbest [category] for [ICP]Comparison[brand] vs [competitor]Decisionpricing and rolloutRetentionsupport and renewalEach prompt is ranked by revenue potential using the POEM model: paid, owned, earned and media.ThriveStack, How To Increase AI Visibility within Days.
Source: ThriveStack, How To Increase AI Visibility within Days.
06 · Tactic 4

How often should you run AI visibility tracking prompts?

Run every prompt at least seven times a day, per engine. One run carries a standard error of 0.370 on the detection rate, which a University of St. Gallen study called essentially uninformative. The error drops below 0.10 at seven runs and below 0.08 at eight.

Fewer than 1 in 100 repeat runs return the same set of brands. Answers shift with server load, with what the live index returns, and with the time of day. A single reading tells a brand almost nothing about where it actually stands.

This is where usage-based pricing matters. citedby prices by prompt run rather than by seat. A brand that just shipped a fix can run its prompt panel daily to catch the change fast, then drop back to weekly once the picture settles. The frequency moves with the news instead of sitting on a fixed plan.
1 run misses, 7 runs cut the error rate below 0.10Standard error of AI brand detection rate, by runs per prompt per day1 run/day0.3707 runs/day< 0.108 runs/day< 0.08Univ. of St. Gallen, arXiv 2604.07585, Apr 2026.
Source: Univ. of St. Gallen, arXiv 2604.07585, Apr 2026.
07 · Tactic 5

How does an AI referral revenue platform close the loop on AI visibility?

Visibility only matters once it turns into pipeline. citedby ties every AI-referred session to first-touch revenue in the CRM, so a brand can see which engine, which prompt and which page led to a signup. Learn more in our study on Revenue Attribution for AI Search and how teams map multi-touch pipeline.

A full breakdown of how ThriveStack improves AI visibility, engine by engine and fix by fix, is in progress as a separate report. This section will link to it once that piece publishes.

No signup required for the sample dashboard.

Start free →Explore with sample data

08 · What to expect

What AI visibility benchmark should B2B brands expect in 90 days?

Across this cohort, AI visibility score rose 26.0% on average, crawl bot traffic rose 55%, and AI-referred revenue rose 30% in 90 days. Individual results ranged from 1.0% to 48% revenue lift. Compare these metrics with our comprehensive AI Visibility Gap Report 2026 benchmarking 500+ SaaS vendors.

This is an early study. It covers one 90-day window for 20 customers. A secular increase, meaning one that holds up and compounds rather than a single good quarter, needs 6 to 12 more months of tracking before anyone should treat these numbers as a benchmark.

The next update to this report will show whether the gains held, grew or faded once the prompt panel runs for a full year across a larger cohort.

See your own AI visibility score move

citedby audits your site, tracks your prompts, and ties AI referrals to revenue.

Start free →
Frequently asked questions

AI visibility case study: FAQ

What is an AI visibility case study?

It is a before and after look at what happens when a brand fixes technical SEO, entity signals, content structure and prompt coverage, then re-measures how often AI engines mention and cite it. This report covers 20 citedby customers over a 90-day pilot and is preliminary.

How is AI visibility score measured?

citedby measures how often a brand appears across many repeated prompt runs on ChatGPT, Perplexity, Gemini and Google AI Mode, rather than where it ranks in any single answer. Research on non-determinism shows rank changes almost every run, while appearance frequency holds steadier and is the more honest signal.

How many prompt runs do you need before trusting an AI visibility number?

At least seven runs per prompt per engine per day, aggregated over a two to four week rolling window. A University of St. Gallen study found a single run carries a standard error of 0.370, which drops below 0.10 at seven runs.

What is the F.A.C.T. diagnosis?

F.A.C.T. stands for Findable, Agent Accessible, Citable and Trustable. citedby scores every customer on all four before recommending fixes, since a page with strong content that a crawler cannot reach still will not appear in an AI answer.

Why does citedby price by usage instead of by seat?

AI answers are non-deterministic, so tracking needs to flex with events. Usage-based pricing lets a brand run its prompt panel daily right after a fix ships, then scale back to weekly once results settle, instead of paying for a fixed cadence year round.

Is 90 days enough to call this an AI visibility benchmark?

No. This is a preliminary case study across one 90-day window and 20 customers. A secular, durable increase needs 6 to 12 more months of tracking across a larger cohort before it should be treated as a benchmark.

Sources

  1. ThriveStack, "AI Visibility Tracking: Why Prompts Return Different Answers": source of the non-determinism, run-frequency and "what moves a brand" data used in this report.
  2. ThriveStack, "AI Brand Visibility Checklist": source of the 73-point, four-category audit referenced in Tactic 1.
  3. ThriveStack, "How To Increase AI Visibility within Days": source of the prompt-creation guidance referenced in Tactic 3.
  4. citedby customer pilot cohort, n=20, preliminary internal data collected by ThriveStack, September 2026. Not independently audited.