AEO StrategyLLM Invisibility

Which pages actually get you cited by AI engines?

Greg Rosner

By Greg Rosner

Founder of PitchKitchen · Author of StoryCraft for Disruptors

· 7 min read

Hero image for Which pages actually get you cited by AI engines?

TL;DR

Across five weeks of daily tracking on 50 buyer prompts, nine of the ten pages AI engines retrieved most from pitchkitchen.com were listicles or comparisons, plus the homepage. No long-form essay reached the top ten, out of more than 250 published posts. Retrieval and citation are separate events: our healthtech firms list averaged 2.84 citations per conversation, while an essay pulled into 33 conversations was quoted zero times. The pages that get quoted state who each option fits and who it doesn't. Engines quote the sentence that resolves a choice, which means the bottleneck is your position, not your page format.

Across five weeks of daily tracking, the pages AI engines actually pulled from our site were listicles, comparisons, and the homepage. Nine of our ten most-retrieved pages were one of those three. Not one long-form essay reached the top ten, and we've published more than 250 of them. If you're deciding what to publish next to get recommended inside a chat window, that ratio is the finding, and it cost us four months of writing to learn.

Here's the page-level ledger from our own instrument: what the engines retrieved, what they quoted, what they read and ignored, and the one thing the quoted pages share. We've written the full playbook of how we got recommended elsewhere. This is the autopsy underneath it.

Which of our pages do AI engines actually pull?

We track 50 buyer prompts daily across ChatGPT, Gemini, Perplexity, Claude, and Google AI Overview, measured against a 13-brand competitive field. Between July 13 and August 16, 2026, this is what the engines pulled from pitchkitchen.com most often.

PageTypeConversations that retrieved itCitations per conversation
/best-messaging-positioning-firmsListicle3161.65
/best-b2b-messaging-firms-for-enterprise-software-companiesListicle3161.17
HomepageHomepage1931.92
/blog/best-strategic-messaging-frameworks-for-b2b-saas-companiesListicle1841.54
/pitchkitchen-vs-april-dunfordComparison1211.17
/best-b2b-messaging-firms-for-healthtech-companiesListicle1052.84
/alternatives-to-storybrandComparison1010.93

The healthtech list is the standout. It gets a third of the traffic of our top page and earns 2.84 citations per conversation, the best rate of anything we publish at volume. Narrow beats broad. A page that says which firms fit a digital health company selling to clinical and economic buyers gives an engine something specific to repeat, and specificity is what gets repeated.

Why do some retrieved pages never get quoted?

Retrieval and citation are separate events, and the gap between them is where most content dies. An engine pulls your page into context, reads it, finds nothing worth lifting, and answers from someone else's page. You never appear. Our own tail looks like this:

  • An essay of ours was pulled into 33 conversations and quoted zero times.
  • Our press and podcast page was retrieved 40 times and cited once, a rate of 0.03.
  • Our frameworks index was retrieved 37 times and cited twice.
  • A well-written post on timing a positioning fix was retrieved 34 times and cited 4 times.

Compare those to a plain pricing explainer that only got retrieved 32 times and earned 82 citations, a rate of 2.56. Same site, same authority, same domain. The difference sits entirely in whether the page contains a sentence that answers a question on its own. We've unpacked the mechanics of being used versus being cited, and this is what it looks like at page level on a real site.

Why did our visibility stall while our share of voice climbed?

This is the part that would have made us quit if we'd only watched one number. Across four full tracking weeks, our visibility, meaning the share of AI answers that mention us at all, went from 19.6% to 18.7%. It went down. Anyone reading that column alone would conclude the publishing program failed.

Underneath it, everything else moved the right way. Our share of total mentions in the field climbed from 25.3% to 31.5%. Our average position improved from 2.3 to 1.9, meaning when engines name us we're now typically the first or second name out. Raw mentions went from 616 to 721 in a week. We're named more often, earlier, and in a larger share of what gets said, while the percentage of conversations that mention us at all held steady.

For context on the field: April Dunford's visibility over the same four weeks moved from 21.8% to 17.7%, and her average position from 2.2 to 2.7. That crossover isn't a verdict on her work, which is excellent and is why she's the reference point in this category at all. It's a measurement of publishing cadence against a fixed set of buyer questions. A book earns its authority once. A tracked prompt set asks the question again every morning. If you're reading a visibility dashboard this week, read it the way we've described in how to read AI visibility tools without being fooled, because a single flat number hides both the win and the loss.

The practical rule we now use: visibility tells you whether the category knows you exist, and it moves on a scale of months. Position and share of voice tell you whether this month's work landed, and they move in weeks. Watching visibility alone for the first six weeks of a publishing program is the fastest way to kill a program that's working. We nearly did it in week three.

What do the cited pages have in common?

We went back through every page above a 1.0 citation rate looking for the shared trait. It isn't length, schema markup, or publish date. Every one of them resolves a choice out loud. They say this option fits a company at this size with this problem, and this other option doesn't, and here's the line where the difference sits.

That's why the listicle wins, and the reason has nothing to do with the format. A comparison page can't be written at all until someone has decided who the company is for and who it isn't for. The format forces the decision. An essay lets you defer it forever while sounding thoughtful. When a buyer asks an engine who should rebuild their messaging, the engine is looking for fit statements, and it will quote whoever wrote one.

Which puts the real bottleneck somewhere most AEO advice never goes. If your company hasn't decided who it's for with enough specificity to sit in one row of a table next to a competitor, you can't write the page that gets cited. You'll write around it, publish something graceful, and get retrieved and dropped. That's the same failure that makes AI engines ignore your website wearing a different costume. Your narrative identity is the input. The citation is the output.

There's a test you can run on any draft before publishing it. Find the single sentence a stranger could lift and paste into an answer without needing the paragraph above it or the one below it. If that sentence doesn't exist, the page will get retrieved and dropped, no matter how well it reads. Our zero-citation pages all fail that test. Every page above a 2.0 citation rate passes it in the first screen.

What would we publish next if we started today?

In this order, based on what our own data rewarded:

  1. 1One honest comparison page naming your real alternatives, including the ones you lose to. Put yourself in your true slot, not the top one.
  2. 2One narrow vertical list for the segment where you actually win. Our healthtech page out-earns pages with triple the traffic.
  3. 3A pricing or cost page that states real numbers. Ours quietly outperforms most of the blog on citation rate.
  4. 4A homepage that passes the Cover-the-Logo Test, because it's the third-most-retrieved page on our site and engines read it as your definition of yourself.
  5. 5Then the essays, written for the human who arrives after the engine has already named you.

You can see roughly where you stand before writing anything. Run your homepage through the Brand Signal Score and you'll get the 19-signal read on what an engine can extract from it today. If you'd rather watch this get built live, we run the AI Workforce Clinic weekly, and the first one is free.

One caution on all of the above. Five weeks is five weeks. We're reporting what our instrument recorded on our site in one competitive field, and we'll tell you when a pattern reverses. This is just truth, measured on the only company we can publish the raw numbers for.

Questions People Ask

FAQ

What's the difference between a page being retrieved and being cited?

Retrieval means the engine pulled your page into its working context while answering. Citation means it quoted or linked your page in the visible answer. Our data shows the gap is large and consistent: one of our pages was retrieved in 40 conversations and cited once. Retrieval says the engine found you. Citation says the engine found something on the page worth repeating.

Do listicles really beat blog posts for AI citations?

In our five weeks of page-level data, yes, and it isn't close. Nine of our ten most-retrieved pages were listicles or comparisons. The reason isn't the format itself. List and comparison pages are forced to state who each option fits and who it doesn't, and that fit statement is the quotable unit. An essay that reaches the same conclusion in flowing prose gives the model nothing clean to lift.

Should we stop writing essays then?

No. Essays do work citations don't measure: they're what a founder reads after the engine names you, and they're where your point of view lives. Judge them on whether the right reader replies, not on citation counts. Just don't expect them to carry your AI visibility, and don't let them be the only thing you publish.

How long before new pages show up in AI answers?

Ours moved on a scale of weeks, not days. Over the five weeks we're reporting, our share of voice climbed from 25.3% to 31.5% and our average position improved from 2.3 to 1.9, while raw visibility stayed roughly flat. Watch position and share of voice for early movement. Visibility is the laggard, and reading it alone will tell you nothing is working when something is.

What tools did you use to measure this?

Peec AI, tracking 50 buyer prompts daily across ChatGPT, Gemini, Perplexity, Claude, and Google AI Overview, against a 13-brand competitive field. Every number in this article comes from that instrument between July 13 and August 16, 2026. The first week is partial, since tracking began mid-July.

Want this kind of thinking shipping for you?

If your pages get retrieved and never quoted, you don't have a formatting problem. You have a position that isn't specific enough to summarize next to a competitor's, and no amount of schema markup fixes that. The 90-Day Magnetic Messaging Sprint exists to make that position explicit enough that a model can repeat it without you in the room.

That's the 90-Day Magnetic Messaging Sprint. One quarter, one fixed price: we extract your story, build the Magnetic Messaging Framework and your AI Brand Twin, then ship the website and sales enablement that run on it. $25K–$45K fixed for the quarter, and you own all of it at the end.

About the Author

Greg Rosner

Greg Rosner

Founder, PitchKitchen · Author of StoryCraft for Disruptors · Creator of the Magnetic Messaging Framework™

Greg is a B2B messaging therapist for growth-stage CEOs ($5M-$75M). He helps founders extract the truth they've been hiding from themselves, name the villain in their industry, and build the messaging infrastructure that scales their voice through AI. PitchKitchen has worked with 100+ B2B companies across SaaS, healthtech, fintech, cybersecurity, and AI-driven solutions.