How the three engines choose sources
All three run a retrieval step against a search index, read the fetched pages, and synthesise an answer with attributions. Google AI Overviews draws on Google's index and appears above organic results for a growing share of queries; in SerionFlow's category, ten of ten money terms tested in September 2026 returned an AI Overview. ChatGPT with search and Perplexity issue their own retrieval queries, often several per prompt, and cite the pages whose passages best answer them. In every case the crawler must fetch the page, the passage must be in the HTML, and the passage must answer the question plainly.
Step one: verify the crawlers can reach you
Check robots.txt for groups blocking OAI-SearchBot, GPTBot, ClaudeBot, PerplexityBot, and Google-Extended, and check your CDN for AI-bot blocking, which Cloudflare can enable by default. Fetch a priority page with each user agent and confirm a 200 with full HTML. SerionFlow's free AI crawler access checker and robots.txt checker do this in a minute. This step alone explains a large share of zero-citation sites.
Step two: put the content in the static HTML
AI crawlers do not execute JavaScript. Fetch your page with JavaScript disabled and count the words; if it is a fraction of what the browser shows, the engines are reading the fraction. When SerionFlow audited itself in September 2026, its pages served 111 to 229 words of static HTML and were cited zero times across six buyer prompts, while competitors with full-HTML comparison pages were cited repeatedly. The fix was to render full page content, headings, tables, and FAQs server-side before writing anything new.
| Step | What to do | Failure it prevents |
|---|---|---|
| 1. Access | Allow AI crawlers in robots.txt and at the CDN; verify with fetches | Engine cannot fetch the page |
| 2. Static HTML | Render full content, headings, tables, FAQs server-side | Engine fetches an empty shell |
| 3. Answer-first pages | One page per question; direct answer in first paragraph; table; FAQ | Engine finds no quotable passage |
| 4. Original data and honesty | Include a fact only you have and a section on limits | Engine prefers a more informative or balanced source |
| 5. Presence on cited sources | Be accurately listed on the third-party pages engines already cite | Engine never encounters your brand in trusted context |
| 6. Measure | Fixed prompt set, monthly re-run, citation share per engine | No way to know what worked |
Step three: write answer-first pages, one per question
Take the buyer questions from your prompt set and from conversational queries in Search Console. For each, publish a page whose H1 is the question, whose first paragraph answers it in forty to sixty words with no preamble, followed by a short TL;DR, sections that each open with a direct statement, a table with real figures, and an FAQ with verbatim questions and short answers. This structure is what all three engines lift passages from. SerionFlow generates drafts in this shape for use-case, how-to, comparison, and listicle intents, but the shape is reproducible by hand.
Step four: add original data and an honest negative
Engines favour sources that add information and that read as balanced. A fact from your own product data or Search Console, dated and explained, gives the page something no competitor page has. A paragraph stating where your product is the wrong choice makes the page read as a source rather than an advertisement, and engines are measurably more willing to cite balanced sources for comparison prompts.
- One unique, dated data point per page.
- One paragraph on limits or when to choose something else.
- Real pricing with an as-of date rather than "contact us".
Step five: be present where the engines already look
Run a citation gap analysis to see which third-party pages are cited for your prompts, then make sure your product is accurately listed on them: comparison sites, community threads where relevant, documentation hubs, and reviewers. Supply facts, not pitches. This is AEO link building; the goal is accurate presence on cited sources rather than link volume.
Step six: measure and iterate
Freeze a prompt set, run it monthly across the three engines, log every cited URL, and track your citation share per engine. Pair it with conversational query impressions in Search Console and AI-related query signals from Bing. Expect a trend over three months. When a prompt stays uncited, compare your page with the cited one on the six steps and fix the gap. Nothing in this method can force a citation; it removes the reasons engines have not to cite you.
Step by step
- 01
Verify crawler access
Check robots.txt and CDN settings for AI-bot blocks and fetch a priority page as each AI user agent, expecting 200 with full HTML.
- 02
Confirm static HTML
Fetch priority pages with JavaScript off and confirm word count, headings, tables, and FAQs are present in the source.
- 03
Publish answer-first pages
For each buyer question, publish a page with the question as H1, a 40–60 word answer first, a TL;DR, a table, and an FAQ.
- 04
Add original data and limits
Include one dated fact only you have and a paragraph on when your product is the wrong choice.
- 05
Earn presence on cited sources
Use a citation gap analysis to find the third-party pages engines cite and get accurately listed on them.
- 06
Measure monthly
Re-run a fixed prompt set across ChatGPT, Perplexity, and AI Overviews, track citation share per engine, and fix uncited pages against the six steps.
Clear answers
Frequently asked questions
Can a small site get cited by ChatGPT or Perplexity?
+
Yes. Engines cite the clearest source for a specific question, and in software categories small vendor comparison and best-of pages are cited routinely. What small sites cannot do is win head terms; what they can do is answer specific questions better than anyone.
Why is my site never cited even though it ranks?
+
Usually because AI crawlers are blocked or the content only exists after JavaScript runs, so engines see an empty page. Fetch your page with JavaScript off and with each AI user agent. If either fails, no content change will help until it is fixed.
How long does it take to get cited?
+
Weeks to months. The page must be published, fetched, and judged the best source on a later answer run. Measure monthly with a fixed prompt set and expect a trend at three months. No method can force a citation.
Does schema markup get pages cited?
+
Not on its own. Accurate schema helps engines understand a page; inaccurate schema hurts trust. Citations follow crawler access, static HTML, and answer-first structure. Add schema after those are in place and keep it matched to visible content.
How does SerionFlow help with getting cited?
+
It generates answer-first pages for buyer questions found in your site, competitor, and Search Console data, serves them as static HTML on your domain via Vercel or Cloudflare routing, submits through Search Console and Bing with IndexNow on Pro and Agency, and monitors sampled AI visibility. It does not promise citations.
Continue exploring
Related SerionFlow resources
More in Guides
- How to test and validate a robots.txt file
- Meta description length and best practices
- Programmatic SEO on Next.js and Vercel
- Programmatic SEO on Shopify: what works and what does not
- Programmatic SEO on Webflow: how to set it up and where it stops
- Programmatic SEO on WordPress: how to set it up and what to watch
- robots.txt for AI crawlers: how to allow or block GPTBot, ClaudeBot, and others
- robots.txt monitoring: how to detect changes and outages
Make the next move obvious
Let SerionFlow turn your market evidence into momentum.
Confirm the market, rank the opportunities that matter, create brand matched pages on your domain, and keep moving with controlled weekly intelligence.