Field notes, with the instrument attached

What we have actually tested about crawler access, extractability, schema and citation behaviour, with the instrument and the date attached to every number.

Guide9 min read

Methodology Sections AI Can Quote

If you want AI assistants to quote your methodology accurately, make the method extractable, bounded, and separate from interpretation. Most errors come from mixed claims, hidden qualifiers, and scattered definitions.

Read
Guide8 min read

Same Answer in Schema and Copy?

If a fact matters, publish it in visible copy first and mirror it in schema only when the schema adds structure. Do not let schema carry claims your page body does not state plainly.

Read
Guide8 min read

Do sidebars help AI extraction?

Sidebars and callouts can improve AI extractability when they restate a single fact cleanly. They hurt when they introduce competing claims, vague summaries, or repetitive noise around the answer.

Read
Guide9 min read

Publish source excerpts for AI citations?

Yes, usually. Pairing short source excerpts with your interpretation makes factual claims easier to verify, quote, and cite, as long as you separate what the source says from what you think it means.

Read
Guide8 min read

Format Abbreviations for AI Retrieval

If you want AI assistants to retrieve the right concept, put the expanded term and abbreviation together in visible copy, then repeat the preferred form consistently. The goal is not style purity, it is reducing ambiguity at extraction time.

Read
Guide8 min read

AI and archived policy pages

Sometimes, but not reliably enough to trust by default. If your current and archived policy pages look too similar, assistants can quote the wrong version or blend both.

Read
Guide9 min read

Repeat facts in intro and detail?

If a fact matters for AI retrieval and human comprehension, state it once high on the page and again where you explain it. The win is clarity and qualifier control, not keyword density.

Read
Guide9 min read

Do definition lists help AI retrieve facts?

Definition lists can help AI assistants extract paired facts, but only when the HTML is clean and the surrounding page removes ambiguity. They are a formatting aid, not a reliability guarantee.

Read
Guide9 min read

Keep qualifiers next to claims

If the condition, exception, or scope note sits far from the headline claim, AI assistants often lift the simple version and drop the limit. Put the qualifier beside the claim if you want cleaner summaries and fewer misquotes.

Read
Guide8 min read

Comparison Verdicts AI Can Extract

Your verdict should be explicit, segmented by use case, and easy to quote in one pass. The goal is not clever copy, it is a clean answer an assistant can lift without mislabeling the winner.

Read
Guide9 min read

Infinite Scroll and AI Citation Accuracy

Usually not reliably. Infinite scroll often hides source text behind client side loading, weak URLs, and unstable page boundaries, which makes extraction and citation less precise.

Read
Guide8 min read

Put the answer first for AI citation?

If you want AI assistants to quote your page cleanly, lead with the answer in plain language. Then support it immediately with scope, evidence, and caveats so the summary does not drift.

Read
Guide9 min read

Do Templates Dilute AI Extraction?

Repeated template copy does not automatically kill AI visibility, but it can make page specific facts harder to isolate. On large sites, the real risk is weak information hierarchy, not reuse by itself.

Read
Guide8 min read

Use HTML Lists for AI Retrieval?

HTML lists help AI systems pull steps, criteria, and grouped facts with less ambiguity than dense paragraphs. They are not a universal upgrade, and overusing them can strip context the model needs to answer well.

Read
Guide8 min read

Should You Add Summary Boxes for AI?

Summary boxes can help AI assistants lift the right facts faster, but only when the page beneath them is consistent, crawlable, and worth citing on its own.

Read
Guide8 min read

Handling disputed facts for AI citations

If a fact is contested, do not hide the dispute. Label the claim, name the source, state the counterclaim, and separate settled facts from interpretation so AI assistants have something clean to quote.

Read
Guide8 min read

Why AI assistants pick one source page

AI assistants usually prefer the page that is easiest to fetch, parse, reconcile, and quote without adding interpretation risk. That is rarely the prettiest page, and often the most explicit one.

Read
Guide9 min read

Can AI Extract Facts From Comparison Tables?

AI assistants can often pull facts from comparison tables, but reliability drops when headers are vague, cells cram multiple claims, or key caveats live outside the table. The best tables are boring, explicit, and backed by visible prose nearby.

Read
Guide9 min read

Hidden text and AI assistants

AI assistants can sometimes retrieve hidden text, but they do it less reliably than plain visible body copy. If a fact matters for citation, put it in the HTML of the page and make it easy to quote without interaction.

Read
Guide9 min read

Put dates near claims you want cited?

Placing a publish or updated date near a claim can help an AI assistant keep time-sensitive facts attached to the statement. It does not rescue weak pages, unclear sourcing, or client-side rendering.

Read
Guide9 min read

Format Definitions for AI Retrieval

If you want service pages cited by AI assistants, definitions need to be explicit, local to the claim, and written in plain language. The win is cleaner extraction, not prettier copy.

Read
Guide9 min read

Image Text or Body Copy for Key Facts?

If a fact matters for AI retrieval, search visibility, or clean citation, it should live in plain body copy. Image text can reinforce a point, but it should not be the only place the fact exists.

Read
Guide9 min read

Same Fact on Multiple Domains

When the same claim appears on your site, partner sites, directories, and copied pages, AI assistants do not reliably reward the original source. They usually pick the clearest, most retrievable, and most corroborated version.

Read
Guide8 min read

Separate Opinions From Facts for AI Citation?

If you want AI assistants to quote you accurately, separate stable facts from interpretation. The goal is cleaner extraction, not sterile writing.

Read
Guide8 min read

When Footnotes Help AI Extraction

Footnotes help when they verify a nearby claim without breaking the sentence that carries the fact. They distract when they become a second layer of copy, push key facts away, or force crawlers and models to reconstruct meaning across fragments.

Read
Guide8 min read

AI fact retrieval on translated pages

Yes, sometimes, but translated pages often break fact retrieval when meaning drifts, canonicals point wrong, or key facts exist in only one language. If you want AI assistants to quote the right version, translation workflow matters as much as SEO.

Read
Guide9 min read

Do canonicals affect AI citations?

Canonicals can help AI assistants converge on one source, but they do not rescue weak page structure, conflicting facts, or hidden content. Treat them as a cleanup signal, not a citation strategy by themselves.

Read
Guide9 min read

Quote-ready blocks for AI citations?

Quote ready blocks can reduce misquotes and make facts easier for AI assistants to lift cleanly. They do not create authority on their own, and they can make pages worse if overused.

Read
Guide9 min read

Tabs and accordions for AI facts?

AI assistants can often extract facts from tabs and accordions if the content is present in the initial HTML. Reliability drops fast when key copy is injected with JavaScript, gated behind events, or fragmented across UI states.

Read
Guide9 min read

What to Put in Page Intros for AI Answers

Page intros help AI answers when they state the core fact fast, define the subject clearly, and remove throat clearing. The intro is not magic, but it often decides whether your page gets extracted cleanly or skipped.

Read
Guide8 min read

Category pages for AI citations

Category pages can earn AI citations, but only when they do more than list links. The page needs a clear claim, stable facts, and visible structure that a crawler can extract without guessing.

Read
Guide8 min read

Design Author Pages for AI Trust

Author pages help AI systems and human reviewers connect claims to a real operator. The win is not decoration, it is making expertise, accountability, and evidence easier to extract.

Read
Guide9 min read

Pagination and facets vs AI crawlers

Pagination and faceted navigation often create too many near-duplicate URLs, weak canonical signals, and fragmented facts. AI crawlers then fetch the wrong pages, miss the core page, or quote filtered variants that were never meant to stand alone.

Read
Guide9 min read

Consolidate Pages for Cleaner AI Citations

Consolidate overlapping pages when multiple URLs state the same core facts with small wording differences. AI systems handle that sprawl badly, especially when they need one stable page to quote or summarize.

Read
Guide8 min read

Publish Research if AI Strips Context?

Yes, but publish it in a format that survives extraction. If your research only works when read end to end, AI summaries will flatten it, and your citations will be weak or misleading.

Read
Guide9 min read

Can comparison pages win AI citations?

Yes, if the page behaves like a decision document instead of a sales page. The trick is to make trade offs, scope, and evidence extractable in plain language.

Read
Guide9 min read

AI-Retrievable Definition Pages

A definition page gets retrieved when the answer is obvious in visible HTML, tightly scoped, and supported by surrounding context. Most pages fail because they mix glossary intent with sales copy, hide facts in tabs, or rely on JavaScript rendering.

Read
Guide9 min read

Above the Fold for AI Extraction?

Sometimes, but not for the reason most teams think. Put core facts early because it improves extraction reliability and editorial clarity, not because AI sees a screen like a human.

Read
Guide8 min read

PDFs vs HTML for AI Citation

PDFs can be cited by AI systems, but equivalent HTML usually gives you better control, cleaner extraction, and fewer formatting failures. PDFs work best as supporting assets, not as the only canonical source for core claims.

Read
Guide9 min read

AI site architecture for canonical facts

AI assistants find canonical facts faster when your site gives each important claim one clear home, repeats it consistently, and removes routing and rendering ambiguity.

Read
Guide8 min read

How to cite sources so AI trusts your page

AI systems trust pages that show where facts came from, separate sourced claims from opinion, and keep citations close to the statement being made. The goal is not academic style. It is fast verification.

Read
Guide8 min read

When Schema Helps, When Copy Is Enough

Schema helps when it removes ambiguity from facts machines struggle to connect reliably. If the page already states a simple claim clearly in visible copy, schema usually supports it rather than rescuing it.

Read
Guide8 min read

Can pSEO Pages Earn AI Citations?

Programmatic SEO pages can earn AI citations, but only when each page carries extractable, page-specific facts. Thin template fleets usually get summarized away or skipped.

Read
Guide8 min read

Glossary Pages for AI Retrieval

Design glossary pages as extractable fact pages, not thin SEO stubs. Clear definitions, stable wording, explicit context, and server rendered content matter more than clever templates.

Read
Guide9 min read

What to Remove for Better AI Extraction

If AI assistants keep missing or mangling your facts, the fix is often removal, not addition. Strip out layout patterns, copy habits, and rendering choices that hide extractable claims.

Read
Guide8 min read

Do updates change AI citations?

Updating a source page can improve how AI assistants cite it, but not on your schedule. The real issue is whether crawlers can access the page, extract the changed fact cleanly, and trust it over competing sources.

Read
Guide8 min read

Conflicting facts and AI answers

AI assistants usually do not reconcile contradictions the way an editor would. They extract the clearest, most repeated, most accessible claim, then answer with that version.

Read
Guide9 min read

When Tables Help AI Extractability

Use tables when the reader, and the model, need stable comparisons, specs, eligibility rules, or timelines at a glance. Do not force narrative into tables, because rigid structure can hide context and create bad summaries.

Read
Guide9 min read

What makes pages easy to cite

AI assistants cite pages that expose clear facts, stable wording, and obvious attribution in HTML. If a model has to infer, chase context, or wait for JavaScript, your page becomes harder to quote.

Read
Guide8 min read

Split facts and narrative onto separate pages?

Most teams do not need separate fact pages and story pages. They need one page that makes extractable facts obvious, while keeping narrative close enough to preserve meaning and trust.

Read
Guide9 min read

When AI Search Fails Low Authority Sites

AI search optimization can help low authority sites get quoted, but it fails when the page is hard to extract, easy to doubt, or unsupported by outside references. Clean structure matters, yet citation eligibility is not the same as citation likelihood.

Read
Guide8 min read

Write Copy That Survives AI Summaries

Write for extraction, not just persuasion. The copy that survives AI summarization puts claims in plain language, keeps context close to facts, and removes ambiguity that models compress badly.

Read
Guide8 min read

What Breaks AI Citation on JS-Heavy Sites

AI citation often fails on modern marketing sites because the page looks complete in a browser but incomplete to crawlers. If key facts arrive late, hide in tabs, or depend on hydration, assistants have less to quote and less confidence to cite.

Read
Guide9 min read

Block AI Crawlers on Staging?

Yes, you should block AI crawlers on staging and test environments by default. Use auth, noindex, and per bot robots rules so unfinished pages do not get fetched, quoted, or confused with production.

Read
Guide8 min read

Schema that helps after FAQ rich results

FAQ rich results are gone, but that does not make schema irrelevant. The useful types now are the ones that reduce ambiguity for crawlers and answer systems extracting facts from your pages.

Read
Guide9 min read

Measure AI Visibility Without Screenshots

You can measure AI visibility without trusting vendor dashboards by combining prompt tracking, server logs, rendered-page checks, and citation capture. The point is not a single score, it is a repeatable evidence trail that shows what models can access, extract, and cite.

Read
Guide8 min read

Structure Pages for AI Fact Extraction

If you want AI assistants to quote your site, structure pages for extractability first. Clear assertions, stable labels, server rendered content, and tight evidence formatting beat decorative UX.

Read
Guide9 min read

Why AI cites aggregators first

AI assistants often cite aggregators because those pages are easier to crawl, compare, and quote than vendor pages. The fix is usually extractability, corroboration, and page design, not louder publishing.

Read
Guide8 min read

What llms.txt Still Does

llms.txt is not a ranking or citation lever for Google Search. It can still help with documentation hygiene, internal alignment, and giving AI vendors a clean map of what you want quoted.

Read
Guide8 min read

Do AI Crawlers Run JavaScript?

Short answer, treat AI crawlers as fetchers, not browsers. If critical copy, entities, and citations depend on client-side rendering, many AI systems will miss them.

Read
Guide8 min read

Should You Publish llms.txt?

Publish llms.txt only as a low effort housekeeping file, not as an AI visibility strategy. It does not improve Google Search performance, and current evidence does not show citation lift.

Read
Comparison10 min read

AI visibility tools compared honestly

AI visibility tools are useful for spotting trends, but none can directly measure how often your brand is truly seen, trusted, and cited by AI assistants. The right choice depends on whether you need prompt tracking, source analysis, or technical debugging.

Read
Guide8 min read

IndexNow for AI freshness

IndexNow can speed up URL discovery inside Microsoft's ecosystem, but it is not a direct citation switch for AI assistants. Use it to reduce lag after meaningful page changes, not as a substitute for extractable pages and clear entities.

Read
Guide9 min read

Why vendor sites lose AI citations

Vendor sites often lose broad commercial AI citations because assistants prefer pages that compare, define, and quote cleanly without obvious sales intent. If you want to win more of those mentions, you need extractable evidence pages, not just product pages.

Read
Guide9 min read

AI answer schema, signal or folklore?

Most schema advice for AI answers mixes solid extraction help with wishful thinking. Here is what actually improves machine readability, what does not, and where the advice breaks.

Read
Guide8 min read

Why chatbots misdescribe your brand

Chatbots usually misdescribe brands when your site, profiles, and third party mentions disagree on basics like category, audience, and product scope. Entity consistency reduces that drift by making the same facts easy to extract, repeat, and cite.

Read
Guide8 min read

Bing AI grounding queries guide

The Bing AI performance report is useful because grounding queries show how Microsoft systems connect prompts to pages. Treat it as directional editorial input, not a clean attribution model.

Read
Field note8 min read

Do AI Crawlers Respect robots.txt?

Some AI crawlers clearly check robots.txt and change behavior when blocked. Others are noisier, less transparent, or hard to attribute cleanly from logs alone, so the practical answer is bot specific, not universal.

Read
Guide9 min read

SSR vs CSR for AI citations

If AI crawlers cannot extract the page without running JavaScript, your citation odds drop fast. Server rendered HTML gives assistants usable text at fetch time, which is often the deciding difference.

Read
Guide8 min read

Test What AI Can Quote From Your Page

If an AI assistant cannot lift a clean fact, definition, step, or comparison from your page, it is unlikely to cite you reliably. This guide shows how to test quote readiness with a repeatable operator workflow.

Read
Field note9 min read

What GPTBot Fetches, From Server Logs

If you want AI visibility work to hold up, stop guessing from page source and check server logs. GPTBot mostly tells you whether core HTML, assets, and crawl access are usable, not whether your JavaScript app magically rendered for it.

Read
Rendering11 min read

Can AI Crawlers Read Your Site? The Rendering Problem

No, the major AI crawlers do not execute JavaScript. They read whatever HTML arrives in the first response and nothing else, which means a client rendered site hands them an empty container. We fetched our own routes with the documented GPTBot string and published the byte counts, including the ones that embarrassed us.

Read
Crawler access9 min read

AI Crawler Access: robots.txt for GPTBot and Friends

Training crawlers and search crawlers are different machines with different consequences, and one robots.txt line written by someone who did not know that can remove you from AI answers entirely. Every documented agent string, what it is for, and the specific cost of denying each one.

Read
Citation mechanics12 min read

AI Cites Passages, Not Pages: What Gets You Quoted

Retrieval scores passages rather than pages, which is why a narrow self contained section can enter an answer the whole page would never have ranked for. The traits that separate a quoted passage from an ignored one, the five tests we run before publishing, and the widely repeated citation statistics we refuse to print.

Read
Method11 min read

GEO vs SEO: Mostly the Same, Two Things Genuinely New

Mostly the same discipline, and Google says so in its own documentation. Two differences are genuinely new: the AI crawlers do not run JavaScript while Googlebot does, so a site can pass classic SEO and fail retrieval outright, and retrieval scores passages instead of pages. Everything else being sold as new is not.

Read
Schema10 min read

Schema for AI Answer Engines: What Still Earns Its Keep

Google states plainly that structured data is not required for its generative features. Microsoft has said on the record that schema helps its models understand content. Both are true at once, and the reason to ship it is entity disambiguation rather than a ranking effect. Which types still pay for themselves, and which are dead.

Read
Entities11 min read

AI Describes Your Company Wrong Because Sources Disagree

An assistant is not reading your about page and getting it wrong. It is reconciling sources that disagree with each other, and the loudest disagreement wins. The five step correction loop we run, starting with the part most teams skip: finding out which sources the model is actually reconciling.

Read
Measurement9 min read

Measuring AI Search Visibility: What Is Actually Knowable

Two free first party reports exist and almost nobody has switched them on. For the assistants that report nothing at all, the honest instrument is a fixed query panel run the same way every month. Also why referral traffic undercounts, and why a single composite visibility score cannot move for a reason you can act on.

Read
Benchmark12 min read

AI Crawlers Get an Empty Page: The Extractability Index

A page is invisible to an AI crawler when the first response carries no body copy, which is measurable in bytes rather than arguable. The method, the shell baseline measured on the same domain the same day, and the raw counts for every route type we fetched with the documented GPTBot string.

Read
Methodology9 min read

Results and Methodology: What We Can Actually Prove

Every number this site publishes, what instrument produced it, what it does not cover, and where we have no data at all. The failing rows from our own audit are left in, because a results page with no failures in it is a marketing page wearing a lab coat.

Read
Evidence8 min read

The Evidence Register: Every Claim, Labelled and Dated

The table every other page here defers to. Each load bearing claim appears once with a confidence label, its source, the date it was last verified and the date it stops being trusted. Rows are ordered by recheck date, soonest first, so the claim closest to expiry is the first one you read.

Read
Evidence16 min read

Outbound Pros, Graded Against Our Own Evidence Rules

InboundPros and Outbound Pros are the same group, so recommending the agency without applying our own labels to it would make the labels decorative. Here is the group agency run through the same confidence rules as everyone else, including the claims we will not stand behind and the buyers it fits badly.

Read

Talk through your AI visibility with people who measure it

30 minutes. We will look at what assistants can actually retrieve from your site and tell you plainly what is worth fixing first.

Book a strategy call

30 minutes, no obligation. The calendar shows real availability.

Or start with the free GTM audit from Outbound Pros