Schema markup for AI answer engines
what still earns its keep
By Jānis Plūme, Founder, Outbound Pros · 10 min read · 2026-08-06
Quick answer
Schema markup is not a ranking input for Google's AI features, and Google states this in its own documentation. Microsoft says the opposite for its own stack, on the record. Both are true, because they are describing different systems. The reason to ship schema in 2026 is entity disambiguation, telling machines unambiguously who you are, who wrote this and what it relates to. Ship Organization, WebSite, Person, BlogPosting with BreadcrumbList, and SoftwareApplication on tools. FAQPage produces no Google rich result since May 2026. The how to type has been dead since 2023. Never fabricate rating markup.
Does schema markup help with AI answers?
It depends entirely on whose AI you mean, and the two largest platforms are on record saying opposite things.
Google, in its guide to optimizing for generative AI features published 15 May 2026: structured data is not required for generative AI search, and there is no special schema.org markup you need to add. That is unambiguous and it comes from the platform itself, not from an inference.
Microsoft, through Fabrice Canel, principal product manager at Bing, speaking at SMX Munich in March 2025: schema markup helps Microsoft's language models understand content, and site owners should push updates through IndexNow. That is the single strongest platform level justification for schema in the AI context available anywhere, and it is worth noticing that it does not come from Google.
Neither statement contradicts the other, because Google is describing an input to its own generative features and Microsoft is describing how its models parse content. Perplexity and Anthropic have published nothing on the subject either way, so anyone telling you Perplexity weights schema is inferring, and should say so.
So what is schema actually for now? Three things, in descending order of value. It disambiguates entities, which is the thing that determines whether an assistant knows your company is your company and not a similarly named one. It remains an eligibility input for whatever rich results still exist in classic search. And it is a machine readable statement of relationships that costs almost nothing to maintain once it is built. What it is not, in 2026, is a lever you pull to increase citations.
Is FAQ schema still worth adding?
Not for a Google rich result, because that feature is fully deprecated. Google restricted FAQ rich results to authoritative government and health sites in August 2023, then removed the feature entirely: the deprecation notice was added to Google's documentation on 8 May 2026, the feature stopped appearing on 7 May 2026, the FAQ search appearance filter and rich result reporting were removed in June 2026, and Search Console API access wound down by August 2026.
Google's documentation notes the markup can stay in place and causes no problem. It simply produces nothing visible.
Here is the distinction that matters and that most coverage of this deprecation missed. The FAQ section as a content structure is strongly supported and worth keeping on every substantial page, because a question followed by a direct answer mirrors the shape of the query a retrieval system is matching against. The FAQPage markup is now a parsing convenience with zero Google upside. Keep the section. Ship the markup if it costs you nothing. Do not let anyone report it as a win, and be suspicious of any 2026 proposal that leads with FAQ schema as a GEO tactic, because it tells you when the author last checked.
Which schema types are worth shipping in 2026?
Six worth shipping, two worth skipping, one worth an honest shrug.
| Type | Ship it | Why |
|---|---|---|
| Organization and WebSite, sitewide | Yes | The entity backbone. Carries sameAs, and parentOrganization or subOrganization where a brand family exists |
| Person, for each author | Yes | Highest value single block if you reuse one identifier everywhere the author appears. Fragmented author identity is a common and invisible failure |
| BlogPosting with BreadcrumbList | Yes | Standard, cheap, still eligible for classic rich results |
| SoftwareApplication, on tool pages | Yes | The best available description of a free calculator or checker |
| FAQPage | Optional | No Google rich result since May 2026. Zero cost, non zero parsing benefit, zero expectation |
| Dataset, on original benchmark pages | Maybe | We ship it on data pages because it is cheap. We have found no evidence either way that any AI surface consumes it, so do not build a case on it |
| The how to type | No | Deprecated in 2023 |
| Review and rating markup | No, unless earned | See the section on what gets you in trouble |
| Anything describing content that is not on the page | No | This is the definition of structured data spam and it carries a manual action |
How do you make a language model resolve your brand correctly?
By saying the same thing about yourself, in the same words, in as many machine readable places as you control, and then getting sources you do not control to agree.
Language models resolve entities by corroboration. They are not looking up a record, they are reconciling many statements. When your LinkedIn page says one company name, your schema says a second spelling, your press coverage says a third, and your product docs use a fourth, you have not given the model a brand, you have given it a disambiguation problem. It will resolve that problem in whichever direction the weight of evidence points, and it may not be your preferred direction.
Three practical rules come out of that.
Pick one display name and one alternate
Ship the canonical name as name and the variant as alternateName in your Organization block. That single field is what tells a machine that two strings are one entity. We run "Outbound Pros" as the display name with "OutboundPros" as the alternate on every property in the group, because both strings are already live across the estate and pretending otherwise would fragment the entity rather than fix it.
Give every author one identifier and reuse it byte for byte
Our founder's Person node uses the same identifier on all five group properties, which is what collapses five author references into one author entity. A mismatched identifier across properties produces five strangers who happen to share a name, which is worse than having no Person markup at all. Worth noting that an identifier is not a link: it is a string that has to match exactly, and pointing it at a page that does not serve real content yet would be worse than leaving the link off, which is why the author name on this page is plain text.
State corporate relationships explicitly
Where a group of properties exists, every child carries parentOrganization pointing at the parent's organization identifier, and the parent carries subOrganization entries for the children. This is ordinary corporate structure expressed in a format a machine can read, and it makes the whole family more resolvable, not less. Hiding the relationship would throw away the main advantage of owning several properties, which is corroboration across sources. We do this across five properties including Outbound Pros, and we publish the pattern because it is genuinely reusable and because a shop that will not show you its own implementation is asking for a lot of trust.
What schema will get you in trouble?
Fabricated review markup, and it is the one schema mistake with legal exposure attached rather than just a ranking cost.
Rating and review blocks describing testimonials that no named customer has approved in writing are a documented Google structured data manual action trigger. That is the cheap version of the consequence. The expensive version is the United States Federal Trade Commission's Rule on the Use of Consumer Reviews and Testimonials, 16 CFR Part 465, published in the Federal Register on 22 August 2024 and in force since 21 October 2024. Penalties are civil and assessed per violation, the rule has not been vacated, and the FTC issued warning letters under it on 22 December 2025, so it is being enforced and not sitting on a shelf. We publish no penalty figure here, because the commonly circulated number is a prior year adjustment and we have not verified the current one. The EU Omnibus Directive requires the same disclosure of connected reviews.
Two sections of that rule govern how anyone should think about testimonial content on their own site. Section 465.5 prohibits a business disseminating testimonials by employees, officers, managers or agents without disclosure when the relationship is not otherwise clear to the audience. Section 465.6 makes it an unfair or deceptive practice to materially misrepresent, expressly or by implication, that a website or entity the business controls provides independent reviews or opinions about a category of businesses including its own.
Read section 465.6 carefully, because the line it draws is narrower and more workable than most people assume. The prohibited act is misrepresenting independence, expressly or by implication. Owning the site is fine. Publishing opinions on it is fine. Recommending your own company on it is fine. Letting a reader believe the site is an independent evaluator when it is not is the violation. Note the phrase "or by implication": silence about ownership on a site that reads like an independent comparison resource is itself the implication.
Which is why our own group ownership appears in the footer of every page here and at the top of the graded write up of the agency, and why no property in this group carries review or rating markup anywhere. Disclosure is not caution in this area. It is the mechanism that converts a possible violation into ordinary first party marketing.
What do we not know about schema and AI?
Two things about schema and AI answers are genuinely unknown: whether any non Google engine weights structured data at all, and whether entity markup changes retrieval or only changes how a correct retrieval gets labelled. Both carry the unknown label in the register of claims behind this site, which means no number and no confident sentence goes on a page about either.
We do not know whether any answer engine consumes Dataset, SoftwareApplication, or most of the long tail of schema.org types. There is no evidence either way, no platform has commented, and we ship them because they are cheap and descriptive rather than because we have seen them work.
We do not know whether Perplexity uses schema markup at all. Perplexity has published crawler documentation and nothing on parsing. Every confident answer you will find on this question is inference presented as fact.
The honest summary is that schema in 2026 is a hygiene and identity layer with one platform endorsement, one platform denial, and silence from the rest. Ship the six types above, spend a day on your entity graph, and put the rest of your effort into the thing that is actually confirmed to matter, which is having something worth quoting on the page.
Frequently asked questions
Does schema markup help you rank in AI Overviews?
No, by Google's own statement. Google's generative AI guide says structured data is not required for generative AI search and there is no special markup to add. Schema still earns its place for entity disambiguation and for classic rich result eligibility, and Microsoft has separately said it helps its own models parse content, but nobody should be selling you schema as an AI Overviews tactic.
Is FAQ schema dead?
The rich result is, the section is not. FAQ rich results stopped appearing on 7 May 2026 and reporting was removed in June 2026. The markup is harmless and produces nothing visible in Google. Keep FAQ sections on your pages because the question and answer shape matches how retrieval queries are formed, and treat the markup as a free extra.
Does Perplexity use schema markup?
Unknown. Perplexity documents its crawlers and has said nothing about how it parses pages. Anyone asserting a definite answer here is guessing, in either direction.
What is the single highest value schema block to add?
A Person block for your authors with one identifier reused everywhere that author appears, assuming you publish under named authors at all. Author identity fragmented across properties is the most common expensive schema mistake we see, and it is invisible until you go looking for it.
Can I mark up testimonials I have collected privately?
Only ones a named client has approved in writing for public use. Anything short of that should not carry review markup, both because of the Google manual action risk and because the FTC rule on consumer reviews is in force and being enforced. Anonymised outcome data with a methodology note is a stronger asset anyway, and it carries no legal exposure.
The checker reads your live schema graph and flags the two failures that matter most: missing or fragmented entity identifiers, and markup describing content that is not on the page.
Last updated: 2026-08-06
See what a crawler sees
on your own site
Paste a URL and get the extractability read: what an assistant can actually retrieve, and what it cannot.
Free. No signup, no email capture.