Short answer: Schema markup will not get you cited by AI search on its own. Google says AI Overviews and AI Mode need no special schema, though it still recommends structured data for rich results. Microsoft has said schema helps its AI systems understand content. OpenAI's crawler and publisher docs and Perplexity's crawler docs say nothing about schema. Add accurate markup that matches the visible page. Treat it as support for good content, not a shortcut to citations.

What does each AI search engine actually say about schema?

I read each company's own documentation and noted what it says and what it leaves out. The table shows the result as of 2 October 2026.

EngineWhat it has said about schemaSourceWhat is unknown
Google (AI Overviews, AI Mode)No special schema.org markup is needed. Structured data is still worth using for rich results, and it must match the visible text.Google Search Central docs, 2025 and 2026Whether markup changes the odds of being cited in an AI answer
Microsoft (Bing, Copilot)Schema helps search engines and AI systems understand content. Listed as a practice that improves eligibility.Microsoft Bing product manager blog, October 2025; SMX Munich talk, March 2025Which types or properties matter, and how much weight they carry
OpenAI (ChatGPT search)Nothing. The docs ask you to allow OAI-SearchBot.OpenAI crawler docs and Help Center FAQWhether ChatGPT search reads JSON-LD at all
PerplexityNothing. The docs ask you to allow PerplexityBot and its IP ranges.Perplexity crawler docsWhether Perplexity reads JSON-LD at all

Only Microsoft has said directly that schema helps its AI systems understand content. Google has said markup is not required for its AI features, though a 2025 Google post called it useful. The other two are silent. Silence does not mean they ignore markup. It means nobody outside those companies can say for sure.

What exactly did Google say about structured data and AI features?

Google's page AI features and your website is direct. It says: "There are no additional requirements to appear in AI Overviews or AI Mode, nor other special optimizations necessary." A few lines later it adds: "There's also no special schema.org structured data that you need to add."

The same page still lists structured data among the basics. One of its listed fundamentals reads: "Making sure your structured data matches the visible text on the page." So Google treats markup as normal SEO hygiene, not as an AI lever.

In May 2026 Google went further. A new page, Google's Guide to Optimizing for Generative AI Features on Google Search, has a section of myths you can ignore. One of them is labelled "Overfocusing on structured data". The text says:

Structured data isn't required for generative AI search, and there's no special schema.org markup you need to add. However, it's a good idea to continue using it as part of your overall SEO strategy, as it helps with being eligible for rich results on Google Search.

An earlier Google post, Top ways to ensure your content performs well in Google's AI experiences on Search (21 May 2025), gives a softer hint. It calls structured data "useful for sharing information about your content in a machine-readable way that our systems consider". It stops short of saying markup helps you get cited.

Google's introduction to structured data explains the wider role. It says Google uses markup "to understand the content of the page, as well as to gather information about the web and the world in general, such as information about the people, books, or companies that are included in the markup." That is the strongest official reason to keep your entity markup clean.

Did Microsoft really confirm that schema helps its LLMs?

Yes, but read the details. In March 2025, Fabrice Canel of Microsoft Bing spoke at SMX Munich. Search Engine Land reported a LinkedIn summary of the talk: "Fabrice Canel confirms that schema markup helps Microsoft's LLMs understand your content". The report says Canel later confirmed this in the comments. That is a conference remark passed on second-hand. I could not find a Microsoft page with that exact sentence.

Canel also retired from Microsoft on 1 July 2026. Treat his remark as what Bing said at the time, not as current policy.

A firmer source came in October 2025. Krishna Madhavan, a Principal Product Manager at Microsoft Bing, wrote Optimizing Your Content for Inclusion in AI Search Answers. It states: "Schema is a type of code that helps search engines and AI systems understand your content." Its checklist says: "Structure your content: Use schema, clear headings, and modular layouts."

The same post is careful about promises. It says: "While there's no secret sauce that guarantees selection in AI answers, there are clear practices that improve eligibility." So Microsoft frames schema as one input, not a guarantee.

What do OpenAI and Perplexity tell site owners?

Both companies talk about crawler access, not markup. The OpenAI crawler documentation says OAI-SearchBot "is used to surface websites in search results in ChatGPT's search features." Sites that block it "will not be shown in ChatGPT search answers, though can still appear as navigational links." I searched the page text for schema, JSON-LD and markup. None of those words appear.

The OpenAI Publishers and Developers FAQ says: "Any public website can appear in ChatGPT search." Its main instruction is to avoid blocking OAI-SearchBot. It also has no mention of structured data.

The Perplexity crawler documentation follows the same pattern. It recommends "allowing PerplexityBot in your site's robots.txt file and permitting requests from our published IP ranges". It also warns that a web application firewall may need to whitelist its bots. There is nothing on schema.

For these two engines, the provable first step is crawler access. Claims that Perplexity "parses schema" circulate widely, but I found no primary source for them.

If AI engines do not require schema, why add it at all?

  • Rich results. Google's AI guide says structured data "helps with being eligible for rich results on Google Search."
  • Entity understanding. Google says it uses markup to learn about the people and companies in it. It also says Organization markup can help it "disambiguate your organization in search results."
  • Bing and Copilot. Microsoft says schema helps AI systems understand content. That is the only engine on record saying so.

My reasoning, not a published rule: an AI answer that names a business has to work out which business it means. Clear markup is a cheap way to state your name, location and profiles in one place. It costs little and carries low risk when it is accurate. That is why I recommend it, while being honest that no engine promises a citation in return.

Which schema types are worth adding to a small business or service site?

For most service businesses, these types cover the useful ground.

  • Organization or LocalBusiness on the home page or about page. Include your name, URL, logo, contact details and address if you serve customers in person. Add sameAs links to your real profiles. Google's Organization documentation says "There are no required properties", so add what is true and relevant.
  • Person for the author or owner. Google's Article documentation recommends an author.url that "uniquely identifies the author of the article", such as a bio page. It adds that "Google can understand both sameAs and url when disambiguating authors."
  • Article or BlogPosting on blog posts and guides, with the author and dates.
  • Service on each service page, describing what you offer and who provides it. This is a schema.org type. Google does not list a Service rich result, so its value is descriptive only.
  • FAQPage only where the questions and answers are visible on the page. Google's documentation changelog says the FAQ rich result "will no longer appear in Google Search starting May 7, 2026." So expect no visual feature from Google. Microsoft still names FAQ among the content types schema can label.

Keep FAQ answers open on the page, not folded away. Microsoft's post warns: "Don't hide important answers in tabs or expandable menus: AI systems may not render hidden content, so key details can be skipped."

What does an Organization block with sameAs look like?

Here is a short JSON-LD example. Replace each placeholder with real details, and list only profiles you control.

<script type="application/ld+json">
{
  "@context": "https://schema.org",
  "@type": "Organization",
  "name": "Example Plumbing Ltd",
  "url": "https://www.example.com/",
  "logo": "https://www.example.com/images/logo.png",
  "email": "hello@example.com",
  "telephone": "+44-20-0000-0000",
  "sameAs": [
    "https://www.example.net/company/example-plumbing",
    "https://www.example.org/profile/example-plumbing"
  ]
}
</script>

Google lists JSON-LD as its recommended format. A local business with a shopfront can use LocalBusiness or a more specific subtype in place of Organization. Add the address in that case.

Why do sameAs links and consistent entity facts matter?

The schema.org definition of sameAs is: "URL of a reference Web page that unambiguously indicates the item's identity." Its examples are a Wikipedia page, a Wikidata entry or an official website. Google describes the property as a URL "with additional information about your organization", such as a social or review profile.

The common thread is identity. Two firms can share a name. A sameAs list ties your site to profiles that only you own. Google's author guidance says the same about people: it reads sameAs and url "when disambiguating authors."

Google's AI guide also points beyond your own site. It says Merchant Center and "Google Business Profiles can help your products and services to be visible in both AI responses and other Google Search results." Its AI features page lists "Checking that your Merchant Center and Business Profile information is up-to-date" as a fundamental.

My reasoning, not a published rule: if your markup, your Business Profile and your social bios give the same name, address and phone number, a system has fewer conflicting facts to resolve. If they disagree, you are relying on the system to guess. No engine has published how it weighs these conflicts. Consistency is simply the safer position.

Which schema mistakes can hurt you?

Bad markup can do more harm than no markup. These are the mistakes Google's rules name directly.

  1. Markup that does not match the page. Google's intro says not to "add structured data about information that is not visible to the user, even if the information is accurate." The general structured data guidelines add: "Your structured data must be a true representation of the page content."
  2. Fake or self-serving reviews. The review snippet guidelines say: "Don't include fake or undisclosed incentivized reviews on your page or in your structured data markup." If a business controls reviews about itself, its LocalBusiness or Organization pages "are ineligible for star review feature." So an aggregateRating for your own business on your own site earns no stars. Google also warns that ratings not by actual users may result in a manual action.
  3. Misleading identity claims. The guidelines say: "Don't impersonate any person or organization, or misrepresent your ownership, affiliation, or primary purpose." A sameAs link to a profile you do not own could fall foul of this.
  4. Empty pages built only to hold markup. Google says: "Don't create blank or empty pages just to hold structured data".

What does a violation cost? Google says a structured data manual action means a page "loses eligibility for appearance as a rich result; it doesn't affect how the page ranks in Google web search." Separately, Google clarified in May 2026 that "our spam policies also apply to generative AI responses in Google Search." Deceptive markup is not a safe experiment for AI visibility.

How can you check what schema your pages already have?

My free AI visibility checker reads your homepage's JSON-LD types and lists its sameAs links. That shows whether your Organization block exists and where it points. Google's Rich Results Test is the next step for validating syntax against Google's features.

Then compare the markup with the page itself. Every fact in the JSON-LD should appear in the visible text. If you want help with the wider picture, my answer engine optimization service covers markup alongside content and crawler access. For how the three disciplines differ, see SEO vs GEO vs AEO.

Frequently asked questions

Does schema markup help you appear in Google AI Overviews?

Google says no special schema is needed for AI Overviews or AI Mode. To be shown as a supporting link, a page must be indexed and eligible to show with a snippet. Google still advises using structured data for rich results and keeping it matched to the visible text. No Google document says markup raises your chance of being cited in an AI Overview.

Does ChatGPT use schema markup?

OpenAI has not said. Its crawler documentation and publisher FAQ never mention schema, JSON-LD or markup. What OpenAI does say is that sites blocking OAI-SearchBot will not be shown in ChatGPT search answers, though they can still appear as navigational links. So the first step for ChatGPT visibility is allowing that crawler in robots.txt and in any firewall rules on your server.

Is FAQPage schema still worth adding in 2026?

Only where the questions and answers are visible on the page. Google stopped showing the FAQ rich result on 7 May 2026, so the markup no longer earns a visual feature there. Microsoft still lists FAQ among the content types schema can label for AI systems. Keep the answers open on the page rather than hidden in tabs or accordions.

What is sameAs in schema markup?

sameAs is a property that lists URLs which identify the same thing as your page describes. Schema.org defines it as a reference page that unambiguously indicates the item's identity. For a business, that usually means its official social and directory profiles. Google says it reads sameAs when telling authors apart, and that Organization markup helps it disambiguate a business.

Can wrong schema markup hurt my site?

Yes. Google can issue a manual action for structured data that breaks its guidelines. That removes rich result eligibility, though Google says it does not change web rankings. Common causes are markup that does not match the page and fake or self-serving reviews. Google also says its spam policies now apply to generative AI responses in Search.

Do I need llms.txt or special AI markup as well as schema?

Not for Google. Its documentation says you do not need new machine-readable files, AI text files or special markup to appear in AI Overviews or AI Mode. Google lists llms.txt among tactics you can ignore for Google Search. Google says it is fine to keep such files for other services. None of the other engines covered here publish a requirement for them.