Quick Answer: To write FAQ pages that appear in AI answers, source real questions from support tickets and search data, phrase each the way a customer types it, and write answers that stand alone. AI engines lift the first sentence out of the page, so it must answer the question by itself.

Most guides on this topic hand you a block of JSON-LD and call it a strategy. That is the wrong half of the problem. Schema tells a machine which text is a question and which text is the answer; it does nothing to make the answer worth quoting. This page is about the words, not the markup: how to find the questions people really ask, how to phrase them the way they are actually typed, and how to write an answer that still makes sense after an engine tears it off the page.

The distinction matters because of how the answer surfaces work. When ChatGPT, Perplexity, or a Google AI Overview builds a response, it does not paste your FAQ section into the reply. It pulls a single passage that answers the sub-question it is solving, drops the rest, and attributes the fragment. Your answer travels alone or it does not travel at all.

The rich result you are writing for is already gone

Before the craft, clear the myth, because it steers a lot of wasted effort. Most owners still write FAQ sections to win the expandable question-and-answer box that used to sit under a Google listing. That box is gone for you. In August 2023, Google announced that FAQ rich results would show only for authoritative government and health websites; for every other domain, the visual result quietly stopped appearing. Chasing a snippet Google no longer serves your site is optimising for a screen that will never render.

So why write FAQ pages at all now? Because the extraction layer moved. Search Engine Land's Jenn Mathews put the shift plainly in her October 2025 piece on the rise and fall of FAQ schema: the rich result died, but structured question-and-answer content is exactly what language models reach for when they assemble a reply. That audience does not care whether your markup validates; it cares whether your answer is clear, complete, and safe to quote. The markup question is real but separate, and we handle it in the companion guide to FAQ schema for AI search. This page stays on the writing.

Where the real questions come from

The biggest quality gap between FAQ pages that get cited and ones that get ignored is where the questions come from. Weak pages invent questions a marketer wishes people asked. Strong pages transcribe questions people already asked, in the customer's own words, so they match the phrasing of real searches.

You are sitting on more of this raw material than you think.

Six places your real questions already live

  • Your support inbox and ticket system. Every repeated email is a question worth publishing. The subject lines are often the question, pre-phrased.
  • Live chat logs. Chat is where people type the way they talk. The awkward, specific, half-spelled questions here are gold, because that is how they query an AI too.
  • Sales call notes. The objection a prospect raises on the third call is a high-intent question. If it decides a sale, it decides a search.
  • On-site search queries. Whatever people type into your own site search is a list of things they expected you to answer and could not find.
  • Google Search Console. The Performance report shows the actual queries already bringing people to your pages, in their exact wording. Sort by impressions to find questions you rank for but answer badly.
  • People Also Ask. Google's own expandable question box surfaces the adjacent questions searchers open next. Use it for coverage, not for copy.

Notice what is not on that list: keyword tools that rank questions by search volume, and competitor pages you might reword. Both hand you the generic phrasing everyone already published, which is the near-duplicate content AI systems have no reason to cite over the original. Reading your own records for an afternoon is unglamorous, and it is the whole difference between a page that sounds like your customers and one that sounds like a guess.

Phrase the question the way a person types it

A question heading is doing retrieval work. When someone asks an assistant a question, the engine matches their phrasing against candidate passages, and your heading is the strongest phrasing signal you control. Write it as the full question a real person would type, not a keyword fragment or a slogan.

Three failure modes show up constantly. The first is the keyword stub: a heading that reads "Pricing" or "Shipping times" instead of the question a human would actually ask. The second is the promotional question, the one no customer has ever voiced, such as "Why are we the best choice for your project?" That is an advertisement wearing a question mark, and both readers and Google's helpful content assessment recognise it. The third is the vague question, where "How does it work?" could belong to any business on earth and matches nobody's specific search.

The same question, guessed and then sourced

Guessed by a marketer: "What are the benefits of professional foundation repair?"

Taken from a real chat log: "Do I need to fix a foundation crack if it is not leaking yet?"

The first matches almost no real query and invites a sales answer. The second is a specific decision a homeowner is stuck on at eleven at night, phrased the way they would ask an assistant. Publish the second, answer it honestly, and you have written something an engine can lift and a person actually searched for.

Keep the question in the customer's register. If they say "sea can" and not "shipping container," use "sea can" in at least one question, because that is the retrieval term they will type. Our AI content optimization guide works through this matching in more depth, but the short rule holds: if you would not say the heading out loud to a customer, do not publish it as a question.

Write the answer so it survives being lifted off the page

Here is the mechanic that governs everything. An AI engine does not quote your FAQ section. It quotes one answer, alone, with no question above it and no paragraph before it. If your answer only makes sense while sitting under its heading, surrounded by the rest of the page, it breaks the moment it is extracted, and a broken passage does not get selected.

The practical rule is to write every answer so a stranger could read it with no idea what page it came from and still understand it. That means restating the subject inside the answer instead of leaning on the heading, and never opening with "It depends," "Yes," "This," or "As mentioned above," because each points at context the engine has already thrown away.

Before and after: an answer that cannot travel, and one that can

Question: How long does foundation repair take?

Before, context-dependent: "It depends on the method. As covered above, the timeline varies quite a bit."

After, self-contained: "Most residential foundation repairs take one to three days on site, depending on the method. Interior crack injection is often finished in a day, while underpinning a settled footing can run three days or more. A structural assessment before the work sets the realistic schedule."

The first answer is useless the instant it leaves the page, because "it" and "as covered above" refer to text the engine discarded. The second names its own subject, gives a specific range, and would read correctly if it appeared in an AI answer with nothing around it. Only one of these is quotable.

This is the discipline behind Vector 4, Embed: writing the exact passage an engine can extract cleanly, the specificity that corporate copy usually sands off. To see which passage an engine reaches for and why, we keep a research note on where AI engines read on a page. A self-contained answer is the price of admission, and most FAQ pages never pay it.

Matt Griffin, Formative Digital: "When we audit an FAQ page that is getting no citations, the answers are almost always fine as English and useless as passages. They start with 'Yes, absolutely,' or they say 'our process' without ever naming the process, so the second you read them on their own they collapse. The fix is not more words. It is rewriting the first sentence so it names its own subject and answers the whole question before the reader needs anything else on the page. That one habit moves more than any schema change we ship."

How long should an FAQ answer be?

An FAQ answer should run roughly 40 to 100 words. That window is not arbitrary. Below about 40 words an answer usually skips the specifics that make it worth quoting, and above about 100 it stops being a clean lift and starts being a passage an engine has to trim, which makes selection less likely. The pages ranking for this topic converge on the same range, and it matches what the answer engines actually pull.

Structure inside that window matters more than the count. Lead with the direct answer, then spend one or two sentences on the qualifier, the exception, or the number that proves you know the subject. Never open with throat-clearing and never bury the answer in the last line, where scanning readers and extracting engines both miss it. If an answer keeps growing past 150 words, you have two questions fused under one heading, and splitting them gives two clean answers instead of one bloated one.

Why the first sentence carries the whole load

The first sentence of an FAQ answer is the most valuable sentence on the page, because it is the one most likely to be read, extracted, and quoted. This is not an AI-era invention. The Nielsen Norman Group has documented for years that people do not read online but scan, and will stop reading at any point and still need the main point. Their prescription is the inverted pyramid: conclusion first, detail after. The answer engines formalised a habit good web writers already had.

There is measurement behind this beyond usability studies. Aggarwal and colleagues, the researchers who defined Generative Engine Optimization, tested content changes against a benchmark of generative-engine queries and reported visibility gains of up to 40%. The moves that worked were not technical tricks; they were quoting real sources, adding relevant statistics, and citing authoritative references. All three live in the writing, and all three land hardest early in the answer rather than buried at the end. Front-load the substance and both the human scanner and the machine extractor get what they came for from the first line.

This maps to Vector 4: Embed

Embed is the Formative Digital vector for writing the answer an engine extracts: a self-contained first sentence, a specific fact, and a passage that reads correctly in isolation. It is editorial work, not markup work, which is why the writing decisions on this page carry more weight than any schema field. The full 12 Vectors methodology sits in our research library.

An FAQ that helps a reader, and an FAQ that pads a page

There is a clean line between an FAQ section that earns its place and one that exists to inflate word count. A helpful FAQ answers questions the page did not already cover, in the reader's language, with information they can act on. A padding FAQ restates the sales pitch as fake questions, repeats what the body already said, or stuffs keywords into headings nobody would ever type.

Google's own AI features guidance draws the line for you. It tells writers to make content for people, to offer a take beyond common knowledge, and not to publish what a generative model could produce on its own. An FAQ padded with generic questions fails all three at once, and the same guidance confirms structured data is not required for AI features, so you cannot markup your way out of thin writing. The padding is not neutral: a block of low-value questions dilutes the useful answers around it and drags the whole page down under the helpful content assessment.

The test is simple to apply. For every question you are about to publish, ask whether a real customer has ever asked it and whether your answer tells them something the rest of the page did not. If either answer is no, cut the question. A tight page of six real answers will out-cite a padded page of twenty every time.

A checklist you can run on every answer

The craft above compresses into a short pass you can run on any answer before it goes live. None of it needs a tool or a developer, only reading your own answer as though you had never seen the page it lives on.

StepCheckWhat good looks like
1. SourceDid a real customer ask this?The question is traceable to a ticket, chat, call, or Search Console query
2. PhrasingIs the heading the full question, in their words?You would say it out loud to a customer without wincing
3. First sentenceDoes it answer completely on its own?No "it depends," "yes," "this," or "as above" opening
4. Self-containmentWould a stranger understand it with no page around it?The subject is named inside the answer
5. LengthIs it between 40 and 100 words?Answer first, one or two sentences of support, no bloat
6. Specific factIs there a number, range, or named detail?Something only a real practitioner would know
7. Non-duplicationDoes it add to what the body already said?New information, not a restatement or a pitch

Before any of those steps, run one decision at the top. Not every question belongs in an FAQ, and forcing them there is its own form of padding.

Decision: should this even be an FAQ entry?

  • Keep it as an FAQ when the question is short, self-contained, and has a direct answer that stands alone in a paragraph.
  • Give it its own page or section when the honest answer needs more than 150 words, images, or steps. A cramped FAQ answer to a big question serves nobody.
  • Cut it entirely when the only reason it exists is to place a keyword or repeat a selling point. That question is padding, and padding costs you.

Want to know which of your answers AI can actually quote?

Send us your domain and we will check what ChatGPT, Perplexity, Google AI Overviews, and Gemini currently pull from your pages, then mark which FAQ answers are self-contained enough to be cited and which collapse the moment they are lifted. You will see the exact passages that are and are not working.

What this actually takes, and what to expect

An honest account of the effort, since most advice on this topic skips it. Sourcing and writing a good FAQ page of ten to fifteen answers is about a half-day for someone who knows the business: an hour or two reading support and chat records, the rest spent writing and cutting. It is real work, and it cannot be handed to a model to finish, because the model does not have your customers' exact questions or your practitioner's exact numbers.

The result does not arrive the week you publish. Content-driven visibility compounds slowly, and pretending otherwise is the tell of a bad agency. One Ontario shipping container dealer we work with logged 4,810 clicks and 627,000 impressions across a 16-month window (Google Search Console, 16-month window, shared with permission). Daily clicks sat in single digits for roughly the first year before the curve turned, reaching seventy-click days with six to nine thousand daily impressions by early summer 2026. Well-written answers are one asset feeding a curve of that shape, not a switch that flips.

Two honest qualifiers belong here. Outcomes vary by trade, market, and how much genuine expertise you put into the answers, so treat the timeline as a pattern, not a promise. And getting quoted by name is a larger project than one FAQ page; if that is the goal, our guide to getting your business mentioned by ChatGPT covers the entity and citation work that sits alongside the writing.

Common questions about writing FAQ pages

How many questions should an FAQ page have?

Enough to cover the questions people genuinely ask, and no more. A page with eight real questions and complete answers outperforms one with thirty invented ones, because padding dilutes the signal and trips Google's helpful content assessment. Start with the questions you can source from real customer contact, and stop when you run out of real ones.

Do I still need FAQ schema if Google removed the rich result?

The schema no longer earns most sites a visual rich result, since Google limited that to government and health domains in August 2023. It can still help machines parse your questions and answers cleanly, which is a markup decision covered in our FAQ schema guide. The writing is what gets extracted; the schema only labels it.

Where do I find the questions people actually ask?

Your support inbox, live chat logs, sales call notes, and on-site search queries are the richest sources, because they record the exact words customers use. Google Search Console shows the queries that already bring people to your site, and the People Also Ask box shows adjacent questions. Skip the questions only a marketer would write.

How long should each FAQ answer be for AI answers?

Aim for roughly 40 to 100 words per answer. That is long enough to answer completely and short enough for an AI engine to lift whole. Lead with the direct answer in the first sentence, then add one or two sentences of support. If an answer runs past 150 words, it is usually two questions wearing one heading.

Can I have ChatGPT write my FAQ answers for me?

You can draft with it, but publishing the output unedited is the fastest way to blend into the answers everyone else generated the same way. Google's AI features guidance says plainly not to publish content a generative model could produce on its own. Use your real customer language and specifics, which a model does not have, and the answer becomes citable.

Get your free AI visibility audit

Formative Digital, Brantford, Ontario

If you have written an FAQ page and it is not showing up in AI answers, the useful first step is seeing which of your answers an engine can actually quote and which fall apart when lifted. We will audit what the search and AI systems currently pull from your site and hand you the specific rewrites that would make more of it citable. Call 226-450-2065 or send the request through the form above.

Request your free AI visibility audit