Gebouwd om geciteerd te worden
The short version: Quotability is not a formatting trick. Answer engines repeat what is original, specific, and attributed, and pass over what only looks tidy. A page can carry every structural signal in the current playbook and still hand a machine nothing worth lifting, because structure only helps a claim that was already worth repeating travel further.
Why does structure get the credit that substance actually earned?
The current advice for getting quoted by an AI answer engine reads like a formatting checklist: a bolded answer block, a question-shaped heading, a tidy list, a sprinkle of schema markup. None of it is wrong, exactly. It is aimed at the wrong layer of the problem, and the evidence for that is more direct than most of the advice admits.
A page can check every box on that list and still say nothing a machine has not already read somewhere else: a bolded answer restating the obvious, a heading shaped like a question nobody actually asks, a list of tips gathered from other lists of tips. The wrapper is not the problem. What is inside it is.
Ahrefs tested roughly 1,885 pages that added schema markup on its own, no new claim, no new number, just a machine-readable label wrapped around content that already existed, and found no reliable citation lift from the addition alone (Search Engine Roundtable, to re-verify). Schema is close to pure structure: a wrapper with nothing new inside it. If structure alone were what answer engines rewarded, this is exactly where the effect should have shown up plainly. It did not.
What did the first real study of this actually find?
A Princeton, Georgia Tech, and IIT Delhi team published one of the first empirical studies of what actually moves a source's standing inside a generated answer, presented at KDD 2024. Working through a range of ways a page might be changed, the result was specific rather than vague: adding cited statistics and direct quotations lifted a source's visibility inside the generated answer more than formatting changes did (arXiv 2311.09735, to re-verify).
Read plainly, that finding puts content ahead of its container. The lift came from giving the answer engine something to attribute: a number with a source behind it, a quotation with a name attached, not a cleaner shell around the same recycled claim. Wrapper changes were the smaller effect. Content changes were the larger one. That is the reverse of the order most guidance still assumes, and it is worth taking seriously precisely because it is one of the few academic studies to test the question directly, rather than infer it from correlation.
That distinction matters because most of what circulates about AI visibility comes from vendors with a product to sell, or a single dashboard describing its own results. An academic study, tested across many pages rather than one client's outcome, is a rarer kind of evidence, and it happens to point away from the wrapper and toward the words inside it.
Why does formatting still get most of the credit?
Partly because it is the easier half to sell and standardize: a template, an audit, a checklist item that either exists on a page or does not. Structure is visible, teachable, and finishable in an afternoon, which makes it a comfortable place to focus.
And partly because formatting and substance tend to travel together, which makes them easy to confuse for one another. A team with the discipline to run original research, name a source, and state a precise number is also, usually, the kind of team that bothers to structure the page well. The structure did not cause the citation. It rode along with the substance that did, and it is easy to mistake the passenger for the driver.
None of this makes structure worthless. It has a real, narrower job: making a claim that is already worth lifting easy for a machine to find and extract without guessing. How to get cited by AI covers that mechanic in full: the answer block, the question-shaped heading, the source attached to every figure. What structure does not do, and was never going to do, is manufacture a claim worth lifting out of one that was not there to begin with.
What actually makes a claim worth lifting?
Strip the formatting question away and three things are left, and each is a property of the claim itself, not the page wrapped around it.
- Originality. A number, a finding, or a stance that exists nowhere else, so an engine composing an answer has nowhere else to get it from.
- Specificity. A precise claim rather than a restated consensus, worded so a machine does not have to guess whether it is the real answer or one of a dozen near-identical phrasings already circulating.
- Attribution. A name, a source, and a date attached to the claim, so it can be repeated without the engine taking on the risk of unverified authority.
A page can have all the formatting in the world and still fail on all three. A page with none of the formatting and all three still tends to get found, quoted, and paraphrased anyway, because the machine's actual job is composing a trustworthy answer, not rewarding a template.
Why do most citations point to pages a brand does not own?
The pattern holds once the frame widens past a brand's own website. Muck Rack's review of AI citation behavior found that the large majority, about 94 percent, of citations point to sources a brand does not own (Muck Rack, to re-verify): press coverage, independent analysis, forum and community discussion.
That figure is difficult to square with a formatting theory of citation. A brand cannot add schema markup to a journalist's article or a stranger's forum comment. What it can do is say something specific, original, and checkable enough that a journalist, an analyst, or a commenter repeats it unprompted, on a page the brand never touched. Once a claim is being repeated independently, in multiple places, an engine composing an answer has more confirmation to draw on than the brand's own page could ever supply alone. The formatting on that one page was never going to decide the other 94 percent.
What does this mean for what gets published next?
Visibility is no longer ranking makes the case for why citation replaced ranking as the finish line, and how a brand should measure that shift once it accepts the premise. This piece has a narrower job: insisting on where the credit belongs once a brand decides to chase that finish line at all.
The reliable move is not a smarter template. It is a truer claim: specific enough to be worth stating, original enough that no one else has already said it, and attributed clearly enough that repeating it costs an engine nothing in credibility. Structure can carry that claim further once it exists. It has never been able to invent it, and no amount of formatting discipline will change that order.
A brand chasing formatting alone tends to plateau: tidy pages, modest results, no clear reason either way. A brand chasing a truer claim has somewhere to go, because an original claim keeps earning fresh mentions each time someone else decides it is worth repeating on a page of their own.
Veelgestelde vragen
Slechts in beperkte mate. Opmaak helpt een machine om een bewering die al de moeite waard is om te herhalen te vinden en eruit te lichten; ze creëert die bewering niet. Ahrefs testte ongeveer 1.885 pagina's waaraan uitsluitend schema-markup was toegevoegd en vond geen betrouwbare toename in citaties, wat dicht in de buurt komt van een gecontroleerde test van opmaak op zichzelf (nog te verifiëren).
Een team van Princeton, Georgia Tech en IIT Delhi ontdekte dat het toevoegen van geciteerde statistieken en directe citaten de zichtbaarheid van een bron binnen een gegenereerd antwoord sterker verhoogde dan aanpassingen aan de opmaak. Inhoud presteerde beter dan de verpakking ervan, wat het omgekeerde is van de volgorde die de meeste adviezen over citeerbaarheid nog altijd veronderstellen (arXiv 2311.09735, nog te verifiëren).
Niet nutteloos, maar niet voldoende op zichzelf. Schema helpt een machine om correct te verwerken wat al op de pagina staat; het voegt geen nieuwe bewering toe. De test van Ahrefs met pagina's waaraan uitsluitend schema was toegevoegd, vond geen betrouwbare toename in citaties. Dat suggereert dat schema alleen loont wanneer het inhoud labelt die al sterk genoeg is om het citeren waard te zijn (nog te verifiëren).
Drie dingen: originaliteit, een feit of standpunt dat nergens anders bestaat; specificiteit, een precieze bewering in plaats van herhaalde consensus; en bronvermelding: een naam, bron en datum eraan gekoppeld, zodat de bewering zonder risico kan worden herhaald. Een pagina kan elke best practice op het gebied van opmaak toepassen en toch op alle drie punten tekortschieten.
Uit het onderzoek van Muck Rack naar het citeergedrag van AI bleek dat ongeveer 94 procent van de citaties verwijst naar bronnen die een merk niet zelf bezit: berichtgeving in de pers, onafhankelijke analyses en discussies binnen een community (nog te verifiëren). Een merk kan het artikel van een ander niet opmaken. Wat het wel kan doen, is iets zeggen dat specifiek en origineel genoeg is dat anderen het ongevraagd herhalen, en dat is precies wat een antwoordmachine uiteindelijk citeert.
Omdat een machine een bewering nog altijd moet vinden en eruit lichten voordat ze die kan herhalen, en een heldere structuur is wat dat mogelijk maakt zonder gissen. Structuur creëert niet de waarde van een bewering; ze levert die waarde af. Een goed gestructureerde pagina die op een zwakke bewering is gebouwd, faalt nog steeds, en een sterke bewering die begraven ligt in een muur ongestructureerde tekst wordt nog steeds gemist.
Het kan helpen dat een pagina in overweging wordt genomen, omdat veel antwoordmachines putten uit dezelfde index als zoekmachines. Maar een goede positie garandeert geen citatie. Wat de citatie bepaalt, is of de pagina een bewering bevat die origineel, specifiek en voldoende toegeschreven is om veilig te herhalen. Positie zorgt ervoor dat een pagina wordt opgemerkt. Inhoud zorgt ervoor dat ze wordt geciteerd.
Ze moet daadwerkelijk bestaan. Originaliteit betekent hier een cijfer, een bevinding of een standpunt dat niemand anders heeft gepubliceerd, geen herschreven versie van iets dat al circuleert. Een merk kan de schijn van originaliteit fabriceren met slimme formuleringen, maar een antwoordmachine die veel bronnen vergelijkt, ontmaskert een herhaalde bewering eerder dan dat ze die beloont.
Citeerbaarheid behandelen als een sjabloon om te installeren in plaats van een bewering om te verdienen. Het toevoegen van een antwoordblok, een vraagvormige kop en schema-markup rond een herhaalde consensus binnen de sector vinkt elk opmaakvakje af, en geeft een machine nog steeds niets wat de moeite waard is om over te nemen, want geen van die vakjes creëert de originele, specifieke en toegeschreven inhoud waar een antwoordmachine daadwerkelijk naar zoekt.
Nee, het betekent het tegenovergestelde: citeer echte bronnen, voeg namen en datums toe, en voeg originele bevindingen toe waar een merk die daadwerkelijk heeft. Het citeren van derden is zelf een vorm van bronvermelding die een bewering veiliger maakt om te herhalen. Het betoog keert zich tegen lege opmaak, niet tegen bewijs. Het vraagt om meer daarvan, niet minder.