Research
the first real study of GEO: what actually earns AI citations
Most advice about ranking in AI answers is folklore. One piece of it is not. In 2024, researchers from Princeton, IIT Delhi, Georgia Tech and the Allen Institute for AI published the study that coined the term generative engine optimization, testing nine content tactics across a benchmark of 10,000 queries. The headline: adding quotations, statistics and cited sources lifted a page's visibility in AI answers by up to 40 percent, while old-school keyword stuffing performed worse than doing nothing at all.
what the researchers actually did
The team built GEO-bench, 10,000 real queries across multiple domains, each paired with the web sources a generative engine would draw on to answer. They then modified those sources using nine tactics, from adding statistics to stuffing keywords, and measured how each change affected the source's share of the generated answer, checking results against a real deployed engine. It is the closest thing this field has to a controlled experiment.
what won
Three tactics stood clear of the rest. Adding relevant statistics, adding quotations from credible sources, and citing sources within the content each lifted visibility on the order of 30 to 40 percent in the study's metrics. A fourth finding deserves more attention than it gets: simply improving fluency, rewriting for clearer prose with no new information added, produced gains around 28 percent. Clear writing is not a nicety. It is machine-legibility.
what lost
Keyword stuffing scored roughly 8 percent below the unmodified baseline. The tactic that defined a decade of bad SEO is actively counterproductive with language models, which parse meaning rather than count matches. If your content still carries that habit, it is now a cost.
the honest catch
One study, one benchmark, an engine built to mimic real ones, and results validated but not exhaustively proven across every platform. Treat the direction as strong evidence and the exact percentages as laboratory numbers rather than guarantees. What makes these findings credible is that they converge with how retrieval systems mechanically work, and nothing published since has overturned them.
Common questions
Do these tactics work on every AI platform?
- The study tested a Bing-style engine and validated on Perplexity. Platforms differ, but the winning tactics all increase clarity and verifiability, which every retrieval system rewards by design.
Should I add statistics to every page?
- Add true, sourced, relevant ones where they genuinely support the answer. The study rewarded substance, and invented or decorative numbers are a trust risk with humans and machines alike.
Where should I start applying this?
- Your highest-intent pages. Add a named source, a real statistic and a clear quotation, and rewrite for plainness. That is the study's entire practical prescription.

Tom Claydon
Co-founder of Nudge
Ask Tom anything about search, he answers on WhatsApp.