AI Search & Google Rankings
The ‘Best Of’ Listicle That Ranks Itself #1: Why Google and ChatGPT Both Keep Rewarding It
We audited Google’s top 10 results for three “best X” queries and checked them against a 5-million-query study of ChatGPT’s hidden searches. The same self-serving format wins in both engines, and neither engine has a mechanism that can stop it.

Nine of the ten pages Google currently ranks for “best GEO tools 2026” are listicles published by vendors or agencies selling GEO tools, and seven of the nine are vendors. At least four of those pages rank the author’s own product first. Zero of the ten come from an independent reviewer, a publication, or a testing lab.
A day earlier, an SEO posted the complaint that prompted this audit to r/bigseo: “It drives me crazy that self promotional and self serving listicles are actually seeing rankings and citations. Like it seemingly has been an effective strategy. And the other models getting duped by this strategy makes sense but even in Google which should be much more sophisticated seemingly gets duped.” Twenty-one replies followed. The most-voted consensus, from a commenter called Proest_off_the_pros: “They rank/get mentioned because they satisfy search intent… it isn’t going to change. Period.”
The resignation is premature, because it skips the interesting question: why do two systems built by some of the best retrieval engineers on earth both keep rewarding a format whose bias is visible in the headline? The answer is not that both engines are fooled. It is that the self-serving listicle is the native food of both retrieval layers, and we can now show the mechanism on each side with data.
The audit: who actually ranks for “best X”
On September 11, 2026, we ran three commercial “best” queries through Google from US locales and recorded what occupied the top 10 organic positions: best GEO tools 2026, best AI SEO tools, and best on-page SEO tools. The GEO query, in a category this blog’s readers work in every day, produced the cleanest result.
| # | Ranking domain | Publisher | Who holds the #1 slot on the page |
|---|---|---|---|
| 1 | writesonic.com | Vendor | Profound (Writesonic ranks itself #2) |
| 2 | aiclicks.io | Vendor | AIclicks (“#1 GEO Tracking Tool”) |
| 3 | geoptie.com | Vendor | Geoptie (“Best All-in-One GEO Platform”) |
| 4 | bluefishai.com | Vendor | Bluefish (“Enterprise GEO Powerhouse”) |
| 5 | visible.seranking.com | Vendor | Includes SE Ranking among the best |
| 6 | nicklafferty.com | Independent consultant | Profound (affiliate-linked list) |
| 7 | buriedagency.com | Agency | n/a (not itemized in snippet) |
| 8 | primacy.com | Agency | Semrush |
| 9 | ziptie.dev | Vendor | ZipTie (disclosed: “ranked #1 below”) |
| 10 | prerender.io | Vendor | Includes Prerender.io |
The pattern held in the other two queries with more noise. For “best AI SEO tools,” Surfer SEO’s own blog ranks its AI Tracker first and eesel.ai ranks itself first, alongside independent lists from Zapier and OneLittleWeb. For “best on-page SEO tools,” SeedProd’s listicle ranks its own All-in-One SEO plugin first.
Two details make the audit more than a SERP gripe. First, the self-ranking is often announced out loud. ZipTie’s page opens with: “Full disclosure: This guide is published by ziptie.ai/, which is ranked #1 below. We’ve applied identical evaluation criteria to ourselves and every competitor.” The disclosure changes nothing about the page’s performance. Second, Google’s own answer layer participates: the AI Overview shown for “best on-page SEO tools” cited SeedProd’s self-ranking listicle among its sources. The engine the r/bigseo poster called “much more sophisticated” was quoting a vendor’s self-assessment inside its own answer box on the day we checked.
Method note: three queries, one day, US results, no personalization. Self-placement was verified from SERP snippets or by fetching the page; cells we could not verify say n/a instead of guessing. Treat the audit as a dated sample of a persistent pattern, not a census.
ChatGPT’s side: the word “best” is injected before your question is searched
Here is where the data explains the behavior. When someone asks ChatGPT a question, the system frequently rewrites the prompt into several hidden background searches, called query fanouts, before it answers. Peec AI GEO specialist Tomek Rudzki analyzed 5 million of these fanouts collected across ChatGPT, Perplexity and Grok between April 1 and April 21, 2026, and published the words the system adds on its own.
A user who types “which project management tool should our team use?” never asks for a ranking. ChatGPT searches for one anyway, because the fanout injector adds “best,” “top,” “reviews,” “comparison” and the year to its background queries. Retrieval is steered toward listicle-shaped pages before the user’s intent is even consulted.
The second lock comes from how the results get merged. Peec AI reports that ChatGPT combines fanout results with Reciprocal Rank Fusion, an algorithm where a page that appears across many sub-searches outscores a page that wins one of them. A “best X” listicle is the only content format that natively matches every injected angle in a single document: it carries the word “best,” a numbered ranking, a mini-review per item, a comparison table, an alternatives section, and the current year in the title. One page scores on every fanout at once. A vendor comparison page written to rank itself first is not gaming RRF. It is the format RRF is structurally easiest to reward.
The model-by-model numbers explain why this story centers on ChatGPT: Peec AI measured an average of 1.4 fanouts per prompt on Perplexity (which mostly just simplifies the query), 2.1 on ChatGPT, and 6.8 on Grok, which runs a full research brief and targets specific sites with the site: operator, including Reddit, Wirecutter, Consumer Reports and G2. The engines differ in volume, but both ChatGPT and Grok spend their fanouts on exactly the territory listicles occupy.
Google’s side: the policy line that doesn’t exist
Google’s complaint is harder to file, because Google’s own rules do not prohibit what these pages do. The spam policies, in their August 28, 2026 revision, define spam by production and hosting patterns. Three policies are relevant, and the self-serving listicle threads the needle of all three:
| Google spam policy | What it requires | A vendor’s single “best X” page |
|---|---|---|
| Scaled content abuse | Mass production of pages, with or without automation, that lack original value | One page, written about a tested category: outside scope |
| Site reputation abuse | Third-party content published on a host site’s domain to exploit its ranking signals | Published on the vendor’s own domain: outside scope |
| Expired domain abuse | Buying an expired domain to exploit its history | Not applicable |
Google did tighten the definition in one way worth noting: the August 28 revision explicitly includes “attempting to manipulate generative AI responses in Google Search” in its definition of spam. But that clause still attaches to the same scaled-manipulation patterns, not to one vendor saying it is the best. Meanwhile, the August 2026 spam update rolled out in 64 hours without Google saying what it hunted, and the r/bigseo complaint was posted on September 10, after the update finished, by someone observing that the listicles are still ranking and still being cited. Our own spot checks the next day agree. We cannot confirm the update touched no listicle farms anywhere; Google will not say, and community reports are anecdotal. What we can say is that the format the complaint names was visible on page one of a commercial SERP after the update completed. For the full breakdown of what the update confirms and what it does not, see our pieces on the update itself and the penalty-risk question for GEO work.
Ranking systems reward intent satisfaction, engagement and authority. A “best X” page satisfies a searcher typing “best X,” which is exactly what the r/bigseo replier meant by “they satisfy search intent.” Detecting that the author has a financial stake in the ranking would require motive inference at single-page scale, and Google has never drawn that line, because it cannot draw it without also condemning ordinary marketing pages.
The loop that keeps feeding itself
Put the two sides together and you get a compounding cycle rather than a bug:
- ChatGPT’s injector adds “best,” “reviews” and the year, so listicles get retrieved even for neutral questions.
- RRF rewards the one format that matches every injected angle, so listicles get cited.
- Vendors see competitors winning citations and bottom-funnel rankings, and publish their own listicle, ranking themselves first.
- Google keeps ranking the pages because searchers keep clicking them, and its answer layers cite them too.
- Freshness injection rewards cheap date-refreshes, and Peec AI’s finding that updating an already-cited page beats publishing a new one hands the advantage to incumbents.
Each step is individually rational, which is why the reply “it isn’t going to change” carries weight. Nothing in either engine’s published mechanics selects against it.
What you do with this
If you sell a product or service and you have no “best [your category]” page on your own domain, this data says your competitors who do are taking three things from you: bottom-funnel Google clicks, ChatGPT citations, and the framing of your own product inside their comparison, limitations included. The decision is whether to play or to hedge.
Playing it defensibly. The pages in our audit that will age well share traits: a disclosed self-placement (the ZipTie approach), genuine testing notes, real limitations listed for the author’s own product, and coverage of the fanout angles under one URL: best, reviews, comparison, alternatives, the current year. That structure is what RRF rewards, and it is also what keeps the page on the right side of the spam policies, because the value is original rather than mass-produced. Keep the date current: the year token fires in 5.44% of fanouts, and an updated cited page beats a new uncited one.
Hedging instead. If you choose not to publish one, the counter-move is to get into the surfaces the engines fetch on their own: G2, Gartner-style review platforms, Reddit threads, Wirecutter-tier publications. Grok’s fanouts target those domains by name with the site: operator, and Peec AI found ChatGPT searches for review content even when the prompt never asks for it; in their Revolut example, Glassdoor ratings and a weak Sitejabber score shaped the AI’s brand description. Your citations then come from pages you don’t control, which is safer under spam policy and more credible in the answer, but slower to move.
Tracking. Whichever side you pick, monitor both layers: your positions on the commercial “best” queries in Google, and your citation share in AI answers, since the two now feed each other. If your rankings move after a future spam update, our coverage of what the August update targeted is the starting point for diagnosing it.
The validation is yours
One r/bigseo commenter, WebLinkr, offered the frame that survives contact with all of this data: “I think it shows that Search engines and LLMs are not fact validators, that verifying/approving content is beyond their reach.” That is the working model. Google ranks what satisfies the query; ChatGPT retrieves what matches the questions it writes for itself; neither inspects who profits from the answer. The self-serving listicle is not a trick both engines keep falling for. It is the format that matches how both engines actually work, and it will keep winning until one of them changes the mechanics, or until readers and buyers do the validation the engines won’t.




