NEWFree AI visibility report. Tracking from $99/month

How the AI Index is built.

The exact method behind every ranking: the questions we ask, the engines we ask them of, how the score is computed, and the rules that keep it honest.

Methodology v1.3 · current as of September 2026

The question we ask

For every category, each AI model is asked the question a real buyer asks: “What are the best {category}? Recommend the top brands or products that people actually use.” Product categories, whose boards rank individual models, ask for specific models instead: “Recommend the specific models people actually buy”. No brand names are supplied either way. Regional editions ask the same question the way a local buyer would, naming the country.

Every monthly refresh asks each AI model the same buyer question once, and the exact run count behind every edition is published in its JSON record. Every figure is therefore reproducible. One engine’s quirky answer cannot create a ranking on its own: the consensus gate below requires at least two AI models to agree before a brand is ranked at all.

The engines

All 9 leading AI models, every refresh: no engine is skipped because it is inconvenient to reach. The exact model version behind every captured answer is recorded with it.

ChatGPTClaudeGeminiPerplexityDeepSeekGrokCopilotGoogle AIGoogle AI Mode

The score: the full formula

Brands are ranked by their AI Recommendation Score (0 to 100). The complete calculation is published here so any position can be recomputed from the answers. There is no editorial weighting and no hidden step: every brand recommendation is extracted from every captured answer, a brand counts at most once per answer however many of its products that answer names, a consensus gate requires at least two AI models, and the score combines share of voice, mention rate and how early the AI models name the brand.

The formulaScore = 55 × min(share of voice ÷ 40%, 1) + 30 × min(mention rate ÷ 80%, 1) + 15 × min(1 ÷ average position, 1)Average position is where the brand first appears in each answer that names it, averaged. Worked example: a brand holding 30% share of voice, named in half of all answers, typically second: 55 × (30 ÷ 40) + 30 × (50 ÷ 80) + 15 × (1 ÷ 2) = 41.25 + 18.75 + 7.5 = 67.5.

Ties: when two brands land on the same score, the tie breaks on average position (earlier wins), then on total answers naming the brand. Ranks are strictly sequential, so three brands tied on 67.5 rank 2, 3 and 4, never joint 2nd.

Why there is no citation component: our July 2026 audit showed citations reward the wrong thing. Review sites that answers cite as sources outranked products that six of eight AI models actually recommended, so citations were removed from the score.

The gates

A brand must be recommended by at least two different AI models to be ranked at all. One engine’s quirk is not a market position.

Every ranking keeps its receipts: the verbatim answers the AI models gave are stored and shown, so any position can be checked against the answers that produced it.

Each monthly refresh is captured as an immutable snapshot. Movement arrows compare against the previous snapshot; past editions are never edited.

What never ranks: entity exclusions

A ranking must contain entities of the category, so two published exclusion rules run before scoring (added 7 August 2026, dated in the changelog below). First, in service-provider categories, the directories and marketplaces that rank providers are never themselves ranked: an entity is excluded when its domain sits on our curated source table as a directory or marketplace. Second, the per-category exclusion list below removes brands AI models name in passing that are not members of the category, such as an ad platform named inside an agency recommendation. Both rules require the entity’s name to identify it as the domain’s owner, so a real provider an answer happened to attribute to a directory’s domain always stays ranked, and a per-category exception list keeps legitimate incumbents in place (Amazon remains a third-party-logistics provider).

Excluded mentions leave the share-of-voice pool, so remaining scores reflect only category members. Nothing is removed from the record itself: every excluded mention stays verbatim in the published answer corpus and its extractedBrands, so any board remains recomputable from its own record. The list changes only with a dated changelog entry.

digital-marketing-agenciesairbnb.com · cineplex.com · google.com · facebook.com · meta.com · amazon.com · zillow.com · tripadvisor.com · semrush.com · ahrefs.com · hubspot.com · klaviyo.com · toptal.com · forbes.com · mailchimp.com · hootsuite.com · buffer.com · sproutsocial.com · later.com · agorapulse.com · canva.com · adobe.com · youtube.com · linkedin.com · tiktok.com
seo-companiesairbnb.com · cineplex.com · google.com · facebook.com · meta.com · amazon.com · zillow.com · tripadvisor.com · semrush.com · ahrefs.com · hubspot.com · klaviyo.com · toptal.com · forbes.com · mailchimp.com · hootsuite.com · buffer.com · sproutsocial.com · later.com · agorapulse.com · canva.com · adobe.com · youtube.com · linkedin.com · tiktok.com · screamingfrog.co.uk · surferseo.com · clearscope.io · moz.com · yoast.com · rankmath.com · brightlocal.com · conductor.com · botify.com · seranking.com
social-media-agenciesairbnb.com · cineplex.com · google.com · facebook.com · meta.com · amazon.com · zillow.com · tripadvisor.com · semrush.com · ahrefs.com · hubspot.com · klaviyo.com · toptal.com · forbes.com · mailchimp.com · hootsuite.com · buffer.com · sproutsocial.com · later.com · agorapulse.com · canva.com · adobe.com · youtube.com · linkedin.com · tiktok.com
content-marketing-agenciesairbnb.com · cineplex.com · google.com · facebook.com · meta.com · amazon.com · zillow.com · tripadvisor.com · semrush.com · ahrefs.com · hubspot.com · klaviyo.com · toptal.com · forbes.com · mailchimp.com · hootsuite.com · buffer.com · sproutsocial.com · later.com · agorapulse.com · canva.com · adobe.com · youtube.com · linkedin.com · tiktok.com
web-design-companiesairbnb.com · cineplex.com · google.com · facebook.com · meta.com · amazon.com · zillow.com · tripadvisor.com · semrush.com · ahrefs.com · hubspot.com · klaviyo.com · toptal.com · forbes.com · mailchimp.com · hootsuite.com · buffer.com · sproutsocial.com · later.com · agorapulse.com · canva.com · adobe.com · youtube.com · linkedin.com · tiktok.com · shopify.com · wordpress.org · wordpress.com · wix.com · squarespace.com · webflow.com · godaddy.com · duda.co · framer.com
branding-agenciesairbnb.com · cineplex.com · google.com · facebook.com · meta.com · amazon.com · zillow.com · tripadvisor.com · semrush.com · ahrefs.com · hubspot.com · klaviyo.com · toptal.com · forbes.com · mailchimp.com · hootsuite.com · buffer.com · sproutsocial.com · later.com · agorapulse.com · canva.com · adobe.com · youtube.com · linkedin.com · tiktok.com · netflix.com · nike.com · spotify.com · coca-cola.com · cocacola.com · uber.com · samsung.com · premierleague.com · warnerbros.com · target.com · starbucks.com · apple.com · airbnb.com · mcdonalds.com · tiktok.com
pr-agenciesairbnb.com · cineplex.com · google.com · facebook.com · meta.com · amazon.com · zillow.com · tripadvisor.com · semrush.com · ahrefs.com · hubspot.com · klaviyo.com · toptal.com · forbes.com · mailchimp.com · hootsuite.com · buffer.com · sproutsocial.com · later.com · agorapulse.com · canva.com · adobe.com · youtube.com · linkedin.com · tiktok.com
accounting-firmsxero.com · quickbooks.intuit.com · turbotax.intuit.com · turbotax.com · intuit.com · freshbooks.com · waveapps.com · sage.com · wealthsimple.com · netsuite.com · oracle.com · zoho.com · freeagent.com · myob.com · myob.com.au · freetaxusa.com · gusto.com · adp.com · google.com · microsoft.com · sap.com · dext.com · hmrc.gov.uk
property-management-companiesappfolio.com · buildium.com · yardi.com · rentmanager.com · turbotenant.com · realpage.com · entrata.com
financial-advisorsforbes.com · nerdwallet.com
solar-companieslg.com · hyundai.com
real-estate-companieszillow.com

Tamper-evident records

Every refresh freezes its receipts twice. At capture time, the complete verbatim answer corpus is hashed with SHA-256 and the hash is stored on the immutable snapshot and published in the record’s JSON download, next to the answers themselves. Anyone can recompute the hash from the published corpus, using the spec shipped in the same file, and confirm the answers were not edited after publication.

Once a monthly sweep is verified complete, the per-category hashes are combined into a single run-level root hash, published here and inside that month’s Index report. One changed character in any archived answer changes its category hash, which changes the root.

2026-09 · Global822f520398b77b22f14ac88f82061f955288dd1441a39c51b4cf4ef807e793eb241 of 241 hashed
2026-09 · Regional6c4838d7903cd2d288121681f46a96e5806bb332bf9d1f8d3e2db8aa1d763172196 of 196 hashed
Show earlier editions (4)
2026-08 · Globalb0bb234c9d1758a6c8081b0852db5397259c220cefc8a631bbcab670adc2c8f4193 of 193 hashed
2026-08 · Regional7c246f1e6c44f05f86675bd4933b251956cafe64024927afc66b23e33d77696b188 of 188 hashed
2026-07 · Global4b039a3e17571d10a76bc9cf38bda66f1c03cd5213f7b4df74a9a4fba01d5725185 of 185 hashed
2026-07 · Regional9db19f97568846f0679d41bcd384375c3dc8d29c18492130b8c35e98d99a2ab2188 of 188 hashed

Non-determinism, sampling and variance

The same AI model, asked the same question twice, does not reliably give the same answer. This is a property of the models, not a flaw in any measurement, and any methodology that does not address it is marking its own homework. We measured it on our own corpus: of 535 engine-prompt pairs asked at least twice over 60 days, only 16.3% produced the identical set of search queries every run (published August 2026, with the full sample). A single point-in-time check of any AI answer is one roll of the dice.

The Index is designed around that fact rather than around pretending it away. Rankings never rest on a single roll: the consensus gate requires at least two independent AI models before a brand ranks at all, so one run’s quirk cannot create a position. Movement is read edition over edition against immutable snapshots, never inside a single run. The exact run count behind every edition is published in its JSON record, the exact model version is recorded on every answer, and model changes are dated in the changelog below with their expected variance noted, so month-over-month movement is always attributable to either the market or the method, never silently to both.

Sampling cadence is deliberate and disclosed: the public Index refreshes monthly in a single verified sweep, and CiteHawk workspaces collect weekly. More runs per refresh would smooth variance further, and the complete verbatim corpus is published precisely so anyone can quantify the remaining variance themselves rather than taking our word for it. That is the trade we choose: fewer, fully published, tamper-evident runs over many unpublishable ones.

Regional rankings

Some categories (banks, insurers, agencies, professional services) get genuinely different answers in different countries, so those categories are also asked as a local buyer would ask, for the United States, the United Kingdom, Australia and Canada. A regional edition is published only when its ranking meaningfully differs from the global one: a different #1, or fewer than three-quarters of the top 10 in common.

The pledge

Rankings are computed from AI responses only. Claiming a brand cannot change its position, and positions are not for sale.

Brand owners can claim their brand to verify identity details and follow their movement. Identity corrections are reviewed and never affect scores. No payment, partnership, or relationship with CiteHawk influences any ranking.

The changelog

Dated version history of the methodology.

2026-09-02mortgage-brokers retired; AI consulting firms added. When asked for mortgage brokers, AI models answer almost entirely with lenders and banks, which already rank on the mortgage-lenders board, so the two boards had become near-duplicates. The mortgage-brokers page now redirects to mortgage-lenders; its frozen editions remain in the archive and stay reachable by their published records. Its slot in the September 2026 edition goes to AI consulting firms, a category buyers are asking AI models about; the new board is captured on this date for the global edition and joins the regional refresh on September 3, with no previous edition and therefore no movement shown. The category count is unchanged.
2026-09-02Adjacent-category members on two service boards: solar-companies ranks installers and energy providers, so panel manufacturers named beside them are excluded and stay eligible on their own product boards; real-estate-companies excludes listings portals, which are directories. Two consumer review sites (Canstar Blue, CHOICE) join the directory table, so the existing directory rule delists them on every service board. Four other boards flagged by the same review were confirmed as they read: cyber-security-companies, financial-advisors, debt-consolidation-companies and hvac-companies keep every entity, because the companies AI models name there are members by any ordinary reading. The lists are published below. Global boards were re-assembled from their archived answers on this date; regional boards apply it from the September 3 regional refresh; frozen editions unchanged.
2026-09-02Product domains pinned on service-firm boards: when an AI model names a product by its parent company’s web address (QuickBooks or TurboTax under intuit.com, NetSuite under oracle.com, Meta Ads under facebook.com), the curated alias list now records the domain the product itself owns, but only on the boards where that product is an intruder (accounting-firms, the marketing-agency boards). This lets the software-product exclusion apply as intended; QuickBooks had stayed on the global accounting-firms board because the name-owns-domain guard read the parent address as a misattribution. turbotax.com joins the accounting-firms exclusion list. Product boards are unchanged. Global boards were re-assembled from their archived answers on this date; regional boards apply it from the September 3 regional refresh; frozen editions unchanged.
2026-09-02Software products excluded from service-firm boards: the rule already applied to seo-companies now covers accounting-firms (Xero, QuickBooks, FreshBooks and similar), web-design-companies (Webflow, Squarespace, Shopify, Wix and similar), property-management-companies (AppFolio, Buildium, Yardi and similar) and every marketing-agency board (Hootsuite, Mailchimp, Buffer, Canva and similar). AI models name these as the tools an agency or firm would use, and the extractor was reading the tool as the recommendation. Each product remains eligible on its own software board, and real firms that also sell software stay ranked. The exclusion lists are published below. Global boards were re-assembled from their archived answers on this date; regional boards apply it from the September 3 regional refresh; frozen editions unchanged.
2026-09-02Empty answers now contribute no entities. When an AI model returns an empty answer (no AI Overview shown, or a provider outage), the entity extractor was still asked to read it and could invent names. 911 archived answers were corrected to name nothing, and the live boards were recomputed: seven published rows on six boards that rested only on invented entities were removed (for example a phone model on water filters and two AI companies on UK credit unions), nine rows had their mention counts corrected, and no board changed its number one. Empty answers still count as answers, so mention-rate denominators are unchanged. Verbatim answer text and published content hashes are untouched.
2026-09-02Exclusion guard correction: the name-owns-domain check that protects a real provider from a mis-attributed domain now compares hyphenated domain labels the same way it compares names. Before this, a listed brand whose domain carries a hyphen (Coca-Cola, coca-cola.com) could not be delisted, and it stayed on the US branding-agencies board after the September 1 client-brand exclusions. No list entry changed. A brand whose own domain is hyphenated may now display that domain in place of an unhyphenated variant; malformed domain fragments in an answer never count as owned. Applies from the September 3 regional refresh; frozen editions unchanged.
2026-09-02The Index expands from 193 to 240 categories within the September 2026 edition: 47 new boards across two new families, Developer Infrastructure and AI Infrastructure, plus gap-fills in existing families. Each new board’s first capture is dated 2 September and publishes its own verbatim corpus and content hash; no existing board, ranking, or published hash changed. The September root hash published after the 1 September sweep covers the 193 boards captured that day; the expansion boards are verifiable individually through their per-board hashes and join the root from the October 2026 edition. New boards have no previous edition, so they show no movement in September.
2026-09-02 · v1.3Position credit is re-indexed after exclusions: when an answer names an excluded entity ahead of category members, the surviving brands now close ranks, so a brand named right after an excluded directory earns first-named credit. Previously positions stayed as extracted, a limitation disclosed in the 2026-08-07 entry. Applies from the next refresh of each board; answer corpora, capture timestamps and published hashes are unchanged. In the same release, a failed brand-extraction call is now recorded on the archived answer and that answer leaves the mention-rate pool instead of counting as an answer that named no brands; every affected answer remains verbatim in the published corpus.
Show earlier changes (15)
2026-09-01TikTok joins the branding-agencies exclusion list in a same-day follow-up: once the first client-brand exclusions were applied it surfaced at rank 13 through the same case-study pattern, with AI models citing its brand work as an example and the extractor reading the example as the recommendation. TikTok remains eligible on its own category’s board. Applies from the September 1 follow-up re-assembly; frozen editions unchanged.
2026-09-01branding-agencies deny list extended with famous client brands (Nike, Uber, Netflix and similar). AI models cite them as branding case studies, and the extractor was reading the example as the recommendation, so consumer giants ranked as agencies and suppressed real agencies’ share of voice. Each remains eligible on its own category’s board. Applies from the September re-assembly; frozen editions unchanged.
2026-09-01Entity resolution and curated aliases: one company mentioned under several names or domains is now counted as one entity before scoring. Resolution merges a name with a domain only when the name owns that domain, a curated alias list (published in the repository of record) folds researched variants such as Manpower and ManpowerGroup, and some aliases are scoped to a single category where a name means different companies on different boards. Merging split mentions moves share of voice, so September rankings may partly reflect this consolidation. Frozen editions are unchanged.
2026-08-10Non-determinism, sampling and variance section added: documents run-to-run answer variance (with our published fan-out determinism measurement), the sampling design behind monthly Index sweeps and weekly workspace collections, and how the consensus gate and immutable editions absorb single-run noise. No scoring change.
2026-08-08Sentiment gating was evaluated and not adopted. A per-mention sentiment classification across five categories’ archived August answer corpora found no negatively-framed brand mentions, and rankings were identical with the gate applied: AI models recommending “the best X” do not name brands negatively. Mentions therefore continue to count regardless of sentiment. No scoring change.
2026-08-07Entity-classification exclusions added: rankings now contain only entities of the category. In service-provider categories, directories and marketplaces that list providers are excluded when the ranked entity is the directory itself (name matches the domain), and a published per-category exclusion list removes brands AI models name in passing that are not members of the category. Excluded mentions leave the share-of-voice pool, so remaining scores reflect only category members; position credit stays as extracted, and every excluded mention remains verbatim in the published answer corpus and its extractedBrands. The July and August 2026 editions were recomputed under this rule on this date; corpora, capture timestamps and published hashes are unchanged. The consensus gate is unchanged.
2026-08-05Rankings in the July 2026 edition, including the launch capture, were recomputed on this date under the current scoring (v1.2) and, for product categories, at model granularity, so month-over-month comparisons read like for like. The answer corpora, capture timestamps and published hashes are unchanged.
2026-08-05Tamper-evident hash-freeze shipped: every refresh’s verbatim answer corpus is SHA-256 hashed at capture, the hash is published in the record’s JSON download, and each verified-complete sweep publishes a run-level root hash on this page and in that month’s Index report. Editions captured earlier were hashed retroactively from the archived corpus on this date. No scoring change.
2026-09-01Google AI Mode joins as the ninth engine from the September 2026 edition. AI Mode is Google’s AI-native search surface; answers are captured from the consumer surface with no API model to pin, the same provenance treatment as Google AI Overviews and Copilot. It enters the shared consensus gate and share-of-voice pool like any other engine, so September movement may partly reflect the wider panel. Editions frozen before September 2026 remain eight-engine records and state their own engine counts.
2026-09-01The SEO companies board is redefined as SEO agencies and service providers only, from the September refresh. SEO software products that AI models name when asked for SEO companies (Screaming Frog, Surfer SEO, Clearscope, Moz Pro, Yoast SEO, Rank Math, BrightLocal, Conductor, Botify, SE Ranking) join the published per-category exclusion list for that board; software keeps its own ranking on the SEO tools board. The frozen July and August editions are unchanged. No scoring change.
2026-09-01Grok answers move from Grok 4.5 to Grok 4.3 from the September refresh, with live web search unchanged. Side-by-side runs put the two models within each other’s normal run-to-run variation; September movement in Grok columns may partly reflect the change.
2026-08-01Models upgraded to consumer parity from the August refresh. August movement will partly reflect this upgrade: that is what this changelog is for.
2026-07-23Per-model splits published on every category page; immutable point-in-time records with the complete verbatim answer corpus and JSON downloads. Regional editions launched (US, UK, AU, CA). No scoring change.
2026-07-02 · v1.1Score becomes presence-only: the citation component was removed after the audit described above, and the two-model consensus gate was added.
2026-06-26 · v1.0The Index launched: the leading AI models asked the buyer’s question monthly, every refresh captured as an immutable snapshot.

What the Index is not

The Index reports what AI models recommend. It is not an endorsement by CiteHawk, and it is not a review site. If an AI model is wrong about a category, the Index will faithfully show you that it is wrong. That is the point.

A cream felt receipt held by a brass clip, sealed in indigo wax

Want this measurement for your own brand?

CiteHawk tracks the same signals for your brand across all 9 AI models, every week, receipts included.

Browse the Index

Free AI visibility report · No credit card · 50 prompts · 10 engines