GoGoChimp Research · White paper · October 2026

How citable are SEO agencies to AI? A 2026 benchmark of agency websites

On 6 October 2026 we crawled 484 agency websites and scored 415 of them for how likely AI engines are to quote them. This is what a sharp prospect would find.

415

agency sites scored

70

median Rubric score

52

median Trusted score, the weakest pillar

14 of 646

service pages signed and sourced

If you sell AI search work, a sharp prospect will check your own website first. On 6 October 2026 we crawled 484 agency sites and scored 415 of them, to see what that prospect would find. The access checks passed almost everywhere, but the trust signals didn’t. On Rubric’s Trusted pillar, 168 of the 415 scored under 50. Only 14 of the 646 service pages we crawled had both a named author and two external sources. So did just 14 of the 570 case studies.

Ten findings

  1. Trusted is the weakest pillar. Its median was 52, against 80 for Known and 71 for Findable, and 168 of 415 agencies scored under 50 on it.
  2. Homepages score about 12 points above the median inner page. The homepage won at 373 of 415 sites (89.9%), with a median of 82 against 70.
  3. Hardly any service page or case study is signed and sourced. Only 14 of 646 service pages (2.2%) and 14 of 570 case studies (2.5%) had both a named author and two external source links.
  4. Named author is the most common failure, at 249 of 413 sites. External sources (224 of 415) and content-type schema (208) come next.
  5. Headings and schema separate the best agencies from the rest. Question or claim-shaped headings failed at 85 of the 104 bottom-quarter agencies and 18 of 102 in the top quarter. Content-type schema failed at 76 against 22.
  6. Few agencies mark up what they sell. Only 130 of 646 service pages (20.1%) declared Service schema.
  7. Deliberate AI-bot blocking is rare. Two of 460 probed agencies paired a robots.txt block with a live 403, both for ClaudeBot. About 1 agency site in 20 (25 of 484) serves a firewall challenge to a standard crawler.
  8. Fewer than half the agency crawls found a comparison page (46.3%), and 10.4% found a machine-readable price.
  9. Agencies beat both comparison groups crawled on the same engine. Their median of 70 compares with 67 for 47 SEO software vendors and 66 for 244 other websites crawled a day later.
  10. No agency reached 90. The middle half scored between 65 and 75.

Data: GoGoChimp’s own crawl of 415 agency sites and 47 SEO-tool vendors on 6 October 2026, plus 244 other websites on 7 October 2026, all scored with Rubric, GoGoChimp’s AI citability auditor. By Chris McCarron, founder of GoGoChimp.

Contents

What’s in this paper

  1. Your clients will ask whether AI can quote them, and your site is the first proof
  2. How did we sample and score agency sites?
  3. Trusted is the weakest pillar, with a median of 52
  4. Service pages and case studies are the weakest pages agencies sell from
  5. A homepage-only check overstates AI citability by about 12 points
  6. What separates the top quarter of agencies from the bottom quarter?
  7. The five most common failures are all markup and attribution fixes
  8. Few agencies mark up the services they sell
  9. Entity links are thin, and some schema only appears after JavaScript runs
  10. Fewer than half the agency crawls found a comparison page
  11. Only 2 of 460 probed agencies blocked an AI crawler on purpose
  12. About one agency site in 20 serves a bot challenge to a standard crawler
  13. Rubric’s engine model rates Gemini the weakest fit at 152 of 415 agency sites
  14. SEO software vendors scored below the agencies in the same crawl
  15. Agencies score four points above the wider web
  16. Country differences are modest, and most samples are small
  17. Agencies cluster in a narrow band, and none reach 90
  18. How did the North of England’s top 50 agencies score?
  19. What should agencies fix first?
  20. What can’t this benchmark tell you?
  21. Appendix A: results by list
  22. Appendix B: agencies by country
  23. Appendix C: page types
  24. Appendix D: schema types declared
  25. Appendix E: top quarter against bottom quarter
  26. Sources

How to read this paper. Agency figures come from GoGoChimp’s Rubric crawl of 6 October 2026. Any other source is named where it’s used. Rubric scores pages from 0 to 100, and 70 is the line at which a page counts as quotable. Rubric predicts citability, it doesn’t count citations. Top performers are named, positively. No other agency is named.

Your clients will ask whether AI can quote them, and your site is the first proof

Your clients are going to ask whether ChatGPT, Claude, Perplexity or Google’s AI answers can quote them. Before they believe your answer, they can run the same test on your own site. It’s the cheapest due diligence a buyer can do.

Agency sites ought to be the easiest test case on the web. They’re written by people who sell search for a living.

The plumbing’s fine. What’s missing is the layer that tells an AI engine what a page is and who stands behind it.

That gap costs you twice: with no author markup, you’ve got a weaker site and a weaker pitch. The client can see you haven’t done for yourself what you’re selling.

How did we sample and score agency sites?

We crawled 533 sites on one engine, at one page cap, on one day. That’s 484 agencies and 49 SEO-tool vendors, all on Rubric scoring version 2026-10-01 at up to 40 pages each. The crawl ran on 6 October 2026 between 14:35 and 20:12. We’ve kept the tool vendors separate. Sites we couldn’t crawl properly are counted, but they’re never scored.

Which agency sites did we crawl?

We crawled every site on the five lists behind our earlier crawls, apart from three Prolific North domains that now redirect to another brand.

ListWho’s on itAgencies attemptedAgencies scored
12-13 Aug seven-country listSEO agencies in the USA, UK, Canada, Australia, Germany, France and Spain, plus the 49 SEO-tool vendors243200
23 Aug English-speaking listAgencies in the US, UK, Canada, Australia, New Zealand, Ireland and South Africa10089
23 Aug international listInternational and mid-tier agencies (Netherlands, Ireland, India, the Nordics, Italy, Poland, Portugal, Singapore, UAE, South Africa, Brazil and a few US, UK and Australian sites)9989
Prolific North Top 50 Digital Agencies 2025Digital agencies in the North of England4742
10 UK SEO agenciesA cross-check list1010

Fourteen agencies sit on two or three lists, so the rows add up to more than 484 and 415. We ran every crawl ourselves rather than through Rubric’s hosted service, with the broken-link check and the Common Crawl lookup switched off.

What does Rubric measure?

Rubric crawls a site up to a page cap and runs every page through the same set of checks. They roll up into three pillars and a score out of 100.

  • Known: can an engine tell what a page is and who’s behind it? Content-type schema and entity markup sit here.
  • Findable: can an engine reach the page and lift a clean answer from it? Answer-first openers and question-shaped headings sit here.
  • Trusted: does the page give an engine reasons to believe it? Named authors, external sources, sourced statistics and freshness sit here.

Rubric predicts citability, it doesn’t count citations. A high score means a page is easy to quote. The score says nothing about an agency’s client work. The per-engine fits in Appendix A are modelled weightings of the same checks.

Which crawls get a score?

Every crawl that clears six failure rules gets a score. We checked each site against them in this order, and the first rule that applied set its status.

  1. Failed: the overall score was 0.
  2. Unreachable: the homepage was a browser error page, such as a DNS, certificate or timeout error.
  3. Block flag: the engine flagged the crawl as access-blocked.
  4. Challenge: more than half the pages were firewall challenges.
  5. Degraded: more than half the pages were errors or non-200 responses.
  6. One page: the crawler reached only the homepage.

In all, 415 of 484 agencies (85.7%) and 47 of 49 tool vendors were scored.

Why an agency wasn’t scoredSites
Firewall challenge pages25
Unreachable (DNS 5, certificate 3, timeout 2, browser error 2)12
Homepage only10
Engine block flag9
Degraded8
Failed5
Total69

Did we retry the sites that failed?

Yes, 24 of them. We gave 14 a second pass, two sites at a time, mostly because their homepage had timed out or thrown a browser error. We tried 20 on their www address after the bare domain failed, and some sites got both. Six of the 24 are scored now, two of them from the www address.

How did we sort the pages by type?

We sorted pages by URL first, so /services/ marks a service page and /case-studies/ or /work/ a case study. Otherwise we used the crawler’s own label, and anything left went into “other page”. The crawler read 15,658 pages at the 415 scored agencies: 26.5% blog posts or guides, 4.1% service pages, 3.6% case studies and 45.4% other pages.

The engine also labels the root of any subdomain it reaches as a homepage, and some of those roots were error pages. So the page-type tables count 537 homepages at 415 sites. Every site-level homepage comparison in this paper uses only the homepage on the address we crawled.

Does the page mix change a site’s score?

Yes, by 5 points at the median. Sites where blog posts made up under a quarter of the crawl had a median of 68 (n=265). Sites where they made up half or more had 73 (n=103). Blog posts score above most inner pages, so a crawl that reaches mostly posts lifts the site’s score. Every site had the same 40-page cap, but the crawler chose which pages it read.

How stable are the scores?

Scores barely move on the same engine. We’d crawled the 42 Prolific North sites on this engine on 5-6 October. A day later, the median absolute move was 0 and no site moved 10 points or more. The list’s median (65.5) and Trusted median (48) came out identical. The 10 UK SEO agencies, crawled at 25 pages on 5-6 October, moved by a median of 1.5.

What changed since our August crawls?

Trusted scores fell under the current engine, which is stricter on named authors and on content-type schema. Our August crawls ran on an older, unversioned engine, and the August figures below are the only ones in this paper. Don’t set them beside anything else here.

We scored 203 sites from the 12-13 August list both times, tool vendors included. Their Trusted median fell from 59 in August to 54 on 6 October. Of those 203, 97 moved 10 points or more on Trusted, while the overall median moved by 1 point.

The August block flag also over-fired. Of the sites it capped as blocked in August, 44 were scored normally on 6 October.

Trusted is the weakest pillar, with a median of 52

Across the 415 agencies, the Trusted median was 52, against 80 for Known and 71 for Findable. Only 61 agencies (14.7%) reached 70 on Trusted, and 168 scored under 50.

Median pillar score, 415 agencies

Known

80

Findable

71

Trusted

52

Trusted is the only pillar with a median below 70, the quotable line. 168 of the 415 agencies scored under 50 on it. Source: GoGoChimp Rubric crawl of 6 October 2026, scoring version 2026-10-01, up to 40 pages per site.

PillarMedianSites at 70+Sites under 50
Known80339 (81.7%)3
Findable71251 (60.5%)1
Trusted5261 (14.7%)168

Trusted came last on every list in the crawl, from 46.5 for the 10 UK SEO agencies to 54 for the English-speaking list. Appendix A has each list.

What does the Trusted pillar check?

The Trusted pillar checks for evidence: a named author, at least two external source links and statistics that carry a source. It also wants at least 1.5 statistics per 100 words and a date inside the last 12 months. I’d put every one of those in a client audit. On agency sites, they’re the signals that go missing.

Why do agencies score lowest on trust?

My read is that agency sites are built as brochures. Most service pages carry nobody’s name and link to nothing outside the site. That’s fine for a buyer who’s already booked a call. But it gives an AI engine nobody to credit and nothing to check the claim against.

Trusted was the weakest of Rubric’s three pillars across the 415 agency sites GoGoChimp crawled on 6 October 2026, with a median of 52. Of the 415, 168 scored under 50 on it.

Where does trust go missing?

Trust goes missing on the pages away from the homepage and the blog. The median homepage scored 67 on Trusted and the median blog post 66. The median service page scored 46, the median case study 45 and the median about, team or contact page 48. They’re page-level medians, which is why they differ from the site-level 52.

How old are agency blog posts?

Just under half of agency blog posts are more than a year old. Of 4,151 blog posts, 3,944 (95.0%) carried a date, and the median dated post was 357.5 days old. Of the dated posts, 1,943 (49.3%) were more than a year old and 1,339 (34.0%) more than two.

Rubric’s freshness check failed at only 89 of 367 sites (24.3%), so the dates tell you more than the check does. Because the crawl stops at 40 pages, a site’s newest post isn’t always among them.

Service pages and case studies are the weakest pages agencies sell from

The median service page scored 70 and the median case study 68, against 75 for a blog post and 80 for a homepage. About, team and contact pages scored lower still, at 66. The crawl reached service pages at 175 of 415 sites and case studies at 157.

Median page score by page type

Homepage

80

Blog post or guide

75

Service page

70

Case study

68

Other page

68

About, team, contact

66

Service pages and case studies, the pages that sell, sit at or below the line. Source: GoGoChimp Rubric crawl of 6 October 2026, scoring version 2026-10-01, up to 40 pages per site.

Page typePagesSitesMedian page scoreKnownFindableTrustedMedian wordsMedian external links
Homepage537415808680678951
Blog post or guide4,151292759071669951
Service page646175707774468840
Case study570157687771455130
About, team, contact1,131347667769484291
Other page7,101388687571487980

The homepage row includes subdomain roots, which is why it’s 80 here and 82 in the homepage section. Appendix C has every page type.

Which checks do service pages and case studies fail?

Named author fails most, on about eight pages in ten. Each cell is the share of applicable pages that came back bad.

CheckBlog post or guideService pageCase study
Named author9%82%78%
Content-type schemanot reported61%60%
Two or more external source links44%60%61%
Answer-first opener52%52%46%
Question or claim-shaped headings43%50%68%
Statistics carry a source28%27%49%

We haven’t reported the blog-post schema figure. The crawler labels pages with Article schema as articles, so it would be partly circular.

How many service pages and case studies are signed and sourced?

Hardly any: 14 of 646 service pages (2.2%) and 14 of 570 case studies (2.5%). We call a page signed and sourced when it passes both the named-author check and the two-or-more external links check.

The byline is the bigger gap. Only 28 of 600 applicable service pages (4.7%) passed the named-author check, and 379 of 646 (58.7%) had no external link at all.

What does the typical agency case study look like?

Most agency case studies are unsigned and unsourced, and one in four runs under 300 words. Of 570 case studies, 347 (60.9%) had no external link and 140 (24.6%) were that short. Only 77 of 554 applicable case studies (13.9%) passed the named-author check, and 83 (14.6%) declared an Article-family schema type.

Why does it matter which pages score low?

A buyer’s specific questions land on these pages. “Can this agency handle technical SEO for a Shopify store?” is a service-page question. “Have they done it for anyone like us?” is a case-study question. Most of those pages are unsigned, and most link to nothing a buyer could check.

In GoGoChimp’s 6 October 2026 crawl of 415 agency sites, 14 of 646 service pages and 14 of 570 case studies had both a named author and two external source links.

A homepage-only check overstates AI citability by about 12 points

The homepage outscored the median inner page at 373 of 415 sites (89.9%). The median homepage scored 82 and the median inner page 70. Site by site, the median gap was 12 points.

Median gap below the homepage, in points

About, team, contact

16

Service page

13

Case study

13

Blog post or guide

7

Median points each page type scored below the same site’s homepage. Source: GoGoChimp Rubric crawl of 6 October 2026, scoring version 2026-10-01, up to 40 pages per site.

The homepage is the agency in its interview suit. The inner pages are the same agency at home on a Sunday, in a dressing gown, eating cereal out of the box.

The service page a buyer lands on is the one in the dressing gown.

Every kind of inner page scores below the homepage at most agencies. For service pages and case studies, it’s nine sites in ten.

Page typeScored below the homepage atMedian gap (points)
About, team, contact322 of 347 sites (92.8%)16
Service page158 of 175 sites (90.3%)13
Case study143 of 157 sites (91.1%)13
Blog post or guide223 of 292 sites (76.4%)7

Why does a short crawl score high?

The fewer pages a crawl reaches, the more the homepage carries the score. One agency’s crawl reached 2 pages and scored 88, which would top our table if we ranked it. We don’t rank any crawl under 10 pages, and we never score a one-page crawl. Thirteen scored crawls reached fewer than 10 pages, and dropping them leaves the median at 70.

What should you test instead?

Test a service page and a recent post, and do the same when you audit a prospect. A homepage score on its own flatters the site.

Across 415 agency sites crawled by GoGoChimp on 6 October 2026, the median homepage scored 82 on Rubric and the median inner page 70.

What separates the top quarter of agencies from the bottom quarter?

The gap comes from content and attribution checks. Question or claim-shaped headings failed at 85 of the 104 lowest-scoring agencies, against 18 of 102 in the top quarter. Entity clarity, content-type schema, external sources, answer-first openers and named author split them the same way. The HTTP, robots.txt and canonical checks didn’t separate them at all, because no agency in the crawl failed them.

Share of sites failing the check, top quarter against bottom quarter

 Top quarter (n=104) Bottom quarter (n=104)
0%25%50%75%100%

Question or claim-shaped headings

18 of 10285 of 104

Content-type schema

2276

Two or more external sources

2677

Answer-first opener

29 of 10272

Named author

35 of 10274

Entity clarity

657

Statistic density

332

Counts are sites failing out of 104 unless shown. The gap sits on attribution and markup, not on access. Source: GoGoChimp Rubric crawl of 6 October 2026, scoring version 2026-10-01, up to 40 pages per site.

The top quarter is the 104 agencies scoring 75 to 88, and the bottom quarter the 104 scoring 46 to 65. Their medians were 78 and 62 overall, and 67.5 and 43 on Trusted.

CheckTop quarter failingBottom quarter failing
Question or claim-shaped headings18 of 10285 of 104
Entity clarity (Organization or Person plus sameAs)6 of 10457 of 104
Content-type schema22 of 10476 of 104
Two or more external source links26 of 10477 of 104
Answer-first opener29 of 10272 of 104
Named author35 of 10274 of 104
Statistic density (1.5 or more per 100 words)3 of 10432 of 104

Counts out of 102 are for checks that didn’t apply at every site. Appendix E has five more checks. Every one of these checks feeds the score that sorts agencies into quarters, so part of each gap is built in. None of it proves what an engine rewards.

Is the gap on the homepage or the inner pages?

The gap is mostly on the inner pages. Top-quarter homepages had a median of 89 and bottom-quarter homepages 76, 13 points apart. Their median inner pages were 78 and 61, 17 points apart.

What do the best agencies do on their inner pages?

The best agencies sign and source most of their inner pages. At conquerradigital.com.au (85, Trusted 92), 34 of 38 applicable inner pages had a named author and 36 of 38 had two or more external source links. All but one had sourced statistics. At firestarterseo.com, the top scorer at 87, 33 of 38 had a named author and 33 of 38 external sources. All 38 carried content-type schema.

In GoGoChimp’s 6 October 2026 crawl, 85 of the 104 lowest-scoring agencies failed Rubric’s question-heading check. In the top quarter, 18 of 102 did.

The five most common failures are all markup and attribution fixes

Named author failed at the most sites, 249 of 413 (60.3%). External source links (224), content-type schema (208), question or claim-shaped headings (205) and answer-first openers (200) came next. No agency failed the HTTP, robots.txt, canonical, internal-link, reachability or word-count checks.

Share of agencies failing the check on most pages

Named author

60.3%

Two or more external sources

54.0%

Content-type schema

50.1%

Question or claim-shaped headings

49.6%

Answer-first opener

48.4%

Named author: 249 of 413. External sources: 224 of 415. Schema: 208 of 415. Headings: 205 of 413. Answer-first: 200 of 413. Source: GoGoChimp Rubric crawl of 6 October 2026, scoring version 2026-10-01, up to 40 pages per site.

A site fails a check here when it comes back bad on more than half the pages it applies to.

CheckPillarSites failingMedian share of pages failing
Named authorTrusted249 of 413 (60.3%)70%
Two or more external source linksTrusted224 of 415 (54.0%)55%
Content-type schema (Article, Service and so on)Known208 of 415 (50.1%)51%
Question or claim-shaped H2 and H3 headingsFindable205 of 413 (49.6%)50%
Answer-first opener (40-70 words)Findable200 of 413 (48.4%)50%

What does missing content-type schema cost?

Without content-type schema, an engine has to work out from the copy what kind of page it’s reading. On the median agency site, 51% of applicable pages failed the check.

What’s wrong with agency headings and openers?

A heading like “Our approach” doesn’t tell an engine what question the section answers. An opener that spends its first 70 words on brand copy gives it nothing to lift. Both are writing problems, and they’re quick to fix.

Why do named authors and external sources matter?

A named author gives an engine a person to credit with the expertise, and external links give it something to cross-check. Without either, the page is an unsigned claim. On the median agency site, 70% of applicable pages failed the named-author check.

What does a signed, sourced post score?

Signed, sourced posts score about 14 points more than posts with neither. Of 4,151 agency blog posts, 1,691 (40.7%) had a named author and two or more external source links, and their median page score was 81. The 195 posts (4.7%) with neither scored a median of 67.

Part of that gap is mechanical. Both checks feed the page score, so a page that passes them scores higher by construction. The gap shows what the two fixes are worth inside Rubric’s scoring. This crawl can’t tell you whether an engine rewards them.

In GoGoChimp’s 6 October 2026 crawl, 249 of 413 agency sites failed Rubric’s named-author check on more than half their pages. None failed the HTTP, robots.txt or canonical checks.

Few agencies mark up the services they sell

Service schema is rare, even on service pages. Only 139 of 415 agencies (33.5%) used Service as its own type on any crawled page, and 130 of 646 service pages (20.1%) declared it. More agencies carried FAQPage (194) than Service.

The table counts sites with each type on at least one crawled page, and Appendix D has every type.

Schema typeAgencies (n=415)
Organization (or a subtype)366 (88.2%)
WebSite333 (80.2%)
BreadcrumbList310 (74.7%)
Person285 (68.7%)
Article family261 (62.9%)
FAQPage194 (46.7%)
Service as its own type139 (33.5%)
Offer114 (27.5%)
Review or AggregateRating106 (25.5%)

Which schema do agency sites carry?

Agency sites mostly carry generic schema. Organization and WebSite top the table, and those two tell an engine whose site it is. Neither says what the agency sells.

Why does agency schema look so alike?

My guess is it’s the CMS. Organization, WebSite and BreadcrumbList are the three commonest types. In my experience, that’s the graph an SEO plugin emits out of the box. Rubric didn’t detect any plugin, and I haven’t checked site by site.

What else is missing from agency schema?

Review markup is missing at most agencies, and a few carry no schema at all. Only 106 agencies (25.5%) declared review or rating markup, and 21 (5.1%) had no schema on any crawled page. HowTo turned up at 15.

In GoGoChimp’s 6 October 2026 crawl, 130 of 646 agency service pages declared Service schema. More of the 415 agencies carried FAQPage (194) than Service (139).

Close to a quarter of agencies (92 of 415, 22.2%) declared no sameAs links, and only 6 linked to Wikidata. The median agency declared 4. JavaScript-only schema turned up at 44 sites (10.6%), and on more than half the crawled pages at 23.

SignalAgencies (n=415)
No sameAs links92 (22.2%)
Median sameAs count4
Wikidata link6
Entity clarity check failing106 (25.5%)
JS-only schema on 1+ page44 (10.6%)
JS-only schema on more than half the pages23
llms.txt present185 (44.6%)

Rubric doesn’t score llms.txt.

The sameAs property ties a brand’s schema to its profiles elsewhere: LinkedIn, Companies House, review sites, Wikidata. It’s how an engine confirms the agency on the page is the entity it’s read about somewhere else.

LinkedIn, then X, then YouTube, and very little else.

ProfileAgencies listing it in sameAs
LinkedIn279 (67.2%)
X207 (49.9%)
YouTube149 (35.9%)
Crunchbase18 (4.3%)
GitHub7 (1.7%)
Wikipedia5 (1.2%)
G25 (1.2%)
Reddit4 (1.0%)
Trustpilot4 (1.0%)
Capterra1 (0.2%)

Rubric only counts a fixed list of platforms. Clutch, Google Business Profile and Companies House aren’t on it, so these figures say nothing about them.

Why does JavaScript-only schema matter?

A crawler that doesn’t run JavaScript never sees schema that exists only in the rendered page. In Vercel’s 2024 analysis, “none of the major AI crawlers currently render JavaScript”, with OpenAI’s three bots, ClaudeBot and PerplexityBot among them (Vercel and MERJ, 2024). It’s a December 2024 study, and serving schema in the HTML costs nothing either way.

Rubric catches JS-only schema by comparing each page’s raw HTML with a rendered copy.

In GoGoChimp’s 6 October 2026 crawl, 92 of 415 agency sites declared no sameAs links in their schema, and only 6 linked to Wikidata.

Fewer than half the agency crawls found a comparison page

Rubric found a comparison page at 192 of 415 agency sites (46.3%) and a how-to-choose page at 124 (29.9%). Only 43 (10.4%) had a machine-readable price on any crawled page.

What counts as a follow-up page?

Rubric sorts a buyer’s follow-up questions into five kinds and looks for a page that answers each, by URL, title and question headings.

Kind of questionWhat Rubric looks forAgencies with a page (of 415)
Comparevs, alternatives, best X for Y192 (46.3%)
Constrainhow to choose, buyer’s guide, checklist, pricing guide, use cases124 (29.9%)
Clarifywhat is, guide to, overview, FAQ, glossary293 (70.6%)
Validatecase study, results, testimonials, reviews308 (74.2%)
Actpricing, get a quote, book a call, contact328 (79.0%)

These counts are lower bounds. A page the crawler didn’t reach in its 40 doesn’t count, so some sites will have pages we’ve missed.

What does a buyer ask an AI before shortlisting an agency?

Buyers ask specific, comparative questions that a homepage can’t answer. These four are illustrative and don’t come from our data.

  • “Which Manchester agencies are good at technical SEO for ecommerce?” (Compare)
  • “How do I choose an SEO agency, and what should I ask on the first call?” (Constrain)
  • “What does a technical SEO retainer cost in the UK?” (Constrain, then Act)
  • “Has Agency A worked with a company like mine?” (Validate)

Agencies have plenty of pages for the last kind and far fewer for the first two.

Why does a missing price matter?

A buyer will ask what it costs, and an engine can only quote a number someone’s published. You don’t have to publish your rates. A starting price or a range on one page gives an engine something of yours to quote.

GoGoChimp’s 6 October 2026 crawl found a comparison page at 192 of 415 agency sites, and a machine-readable price at 43.

Only 2 of 460 probed agencies blocked an AI crawler on purpose

Two agencies paired a robots.txt block on an AI bot with a live 403 to that same bot, and both were blocking ClaudeBot. That pairing is what we count as a deliberate block. We probed 460 agencies, and we’ve counted the two rather than named them.

How many agencies block AI bots in robots.txt?

Thirteen do. Twelve of them block training crawlers only, and one also blocks a search or user-triggered bot. All 13 block CCBot. GPTBot is blocked at 8, Google-Extended at 6, ClaudeBot at 5 and PerplexityBot at 1.

Why isn’t a 403 proof of a block?

A firewall can turn away any bot it can’t verify, whatever the site’s AI policy. A user-agent probe sends a request labelled as GPTBot or ClaudeBot and records the reply. Ours came from our own IP address, which isn’t on any vendor’s network. Of the 460 agencies we probed, 70 (15.2%) returned a 403 to at least one AI agent.

Of those 70, 26 refused every bot we sent, Googlebot and Bingbot included. Our Googlebot probe got a 403 at 35 agencies, more often than our OAI-SearchBot probe (30). That’s what a firewall turning away unverified bots looks like.

If a client’s tool says their site blocks AI, check the firewall before the robots.txt. If your own audit says it, re-probe before it goes in a deck. That includes ours.

Of 460 agency sites GoGoChimp probed on 6 October 2026, 70 returned a 403 to at least one AI user agent. Only 2 paired it with a robots.txt block on the same bot.

Which AI crawlers should an agency allow?

Allow all the search crawlers and answer-time fetchers, and make training a separate decision. The vendors’ own documentation splits their bots that way.

KindBotWhat the vendor says
SearchOAI-SearchBot“OAI-SearchBot is for search” (OpenAI)
SearchClaude-SearchBotBlocking it “prevents our system from indexing your content for search optimization” (Anthropic)
SearchPerplexityBot“designed to surface and link websites in search results on Perplexity” (Perplexity)
SearchGooglebotIts rules “affect Google Search (including Discover and all Google Search features)” (Google)
Answer-time fetchChatGPT-User“Because these actions are initiated by a user, robots.txt rules may not apply” (OpenAI)
Answer-time fetchClaude-UserUsed “When individuals ask questions to Claude” (Anthropic)
Answer-time fetchPerplexity-User“this fetcher generally ignores robots.txt rules” (Perplexity)
TrainingGPTBotDisallowing it “indicates a site’s content should not be used in training generative AI foundation models” (OpenAI)
TrainingClaudeBotRestricting it “signals that the site’s future materials should be excluded from our AI model training datasets” (Anthropic)
Control tokenGoogle-ExtendedCovers Gemini training and grounding, and “does not impact a site’s inclusion in Google Search” (Google)

Google-Extended isn’t a crawler of its own. Google calls it a “standalone product token” that’s used “in a control capacity”, and Google’s usual crawlers do the fetching (Google).

Most of the refusals we saw weren’t aimed at AI search. Besides the 26 agencies that refused every bot, 33 of the 70 refused only ClaudeBot, which is Anthropic’s training crawler. That’s 59 of the 70.

On an agency site, I’d let the four search crawlers and three fetchers through both robots.txt and the firewall. The fetchers may ignore robots.txt anyway, by OpenAI’s and Perplexity’s own account, so the firewall decides what they get.

About one agency site in 20 serves a bot challenge to a standard crawler

A firewall challenge page met our crawler on most URLs at 25 of the 484 agency sites we attempted (5.2%). We didn’t score those sites.

A challenge page has nothing on it to quote, whoever it’s meant to stop.

We can’t see from outside whether verified AI crawlers get the same page. Put that question to your host or firewall provider: what do GPTBot, ClaudeBot and PerplexityBot receive on their first request? If it’s a challenge page, that’s what they read.

Firewall defaults have shifted under agencies, too. In July 2025, Cloudflare said it was “changing the default to block AI crawlers unless they pay creators for their content” (Cloudflare, 2025). In July 2026 it said “non-verified bots are still default blocked” (Cloudflare, July 2026). On 15 September 2026, it said its Block settings “now apply to mixed-use crawlers, including Applebot, Bingbot, and Googlebot”, so they affect search as well as training (Cloudflare, September 2026). We can’t tell which provider or setting any of our challenged sites used.

On 6 October 2026, 25 of the 484 agency sites GoGoChimp tried to crawl served a firewall challenge page on most URLs, about 1 in 20.

Rubric’s engine model rates Gemini the weakest fit at 152 of 415 agency sites

Rubric’s model rated Gemini the weakest engine fit at 152 of 415 agency sites, and AI Overviews at 111. ChatGPT was weakest at 63, Perplexity at 51, Claude at 20 and Copilot at 18. These fits are modelled, and we didn’t query any engine.

In Rubric’s current engine, Gemini’s heaviest checks are content-type schema, entity clarity and indexability. For AI Overviews, they’re answer-first openers, question headings and indexability. Three of those are among the five checks agencies failed most often, so the two fits that lean on them sit lowest. Appendix A has the median fit for each engine.

I wouldn’t chase one engine’s fit. The checks that pull Gemini and AI Overviews down are the same ones the fixes below deal with.

SEO software vendors scored below the agencies in the same crawl

SEO-tool vendors scored a median of 67 (n=47) and agencies 70 (n=415), on the same engine, page cap and day. The gap sits in Known, at 73 against 80. Findable went a point the other way (72 against 71), and Trusted was close (51 against 52).

SEO-tool vendors (n=47)Agencies (n=415)
Median6770
Mean67.569.7
Known7380
Findable7271
Trusted5152
Scoring 70+13 (27.7%)222 (53.5%)

I didn’t expect that.

I suspect product sites lean on app pages, docs and pricing tables. Those carry less of the content-type and entity markup Rubric checks under Known, though I haven’t tested it. The best tool sites still score well: keyword.com at 86, diib.com at 82 and yoast.com at 81.

In GoGoChimp’s 6 October 2026 crawl, SEO software vendors scored a median of 67 on Rubric (n=47). SEO agencies in the same crawl scored 70 (n=415).

Agencies score four points above the wider web

Agencies scored a median of 70, against 66 for 244 other websites we crawled on 7 October, a day later, with the same engine and the same 40-page cap. Those sites are Rubric’s benchmark corpus: online shops, blogs and publishers, SaaS companies and other businesses. Because both crawls ran under the same rules, the scores compare directly.

Median overall and Trusted score

 Agencies (n=415) Other websites (n=244)

Median score

70
66

Known

80
71

Findable

71
68

Trusted

52
49

Other websites: Rubric’s benchmark corpus, re-crawled on 7 October 2026 on the same engine and 40-page cap. Source: GoGoChimp Rubric crawl of 6 October 2026, scoring version 2026-10-01, up to 40 pages per site.

Agencies (n=415)Other websites (n=244)
Median7066
Scoring 70+222 (53.5%)55 (22.5%)
Known8071
Findable7168
Trusted5249
Homepage ahead of inner pages373 (89.9%)225 (92.2%)

More than half the agencies cleared the 70 line, against fewer than a quarter of the other sites. The biggest gap is on Known, at nine points. On Trusted, the agency lead shrinks to three.

Which habits do agencies keep better than other sites?

Agencies sign, structure and mark up their pages more often than other sites do. They don’t cite outside sources much more often.

Check failing on most pagesAgenciesOther websites
Named author249 of 413 (60.3%)198 of 242 (81.8%)
Answer-first opener200 of 413 (48.4%)197 of 242 (81.4%)
Question or claim-shaped headings205 of 413 (49.6%)181 of 242 (74.8%)
Content-type schema208 of 415 (50.1%)171 of 243 (70.4%)
Two or more external source links224 of 415 (54.0%)138 of 244 (56.6%)

So agencies practise more of what they preach than the average site, except on sources. Sources are one of the Trusted signals, and missing them helps keep both groups under the line on trust.

In GoGoChimp’s October 2026 crawls, 415 agency websites scored a median of 70 on Rubric, against 66 for 244 other websites crawled a day later on the same engine.

Country differences are modest, and most samples are small

Agency medians ran from 73 in Australia (n=36) to 63.5 in South Africa (n=10), and the UK sat at 69 (n=89). With 10 to 92 scored sites per group, read the table as description and ignore a 1 or 2 point gap between neighbours.

CountrySitesMedianTrusted median
Australia367357
USA437151
Germany and Austria227158
Canada3470.554.5
France1869.550.5
Nordics1069.550
UK896951
Spain176954
South Africa1063.557
Generic domain, no country927151.5

We took the country from the 12-13 August list where it gave one, and otherwise from the domain ending. The table shows groups with 10 or more scored agencies, and the other 44 sit in smaller ones. Appendix B adds Known and Findable.

Agencies cluster in a narrow band, and none reach 90

The middle half of the 415 agencies spans 10 points, from 65 to 75. No agency reached 90, and 28 scored in the 80s. Another 194 sat in the 70s and 161 in the 60s.

Agencies by score band (n=415)

0

90-100

28

80-89

194

70-79

161

60-69

29

50-59

3

Under 50

355 of the 415 agencies scored between 60 and 79. Source: GoGoChimp Rubric crawl of 6 October 2026, scoring version 2026-10-01, up to 40 pages per site.

For a mid-table agency, the top 10 was 12 points away, with tenth place on 82 against a median of 70.

Which agencies scored highest?

We only rank crawls of 10 pages or more. Four agencies tied for tenth, so the table has 13.

RankSiteScore
1firestarterseo.com87
2aztekweb.com86
3conquerradigital.com.au85
4=netzbekannt.de84
4=seokratie.de84
4=soupagency.com.au84
4=thatware.co84
8=kiwop.com83
8=rodanet.com83
10=azurodigital.com82
10=blennd.com82
10=mediaforce.ca82
10=titanblue.com.au82

How high could agencies go?

Agencies would reach a median of 95 on Rubric’s model if every flagged fix were made. That’s a median gain of 24 points, with 380 of 415 projected at 90 or more. The projection is a ceiling, not a forecast, because it assumes every fix lands on every page.

No agency in GoGoChimp’s 6 October 2026 crawl of 415 agency sites reached 90 on Rubric. The middle half scored between 65 and 75.

How did the North of England’s top 50 agencies score?

We have published a fuller write-up of this subset: a regional deep-dive of this benchmark.

The 42 scored northern sites had a median of 65.5, against 70 for all 415 agencies, and a Trusted median of 48, against 52. They’re a subset of the same 6 October crawl, so the two compare directly.

The middle half of the northern agencies scored between 64 and 70, and 12 of the 42 (28.6%) reached 70 or above. Only 2 reached 70 on Trusted. The homepage beat the median inner page at 37 of the 42, by a median of 12.5 points.

The list is the Prolific North Top 50 Digital Agencies 2025, and it ranks digital agencies rather than SEO agencies. Part of the gap to the full benchmark may be the kind of agency, and this data can’t tell you how much.

Which northern agencies scored highest?

Three sites tied for tenth on 70, so the table has 12.

RankSiteScore
1=embryo.com80
1=seoworks.co.uk80
3=havasmarket.co.uk75
3=idhlagency.com75
5visualsoft.co.uk74
6=reasondigital.com73
6=riseatseven.com73
6=unrvld.com73
9velstar.agency71
10=addpeople.co.uk70
10=connective3.com70
10=housedigital.co.uk70

The joint top site, seoworks.co.uk, had a Known score of 91. The two northern sites that reached 70 on Trusted are both in the top 10: riseatseven.com at 76 and embryo.com at 72.

What are the top northern sites doing right?

The top northern sites source or sign most of their inner pages. At riseatseven.com, all 38 applicable inner pages had two or more external source links and sourced statistics. At embryo.com, all 18 applicable inner pages had both. At reasondigital.com, 31 of 37 applicable inner pages carried a named author and 32 of 37 carried content-type schema. At seoworks.co.uk, 23 of 30 had a named author and 24 of 30 had sourced statistics.

GoGoChimp crawled the Prolific North Top 50 Digital Agencies 2025 on 6 October 2026. The 42 scored sites had a median Rubric score of 65.5 and a Trusted median of 48, against 70 and 52 for all 415 agencies.

What should agencies fix first?

Focus on the service pages and case studies, because that’s where trust goes missing. Put a named author on them first, since it’s the most common failure and Trusted is the weakest pillar. Then add sources and content-type schema. All nine fixes are markup or writing jobs. The wider playbook for client sites is in our AI search optimisation hub.

FixRubric checkAgencies (n=415)
1. Name the authorNamed author249 of 413 sites failing
2. Mark up the page typeContent-type schema208 sites failing
3. Cite two or more sourcesExternal source links224 sites failing
4. Ask the question in the headingQuestion or claim-shaped H2-H3s205 of 413 sites failing
5. Answer firstAnswer-first opener200 of 413 sites failing
6. Add sameAs linksEntity clarity106 sites failing
7. Serve schema in the HTMLJS-only schema (raw HTML against rendered)JS-only schema on more than half the pages at 23 sites
8. Sign, date and source case studiesNamed author plus external source links, on case-study pages14 of 570 case studies passed both
9. Publish a comparison page and a how-to-choose pageFollow-up coverage (crawled pages only)Comparison page found at 192 sites, how-to-choose page at 124

Rows 1 to 6 count sites that fail the check on more than half their applicable pages. Rows 7 to 9 count sites or pages, as each cell says.

1. Every article and guide needs a named author

Add a visible byline and Person schema to every post and guide, and to any page that makes an expert claim. Give the Person sameAs links to the author’s profiles.

Give each author a page marked up as a Person, and point each post’s Article schema at it. Bylines go missing most on service pages, so sign each one with the person who leads it. To check, search a post’s source for “author” and match it to the byline.

2. Each page should say in schema what it is

Use Article schema on posts, guides and case studies, and Service schema on service pages.

Set it once in each template so new pages inherit it, and give each Service a provider that points to your Organization. Then view-source on one page per template and read the @type values.

3. Pages that make claims need two external sources

Link to the study you’re quoting, the client’s public result or the standard you’re citing. I understand the instinct to keep visitors on the site. But to an AI engine, a page that links nowhere has nothing to cross-check.

Put the links in the body copy, next to the claim, and start with the service pages, where most linked nowhere. To check, count the outbound links in the main content against Rubric’s floor of two.

4. Headings should ask the buyer’s question

Swap label headings for the question the section answers. “Our process” becomes “How long does a technical SEO audit take?” Our guide to mapping the follow-up questions AI asks about a topic shows how to find the questions worth asking.

Take the questions from your sales calls, in the buyer’s words. Half the service pages we crawled came back bad on question or claim-shaped headings, and case studies did worse, at 68%. To check, read the headings alone and ask whether a buyer could tell what each section answers.

5. The answer belongs in the first 40 to 70 words

Answer each heading in the first paragraph under it, in 40 to 70 words. Brand copy can follow.

Rubric’s check uses that window. Read the first 70 words under each heading and ask whether they’d make sense quoted on their own.

Link to LinkedIn, Companies House, your review profiles and Wikidata if you qualify for an entry.

Put one Organization block on the homepage with an @id, and point every other block at it. To check, search the homepage source for “sameAs”.

7. Schema belongs in the HTML your server sends

If a tag manager or a JavaScript framework injects your schema, move it server-side, into the page template. Test with view-source rather than the browser inspector, which shows the page after JavaScript has run. If you audit client shops, the same rule applies to price and stock. Our guide to why AI won’t recommend a product page covers that side.

From a terminal, curl https://www.example.com/services/ should return the JSON-LD in its output.

8. Every case study needs a byline, a date and a source the client can check

Sign each case study with the person who led the work, and date the page and the period the work covered. Link at least two sources, one of them something the client could confirm, such as their own announcement or their live site. Then add Article schema with the author pattern below.

To check, view-source for Article and author and count the body links. If a client won’t allow a link, name them and the person who signed off the result.

9. Buyers need a comparison page and a how-to-choose page

Write one page comparing your approach with the options a buyer weighs, such as an in-house hire or a freelancer. Write another on how to choose an agency in your speciality, including when you’re the wrong fit.

Ask ChatGPT and Claude the comparison question before and after you publish, and note which pages they cite.

What does the markup look like?

Three short JSON-LD blocks cover it: Person plus Article on a post, Service on a service page, and Organization with sameAs on the homepage. Every name, URL and profile in them is made up, so swap in your own.

A post, with the author as a Person:

{
  "@context": "https://schema.org",
  "@graph": [
    {
      "@type": "Person",
      "@id": "https://www.example.com/team/jamie-example#person",
      "name": "Jamie Example",
      "jobTitle": "Head of Technical SEO",
      "url": "https://www.example.com/team/jamie-example",
      "worksFor": { "@id": "https://www.example.com/#organization" },
      "sameAs": ["https://www.linkedin.com/in/jamie-example"]
    },
    {
      "@type": "Article",
      "headline": "How long does a technical SEO audit take?",
      "datePublished": "2026-10-06",
      "author": { "@id": "https://www.example.com/team/jamie-example#person" },
      "publisher": { "@id": "https://www.example.com/#organization" },
      "mainEntityOfPage": "https://www.example.com/blog/technical-seo-audit-time"
    }
  ]
}

A service page:

{
  "@context": "https://schema.org",
  "@type": "Service",
  "name": "Technical SEO audit",
  "description": "A 10-day audit of crawling, indexing, site speed and structured data, with a ranked fix list.",
  "provider": { "@id": "https://www.example.com/#organization" },
  "url": "https://www.example.com/services/technical-seo-audit",
  "offers": { "@type": "Offer", "price": "3000", "priceCurrency": "GBP" }
}

The homepage, with sameAs:

{
  "@context": "https://schema.org",
  "@type": "Organization",
  "@id": "https://www.example.com/#organization",
  "name": "Example Agency",
  "url": "https://www.example.com/",
  "sameAs": [
    "https://www.linkedin.com/company/example-agency",
    "https://x.com/exampleagency",
    "https://www.youtube.com/@exampleagency",
    "https://find-and-update.company-information.service.gov.uk/company/00000000"
  ]
}

Every block points at the same Organization @id, which lets an engine join them into one entity. Use the same Article pattern on case studies. The Offer is optional, and its £3,000 is as made up as everything else. It only shows where a price goes in the markup, so don’t read it as a market rate.

What does an answer-first service page look like?

A good one names the company, the person and the scope in its first 50-odd words. The before-and-after below uses the same made-up Example Agency and Jamie Example.

BeforeAfter
HeadingWhy choose usWhat does a technical SEO audit from Example Agency include?
OpenerAt Example Agency, we’re a team of passionate search specialists who put clients first. We believe great SEO starts with great relationships, and we’ve been helping brands grow since 2012.A technical SEO audit from Example Agency takes 10 working days. Jamie Example, our head of technical SEO, checks crawling, indexing, site speed and structured data on up to 5,000 URLs. You get a ranked list of fixes, each tied to the pages it affects, and a 60-minute call to go through it.

The after version answers a buyer’s question in 53 words with three figures. It’s built on four habits any service page can copy:

  1. One direct statement of what the service is, in the first sentence.
  2. The company’s name, and the name of the person who delivers the work.
  3. Headings phrased as the questions a buyer asks.
  4. At least one figure, such as a turnaround, a scope or a price range.

What does a 30-day order of work look like?

Fix the templates first, then the pages a buyer reads, using only the fixes above.

WeekWork (fix number)
1Move schema into the server HTML (7). Add Organization schema with sameAs to the homepage (6). Build author pages with Person schema (1).
2Add Article schema to the post and case-study templates and Service schema to the service template (2). Add bylines to posts and guides (1).
3Give your two busiest service pages question headings, answer-first openers, two sources and a named lead (1, 3, 4, 5). Rewrite the busier one with the four habits above.
4Bring your five newest case studies up to standard (8). Publish a comparison page and a how-to-choose page (9). Re-crawl and read the inner pages first.

I haven’t put a score gain against any week, because it depends on how many pages each fix reaches.

How do you check a prospect’s site in 20 minutes?

Check one service page and one case study, at about two minutes a step.

  1. Leave the homepage out, because it scores well above the inner pages.
  2. Type view-source: before each URL in Chrome and search for application/ld+json. You’re looking for Article on the case study and Service on the service page.
  3. Look for a byline with a real name, then search the source for “author”.
  4. Count the external links in the body copy. Two is the floor.
  5. Read the headings on their own, and mark each as a question, a claim or a label.
  6. Read the first 70 words under the H1. Do they say what the page offers, or talk about the agency?
  7. Search the homepage source for “sameAs” and count the profiles it lists.
  8. Request a page as a bot, such as curl -A "GPTBot" https://www.example.com/services/, and see whether you get the page or a challenge. A 403 isn’t proof the real bot is blocked, since your request doesn’t come from its network. A challenge page is still worth raising.
  9. Write one line per check. That’s your opening slide.

To see your own inner pages against these checks, run a Rubric audit and read the page-level results before the headline score.

What can’t this benchmark tell you?

This benchmark can’t tell you how often any of these sites is quoted, or why a site blocks what it blocks. It’s first-party crawl data from convenience lists, scored on one engine at one page cap on one day.

  • None of the lists is a random draw. The 12-13 August list includes 49 SEO-tool vendors (47 scored), and we report them separately.
  • Fourteen agencies sit on more than one list, so the per-list counts add up to 430.
  • Thirteen scored crawls reached fewer than 10 pages. Dropping them leaves the median at 70, and none of them is ranked.
  • A score partly reflects which pages the crawler reached. Blog-heavy crawls scored 5 points higher at the median (see “Does the page mix change a site’s score?”).
  • The engine labels subdomain roots as homepages, and some of them were error pages. Site-level homepage comparisons use the homepage on the crawled address, but the page-type tables include the subdomain roots.
  • We sent the AI-crawler probes from our own IP. From outside, we can’t see whether verified AI crawlers get the response our probes got.
  • The block flag is Rubric’s own. We didn’t score the 9 agencies it flagged, and we haven’t checked each flag by hand.
  • Page types come from a URL rule, then the crawler’s own label, and “other page” is a mixed bucket.
  • Follow-up coverage counts only crawled pages, so its figures are lower bounds.
  • Engine fits and projected scores come from Rubric’s model. They aren’t measured citations.
  • We didn’t report how many blog posts carry Article schema. The crawler labels pages with Article schema as articles, so that rate would be partly circular.
  • An earlier draft’s observation about quotable passages on legal pages didn’t survive a hand-check, so we’ve dropped it.
  • Rubric is GoGoChimp’s product, and GoGoChimp is an agency. We’ve reported where the tool’s own block flag misfired, and we haven’t named any low scorer.

Appendix A: results by list

Every column comes from the same 6 October 2026 crawl, on one engine at one page cap, so the columns can be compared. Fourteen agencies sit on more than one list, so the list columns overlap. Engine fits are Rubric’s modelled weightings.

All agencies12-13 Aug list23 Aug English-speaking23 Aug internationalProlific North Top 5010 UK SEO agenciesSEO-tool vendors
Attempted48424310099471049
Scored415 (85.7%)2008989421047
Median7071716965.56867
Mean69.770.570.868.666.669.967.5
IQR65-7566-7566.5-7765-7364-7065.5-75.7563-70
Min / max46 / 8847 / 8846 / 8755 / 8454 / 8064 / 8053 / 86
Scoring 70+222 (53.5%)116 (58.0%)58 (65.2%)40 (44.9%)12 (28.6%)3 (30.0%)13 (27.7%)
Known median (share at 70+)80 (81.7%)80 (84.0%)83 (84.3%)80 (82.0%)74.5 (66.7%)80.5 (80.0%)73 (68.1%)
Findable median (share at 70+)71 (60.5%)71 (61.0%)72 (64.0%)70 (56.2%)70.5 (61.9%)71.5 (60.0%)72 (66.0%)
Trusted median (share at 70+)52 (14.7%)53.5 (20.5%)54 (11.2%)51 (11.2%)48 (4.8%)46.5 (10.0%)51 (10.6%)
Trusted under 5016877303922721
ChatGPT fit67696965656465
Perplexity fit6869696866.566.569
AI Overviews fit666668656162.560
Gemini fit66677066606355
Copilot fit68686968656563
Claude fit7071706966.56866
Bands: 90-100 / 80-89 / 70-79 / 60-69 / 50-59 / under 500/28/194/161/29/30/17/99/69/14/10/9/49/23/6/20/2/38/43/6/00/2/10/27/3/00/1/2/7/0/00/3/10/31/3/0

Appendix B: agencies by country

Country comes from the 12-13 August list’s country label where it had one, and otherwise from the domain ending. Groups are shown where 10 or more agencies were scored, and the other 44 scored agencies sit in smaller groups. Read it as description only.

CountrySitesMedianKnownFindableTrusted
Australia367384.57357
USA4371817051
Germany and Austria227175.57158
Canada3470.580.57054.5
France1869.5767150.5
Nordics1069.5857150
UK8969777151
Spain1769837154
South Africa1063.5786657
Generic domain, no country9271807251.5

Appendix C: page types

Page-level medians, scored agencies only (n=415, 15,658 pages). Page types come from the URL first, then the crawler’s own label, and “other page” is a mixed bucket. The homepage row includes subdomain roots, so its median differs from the site-level homepage median of 82.

Page typePagesShare of pagesSites with 1+Median page scoreKnownFindableTrustedMedian wordsMedian external links
Other page7,10145.4%388687571487980
Blog post or guide4,15126.5%292759071669951
About, team, contact1,1317.2%347667769484291
Listing or index6654.2%32978897403980
Service page6464.1%175707774468840
Legal and utility5923.8%272657070481,124.50
Case study5703.6%157687771455130
Homepage5373.4%415808680678951
Product page2651.7%33818279892,4052

Share of applicable pages marked bad, by page type

CheckBlog post or guideService pageCase studyAbout, team, contactOther page
Named author9%82%78%81%78%
Content-type schemanot reported61%60%65%69%
Two or more external source links44%60%61%48%53%
Answer-first opener52%52%46%56%51%
Question or claim-shaped headings43%50%68%68%50%
Statistics carry a source28%27%49%31%33%
Statistic density26%43%16%25%30%
Entity clarity4%29%35%32%35%

Signed, sourced, dated and sized: blog posts, service pages and case studies

MeasureBlog post or guideService pageCase study
Pages4,151646570
Named author and 2+ external links1,691 (40.7%), median 8114 (2.2%), median 75.514 (2.5%), median 77.5
Neither195 (4.7%), median 67313 (48.5%), median 67261 (45.8%), median 64
Passed named author (applicable pages)3,489 of 4,147 (84.1%)28 of 600 (4.7%)77 of 554 (13.9%)
No external link1,808 (43.6%)379 (58.7%)347 (60.9%)
Under 300 words1,040 (25.1%)72 (11.1%)140 (24.6%)

Part of each median gap is mechanical, because both checks feed the page score. Blog-post schema figures are left out because they’d be partly circular.

Appendix D: schema types declared

Sites with the type on at least one crawled page.

TypeAgencies (n=415)
Organization (or a subtype)366 (88.2%)
WebSite333 (80.2%)
BreadcrumbList310 (74.7%)
Person285 (68.7%)
Article family261 (62.9%)
FAQPage194 (46.7%)
Service or ProfessionalService162 (39.0%)
Service as its own type139 (33.5%)
Offer114 (27.5%)
Review or AggregateRating106 (25.5%)
VideoObject65 (15.7%)
HowTo15 (3.6%)
No schema on any crawled page21 (5.1%)
Page-level measureAgencies
Service pages declaring Service130 of 646 (20.1%)
Case studies declaring an Article-family type83 of 570 (14.6%)

We don’t report how many blog posts declare Article schema, because the crawler labels pages with Article schema as articles.

Appendix E: top quarter against bottom quarter

The top quarter is the 104 agencies scoring 75 to 88 and the bottom quarter the 104 scoring 46 to 65.

MedianTop quarter (n=104)Bottom quarter (n=104)All agencies (n=415)
Overall786270
Known886880
Findable776471
Trusted67.54352
Homepage897682
Inner pages786170

Sites failing each check (bad on more than half the applicable pages). These are the 12 checks with the widest gap between the two quarters. HTTP status, robots.txt, canonical, internal links, reachability and word count failed at no agency in the crawl. The gaps are correlational, and partly mechanical, because each check feeds the score.

CheckTop quarter failingBottom quarter failing
Question or claim-shaped H2-H3s18 of 10285 of 104
Content-type schema22 of 10476 of 104
Entity clarity (Organization or Person plus sameAs)6 of 10457 of 104
Two or more external source links26 of 10477 of 104
Answer-first opener (40-70 words)29 of 10272 of 104
Named author35 of 10274 of 104
Fresh (updated in the last 12 months)11 of 9734 of 84
Statistic density (1.5 or more per 100 words)3 of 10432 of 104
Distinct content (no near-duplicates)1 of 10325 of 103
Exactly one H14 of 10427 of 104
Statistics carry a source16 of 10234 of 104
Meta description (50-160 characters)3 of 10420 of 104

Sources

About the author

Chris McCarron founded GoGoChimp in Glasgow in 2013 and has 13 years of conversion rate optimisation experience. He was previously Head of Growth at Roadtrippers and FOMO, and he wrote CITED, a book on getting your business recommended by AI search.

Chris McCarron on LinkedIn

About Rubric

Rubric is GoGoChimp’s AI-citability auditor. It crawls a site the way AI crawlers read it and runs 44 checks across three pillars, Known, Findable and Trusted, weighted for six AI engines: ChatGPT, Perplexity, Google AI Overviews, Gemini, Microsoft Copilot and Claude.

Rubric predicts citability, it doesn’t count citations. 70 is the quotable line.

Check your agency against the benchmark

Run Rubric on your own site and compare your pillar scores and failing checks with the figures in this study.

 Free AI SEO audit

See how your own site scores.

The Rubric audit is free and needs no sign-up.

yourstore.com · 26 pages

OverviewAction planIssuesEngines

Citability score

67

3 points below quotable

70 is the quotable line.

By engine

Perplexity

76

Claude

74

ChatGPT

65

Copilot

63

AI Overviews

61

Gemini

57

Do these first

Add Organization schema with sameAs

Name an author on every article

© 2026 GoGoChimp. All rights reserved. Call: 0141 463 6875 - Address: 8 Cheviot Drive, Newton Mearns, Glasgow, G77 5AS
Nominated — Digital Doughnut Digital Marketing Agency of the Year 2021
Shopify Partner — GoGoChimp
Select the comment + the next block ONLY (3 lines total). Paste EVERYTHING below in its place. --> '"'"'""')})}}) "'"')}})