Parse
Work with usPricing
Sign inCheck your brand
Research/The directory tax is a myth: G2 and Capterra are #1 in 0 of 835 categories

The directory tax is a myth: G2 and Capterra are #1 in 0 of 835 categories

The common generative-engine-optimization advice says one directory listing can control whether AI recommends you. In Parse's category evidence, the directory chokepoint mostly is not there.

By Dimitry Apollonsky · July 8, 2026 · 8 min read

Median share of evidence from the top sources
  • Top 1 source7.4%
  • Top 3 sources17.4%
  • Top 5 sources24.8%
  • Top 10 sources37.1%
The median category draws from hundreds of domains. Even the top 10 domains carry only 37.1% of evidence.
▸Contents
  • There is no single source chokepoint in the median category
  • The major directories are almost never the #1 source
  • Three software brands in four are recommended with no directory citation
  • Directory citations are a tiny slice of recommendation evidence
  • When a category does lean on one source, it is usually community or video
  • The category leaders are social, editorial, and owned sites
  • Mainstream software categories are spread out too
  • Both engines spread evidence across many sources
  • True chokepoints exist, but they are outliers
  • The practical playbook is a spread of sources, not a directory tax
  • Get the data
  • Sources
  • Related research
Contents
  • There is no single source chokepoint in the median category
  • The major directories are almost never the #1 source
  • Three software brands in four are recommended with no directory citation
  • Directory citations are a tiny slice of recommendation evidence
  • When a category does lean on one source, it is usually community or video
  • The category leaders are social, editorial, and owned sites
  • Mainstream software categories are spread out too
  • Both engines spread evidence across many sources
  • True chokepoints exist, but they are outliers
  • The practical playbook is a spread of sources, not a directory tax
  • Get the data
  • Sources
  • Related research

We analyzed 5.42M evidence links, the cited pages that sit behind AI brand recommendations, across 835 monitored categories, 49,638 brands, 7,611 prompts, and 101,677 source domains.

0
categories where G2 or Capterra is the #1 source
7.4%
median top-source share per category
260
median distinct source domains per category
14.2%
brands with any directory citation

There is no single source chokepoint in the median category

Across 835 monitored categories, the median category's single most-cited source domain accounts for only 7.4% of recommendation evidence. The median category draws on 260 distinct domains.

That is the opposite of a tax. In most categories, AI recommendations are assembled from a wide range of sources, not unlocked by one mandatory listing.

7.4%
Median top-source share
260
Median distinct domains
835
Categories measured

Takeaway

For most categories, the source strategy is building a spread of sources, not paying one gatekeeper.

The major directories are almost never the #1 source

G2, Capterra, GetApp, Software Advice, SourceForge, TrustRadius, and Trustpilot are each the #1 source in zero of the 835 categories. Clutch is #1 in exactly one. By source type, a directory is #1 in only 8 categories.

Where the major directories do appear, their average shares are small: G2 appears in 333 categories at 0.3%, Capterra in 186 at 0.2%, GetApp in 174 at 0.4%, and Clutch in 35 at 0.9%.

Directory gatekeeper test
clutch.co faviconclutch.co1
g2.com favicong2.com0
capterra.com faviconcapterra.com0
getapp.com favicongetapp.com0
softwareadvice.com faviconsoftwareadvice.com0
sourceforge.net faviconsourceforge.net0
trustradius.com favicontrustradius.com0
trustpilot.com favicontrustpilot.com0

Three software brands in four are recommended with no directory citation

The charitable test is whether a directory citation is present anywhere behind recommended brands. Even there, the directory story is weak. Among all brands with at least 20 evidence links, 14.2% carry any directory citation. In software categories, where the directory premise is strongest, 24.0% do.

  • Brands in software categories24.0%
  • All brands14.2%
Share of brands with at least one directory citation.

Takeaway

Directory presence may help some brands, but absence is common even among brands AI already recommends.

Directory citations are a tiny slice of recommendation evidence

Directory citations make up 0.74% of the recommendation evidence for brands in software categories, and 0.51% across all brands. That is not enough to explain recommendation visibility as a directory-led system.

  • Brands in software categories0.74%
  • All brands0.51%
Directory citation share of recommendation evidence.

When a category does lean on one source, it is usually community or video

The domains that do become #1 sources are more often YouTube logoYouTube and Reddit logoReddit than classic directories. YouTube logoYouTube is the #1 source in 169 categories; Reddit logoReddit is #1 in 142. Together they lead 311 categories, or 37% of the measured set.

  • youtube.com faviconyoutube.com169
  • reddit.com faviconreddit.com142
  • All directory type8
Number of categories where each domain is the #1 source.

The category leaders are social, editorial, and owned sites

By source type, social sources lead 321 categories, general editorial leads 230, and a brand's own site leads 129. Directories lead only 8. That mix says your source work should start with the pages, communities, and owned assets AI already uses in the category.

Source type of the #1 domain
Social321
General/editorial230
Brand or vendor site129
News and finance34
Ecommerce30
Directory8

Mainstream software categories are spread out too

The finding is not only a side effect of small, obscure categories. CRM and Sales Pipeline Software draws on 852 domains, with YouTube logoYouTube top at 8.3%. Payroll and HRIS draws on 853 domains, with Gusto logoGusto's own domain top at 8.8%. Project management draws on 517 domains, with YouTube logoYouTube top at 6.2%.

Example category source concentration
Payroll and HRIS853gusto.com favicongusto.com8.8
CRM and Sales Pipeline Software852youtube.com faviconyoutube.com8.3
Small-business accounting806youtube.com faviconyoutube.com9.9
Email marketing715reddit.com faviconreddit.com6.6
Project management517youtube.com faviconyoutube.com6.2

Both engines spread evidence across many sources

The median per-category top-source share is 6.3% for ChatGPT logoChatGPT and 8.8% for Google logoGoogle's answer engines. The engine that cites fewer sources is not more gatekept; if anything, ChatGPT logoChatGPT is slightly more spread out among the categories that qualified.

  • Google logoGoogle answer engines8.8%
  • ChatGPT logoChatGPT6.3%
Median top-source share by engine family.

True chokepoints exist, but they are outliers

Across all 835 categories, the top-source share exceeds 14% only in the top 10% of categories. The single most concentrated category reached 53.2%, an accessories category dominated by YouTube logoYouTube. Chokepoints happen, but the median category is not built that way.

13.9%
90th percentile top-source share
53.2%
Most concentrated category
7.4%
Median top-source share

The practical playbook is a spread of sources, not a directory tax

Directory work is not useless. It is just too narrow to be the default explanation for AI recommendations. In many categories, the pages that matter most will be YouTube logoYouTube videos, Reddit logoReddit threads, editorial comparisons, a brand's own pages, and category-specific publishers.

Start by mapping the actual top domains in your category, then prioritize the small set of sources AI repeatedly uses. Treat directories as one kind of source among many, not the toll booth for the whole market.

Get the data

Dataset CSVHeadline source concentration and directory-presence figures.

Sources

  1. Companion analysis: directory tax in AI recommendations
  2. Search Engine Land: AI platforms and generative engine optimization coverage
  3. Semrush: AI citation source patterns

Related research

Which domains AI cites most in its answers
A source-domain cut of AI answers: which domains supply the evidence behind AI recommendations, and why community platforms are not a side channel.
What pages AI cites most: best-of listicles dominate
Domain studies tell you AI loves Reddit, YouTube, and Wikipedia. The page level shows what it actually pulls from them: six in ten of AI's most-cited pages are “Best X” listicles.
YouTube vs Reddit: which source AI cites most
YouTube ranks first among cited source domains, but only 8.4% ahead of Reddit, and Reddit reaches more distinct questions. A look at the top seven.

About this research

Dimitry Apollonsky

Founder, Parse

I built Parse to track where AI answers really come from: the sources they cite and the brands they name. DM me on LinkedIn to talk shop.

Find the sources that actually back recommendations in your category.

Run a free check against live AI answers — no account needed.

Parse

Parse indexes AI recommendations so brands know where they stand.

Products

  • Brands
  • Markets
  • Integrations
  • Work with us
  • Pricing
  • MCP

Resources

  • Research
  • Methodology
  • Blog

© 2026 Parse. All rights reserved.

LegalPrivacy PolicyTerms of Service