Which sources ChatGPT, Perplexity and Gemini cite when people ask who to hire (2026)
Across 3,118 answers from ChatGPT, Perplexity and Gemini to 280 "who should I hire" and "what should I use" questions, 50.2% of the 39,763 cited links went to company websites, 25.9% to directories and review sites and 11.8% to "best X" listicles. Reddit was 2.2%, LinkedIn 1.0%, YouTube 0.2% and Wikipedia 0.1%.
By Matt Turley3,118 answers39,763 citations280 questions3 engines
I run these questions every week for the businesses I work with, and for a few category scans. This page pools a week of it. Every number below is a count from that store.
Company websites and directories carry 76.1% of the citations. Company and vendor sites got 50.2% of all 39,763 citations and directories or review sites got 25.9%. “Best X” listicles added 11.8%. Reddit was 2.2% of citations, though 16.0% of answers cited it at least once.
The three engines read the web differently. Perplexity cited 20.0 sources per answer, ChatGPT 7.4 and Gemini 10.0. ChatGPT cited Reddit in 22.3% of its answers, against 13.4% for Perplexity and 10.0% for Gemini. Perplexity cited LinkedIn in 16.3% of its answers, ChatGPT in 0.5% and Gemini in 0.1%.
Your category decides where you need to be listed. Directories were 50.3% of citations for travel and attractions questions and 37.1% for local services, but 1.3% for dev tools and SaaS tools, where listicles (25.8%) and docs or code (15.2%) did the work.
Being cited is not being named. When an answer cited a company's own website, the answer named that company in its text 26.7% of the time (4,505 of 16,901). The rest of the time the site was a source and the business went unmentioned.
The same question gives different names each time. Asked three times in a row, the set of businesses named changed in 100% of 120 local-service question and engine pairs. Only 30.2% of the businesses named showed up in all three runs, and 43.8% showed up in just one.
Chart 1 · All engines
What kind of site gets cited
Share of all 39,763 citations by source type. A company blog post counts as a company site; a “top 10” article on any site counts as a listicle.
The bar shows each engine's citations split by source type. The table shows how often an answer cited at least one source of that type.
ChatGPT1,206 answers · 7.4 citations each
48%22%12%11%
Perplexity1,179 answers · 20.0 citations each
47%30%11%
Gemini733 answers · 10.0 citations each
64%16%14%
Company and vendor sites Directories and review sites Listicles ("best X") Reddit, forums, LinkedIn, social, video News, docs, Wikipedia, reference
Share of answers citing at least one source of each type
Source type
ChatGPT
Perplexity
Gemini
Directories and review sites
55.3%
73.0%
59.9%
Listicles ("best X")
36.7%
48.3%
55.4%
Reddit
22.3%
13.4%
10.0%
LinkedIn
0.5%
16.3%
0.1%
YouTube
0.7%
0.8%
4.8%
Wikipedia
3.8%
0.0%
0.3%
Chart 3 · By category
Where to get listed depends on what you sell
Local services and travel lean on directories and review sites. Software leans on listicles and docs. Questions about hiring a service firm are answered mostly from company sites.
Local services116 questions · 1,032 answers
52%37%
Travel and attractions40 questions · 757 answers
37%50%
Dev tools and SaaS tools51 questions · 552 answers
46%26%23%
Software and dev services26 questions · 321 answers
67%11%11%
AEO and AI-visibility services17 questions · 231 answers
60%16%12%9%
B2B consulting30 questions · 225 answers
61%15%13%
Company and vendor sites Directories and review sites Listicles ("best X") Reddit, forums, LinkedIn, social, video News, docs, Wikipedia, reference
Chart 4 · Run to run
Ask twice, get a different answer
Each question went to each engine three times in a row. These bars show how often the three answers did not match.
Named businesses changed120 local-service question and engine pairs100.0%
Perplexity returned the same set of cited domains on 306 of 360 repeats, while its wording and the names in it still moved. ChatGPT and Gemini changed their sources almost every time. Across all three engines, 49.1% of the domains cited in a set of three runs were cited in all three.
Top domains
The 15 most-cited sites
Counted across all categories, so the list reflects our question mix: travel questions put booking sites near the top. Our clients' and prospects' own sites are left out of this table, and so are market-specific travel sites.
Most-cited domains, 5,232 distinct domains in total
Domain
Type
Citations
ChatGPT
Perplexity
Gemini
tripadvisor.com
Directories and review sites
1,178
260
918
0
reddit.com
Reddit
868
519
255
94
clutch.co
Directories and review sites
832
126
585
121
viator.com
Directories and review sites
757
71
618
68
designrush.com
Directories and review sites
601
25
536
40
getyourguide.com
Directories and review sites
558
1
479
78
github.com
Docs and code
517
215
283
19
expertise.com
Directories and review sites
470
116
340
14
linkedin.com
LinkedIn
401
3
398
0
ecosystem.hubspot.com
Directories and review sites
296
27
255
14
getmaxim.ai
Listicles ("best X", "top N", alternatives, vs)
278
30
225
23
agencies.semrush.com
Directories and review sites
221
21
154
46
thumbtack.com
Directories and review sites
204
16
160
28
angi.com
Directories and review sites
202
31
96
75
themanifest.com
Directories and review sites
168
19
50
99
What I take from it
What should a business do with this?
Start with the sites the engines already read for your category. If you run a local service or a travel business, that means the directories and review sites: in this sample they were 37.1% and 50.3% of the citations. If you sell software, it means the “best X” and “X alternatives” articles and your own docs.
Then fix your own site so it names you plainly. Company sites got half the citations, but a cited company was named only 26.7% of the time. ChatGPT named the cited company most often (46.6%), Gemini next (30.0%) and Perplexity least (16.3%), partly because Perplexity cites so many more pages per answer.
Last, never judge one answer. With names changing on every run, a single screenshot of ChatGPT naming you, or not naming you, says very little. Ask each question at least three times and count a win only when you show up in most runs.
Method
How I measured it
Engines and models
ChatGPT: the OpenAI Responses API, model gpt-4o, with the web search tool required on every call. 1,206 answers.
Perplexity: the Perplexity API, model sonar, which searches live on every call. 1,179 answers.
Gemini: Google Vertex AI, model gemini-2.5-flash, with Google Search grounding on. 733 answers. Gemini was not switched on for every question set that week, so it has fewer answers.
These are the APIs, not the consumer apps. Nobody was signed in, so no answer was shaped by a personal history. The question was the whole input, with no system prompt.
Questions, runs and dates
280 distinct questions, written the way a buyer types them: “best HVAC company in [city]”, “who should I hire to finish my app”, “best LLM gateway”. They come from 25 question sets we track for clients, prospects and category scans, grouped into six categories. Each question went to each engine three times in a row. Answers were collected from 2026-09-24 to 2026-09-30 across 41 probe runs.
Cleaning
The store held 8,042 rows. 3,895 were answers from one of the three engines, from a live probe, with the model recorded and the cited links kept. Older imports that did not keep cited links were left out.
126 answers about a local election were dropped: they are not a buying question.
651 rows were same-day cached copies of an answer already counted (two question sets asking the same question on the same day), so they were dropped. A further 0 rows repeated the exact text of an answer already counted and were dropped too.
That left 3,118 answers. 26 of them cited nothing.
How a citation is counted
A citation is a URL the API itself returns as a source: OpenAI's URL citation annotations, Perplexity's citations list, and Gemini's grounding sources (resolved from Google's redirect links to the real page). The same URL twice in one answer counts once. Links that only appear inside the answer text are not counted.
How a source type is assigned
By the URL, with a written rulebook: known lists of directories, review sites, marketplaces, news outlets, forums, video and social sites; “best”, “top N”, “alternatives” and “vs” in the page path for listicles; docs subdomains and paths for docs. I read the most-cited domains the rules left untyped and added the directories, job boards, local news sites and research papers among them by hand. Anything left over counts as a company or vendor site, so that bucket also holds some independent blogs.
Named vs cited
For each cited company site, I took its domain name without the ending (“acmeair” for acmeair.com) and checked whether it appears in the answer text, ignoring spaces, case and punctuation, after removing the links. This used the full answer text for 2,950 answers; answers where only a snippet was stored were left out of this one measure.
Run to run
Cited domains: every question and engine with three fresh runs in the same probe (970 sets) was checked for an identical set of cited domains. Named businesses: on four local-service leaderboards (360 answers, probed on 2026-09-26 and 2026-09-28), a list of the business names in the answers was built first, then each run was checked for which of them it named.
Privacy
The questions were written for businesses I work with and for prospects. This page reports categories and domains only. Our clients' and prospects' own sites (17 domains, 594 citations) are counted in every total but left out of the domain table. Travel sites cited only for the travel questions are left out of the table too.
Limitations
What this study cannot tell you
The questions are not a random sample. They were chosen because our clients, prospects and category scans care about them. Local services (1,032 answers) and travel (757 answers) are a large share, and that shapes the totals and the top-domain list.
APIs are not the apps. The ChatGPT, Perplexity and Gemini apps can use other models, other search settings and your history. Treat these numbers as what the engines read, not a screenshot of what you will see.
One week, one set of models. Engines change their search stacks often. A re-run next quarter may differ, which is why the script is saved and dated.
Source types come from rules, not a person reading every page. Some pages will be in the wrong bucket, mostly between company sites, blogs and listicles.
The named check is strict. It looks for the domain name in the answer, so a business whose name differs from its domain is missed by this check.
The name check across runs covers local services only, the four categories where we had a verified list of the businesses named.
Gemini has fewer answers (733) than ChatGPT (1,206) and Perplexity (1,179), because it was not switched on for every question set that week.