Original research · September 24 to 30, 2026

Which sources ChatGPT, Perplexity and Gemini cite when people ask who to hire (2026)

Across 3,118 answers from ChatGPT, Perplexity and Gemini to 280 "who should I hire" and "what should I use" questions, 50.2% of the 39,763 cited links went to company websites, 25.9% to directories and review sites and 11.8% to "best X" listicles. Reddit was 2.2%, LinkedIn 1.0%, YouTube 0.2% and Wikipedia 0.1%.

By Matt Turley3,118 answers39,763 citations280 questions3 engines

Five findings

What the answers actually read

I run these questions every week for the businesses I work with, and for a few category scans. This page pools a week of it. Every number below is a count from that store.

  1. Company websites and directories carry 76.1% of the citations. Company and vendor sites got 50.2% of all 39,763 citations and directories or review sites got 25.9%. “Best X” listicles added 11.8%. Reddit was 2.2% of citations, though 16.0% of answers cited it at least once.
  2. The three engines read the web differently. Perplexity cited 20.0 sources per answer, ChatGPT 7.4 and Gemini 10.0. ChatGPT cited Reddit in 22.3% of its answers, against 13.4% for Perplexity and 10.0% for Gemini. Perplexity cited LinkedIn in 16.3% of its answers, ChatGPT in 0.5% and Gemini in 0.1%.
  3. Your category decides where you need to be listed. Directories were 50.3% of citations for travel and attractions questions and 37.1% for local services, but 1.3% for dev tools and SaaS tools, where listicles (25.8%) and docs or code (15.2%) did the work.
  4. Being cited is not being named. When an answer cited a company's own website, the answer named that company in its text 26.7% of the time (4,505 of 16,901). The rest of the time the site was a source and the business went unmentioned.
  5. The same question gives different names each time. Asked three times in a row, the set of businesses named changed in 100% of 120 local-service question and engine pairs. Only 30.2% of the businesses named showed up in all three runs, and 43.8% showed up in just one.

Chart 1 · All engines

What kind of site gets cited

Share of all 39,763 citations by source type. A company blog post counts as a company site; a “top 10” article on any site counts as a listicle.

Company and vendor sites50.2% 19,958
Directories and review sites25.9% 10,282
Listicles ("best X", "top N", alternatives, vs)11.8% 4,698
News and media3.5% 1,401
Docs and code3.4% 1,350
Reddit2.2% 889
LinkedIn1.0% 415
Government, .edu and other reference0.9% 361
Other forums and Q&A0.5% 218
YouTube0.2% 93
Wikipedia0.1% 54
Other social0.1% 44

Chart 2 · By engine

Each engine has its own habits

The bar shows each engine's citations split by source type. The table shows how often an answer cited at least one source of that type.

ChatGPT1,206 answers · 7.4 citations each
Perplexity1,179 answers · 20.0 citations each
Gemini733 answers · 10.0 citations each
Company and vendor sites Directories and review sites Listicles ("best X") Reddit, forums, LinkedIn, social, video News, docs, Wikipedia, reference
Share of answers citing at least one source of each type
Source typeChatGPTPerplexityGemini
Directories and review sites55.3%73.0%59.9%
Listicles ("best X")36.7%48.3%55.4%
Reddit22.3%13.4%10.0%
LinkedIn0.5%16.3%0.1%
YouTube0.7%0.8%4.8%
Wikipedia3.8%0.0%0.3%

Chart 3 · By category

Where to get listed depends on what you sell

Local services and travel lean on directories and review sites. Software leans on listicles and docs. Questions about hiring a service firm are answered mostly from company sites.

Local services116 questions · 1,032 answers
Travel and attractions40 questions · 757 answers
Dev tools and SaaS tools51 questions · 552 answers
Software and dev services26 questions · 321 answers
AEO and AI-visibility services17 questions · 231 answers
B2B consulting30 questions · 225 answers
Company and vendor sites Directories and review sites Listicles ("best X") Reddit, forums, LinkedIn, social, video News, docs, Wikipedia, reference

Chart 4 · Run to run

Ask twice, get a different answer

Each question went to each engine three times in a row. These bars show how often the three answers did not match.

Named businesses changed120 local-service question and engine pairs100.0%
Cited domains changed, ChatGPT371 question sets asked 3 times99.7%
Cited domains changed, Perplexity360 question sets asked 3 times15.0%
Cited domains changed, Gemini239 question sets asked 3 times98.7%

Perplexity returned the same set of cited domains on 306 of 360 repeats, while its wording and the names in it still moved. ChatGPT and Gemini changed their sources almost every time. Across all three engines, 49.1% of the domains cited in a set of three runs were cited in all three.

Top domains

The 15 most-cited sites

Counted across all categories, so the list reflects our question mix: travel questions put booking sites near the top. Our clients' and prospects' own sites are left out of this table, and so are market-specific travel sites.

Most-cited domains, 5,232 distinct domains in total
DomainTypeCitationsChatGPTPerplexityGemini
tripadvisor.comDirectories and review sites1,1782609180
reddit.comReddit86851925594
clutch.coDirectories and review sites832126585121
viator.comDirectories and review sites7577161868
designrush.comDirectories and review sites6012553640
getyourguide.comDirectories and review sites558147978
github.comDocs and code51721528319
expertise.comDirectories and review sites47011634014
linkedin.comLinkedIn40133980
ecosystem.hubspot.comDirectories and review sites2962725514
getmaxim.aiListicles ("best X", "top N", alternatives, vs)2783022523
agencies.semrush.comDirectories and review sites2212115446
thumbtack.comDirectories and review sites2041616028
angi.comDirectories and review sites202319675
themanifest.comDirectories and review sites168195099

What I take from it

What should a business do with this?

Start with the sites the engines already read for your category. If you run a local service or a travel business, that means the directories and review sites: in this sample they were 37.1% and 50.3% of the citations. If you sell software, it means the “best X” and “X alternatives” articles and your own docs.

Then fix your own site so it names you plainly. Company sites got half the citations, but a cited company was named only 26.7% of the time. ChatGPT named the cited company most often (46.6%), Gemini next (30.0%) and Perplexity least (16.3%), partly because Perplexity cites so many more pages per answer.

Last, never judge one answer. With names changing on every run, a single screenshot of ChatGPT naming you, or not naming you, says very little. Ask each question at least three times and count a win only when you show up in most runs.

Method

How I measured it

Engines and models

  • ChatGPT: the OpenAI Responses API, model gpt-4o, with the web search tool required on every call. 1,206 answers.
  • Perplexity: the Perplexity API, model sonar, which searches live on every call. 1,179 answers.
  • Gemini: Google Vertex AI, model gemini-2.5-flash, with Google Search grounding on. 733 answers. Gemini was not switched on for every question set that week, so it has fewer answers.

These are the APIs, not the consumer apps. Nobody was signed in, so no answer was shaped by a personal history. The question was the whole input, with no system prompt.

Questions, runs and dates

280 distinct questions, written the way a buyer types them: “best HVAC company in [city]”, “who should I hire to finish my app”, “best LLM gateway”. They come from 25 question sets we track for clients, prospects and category scans, grouped into six categories. Each question went to each engine three times in a row. Answers were collected from 2026-09-24 to 2026-09-30 across 41 probe runs.

Cleaning

  • The store held 8,042 rows. 3,895 were answers from one of the three engines, from a live probe, with the model recorded and the cited links kept. Older imports that did not keep cited links were left out.
  • 126 answers about a local election were dropped: they are not a buying question.
  • 651 rows were same-day cached copies of an answer already counted (two question sets asking the same question on the same day), so they were dropped. A further 0 rows repeated the exact text of an answer already counted and were dropped too.
  • That left 3,118 answers. 26 of them cited nothing.

How a citation is counted

A citation is a URL the API itself returns as a source: OpenAI's URL citation annotations, Perplexity's citations list, and Gemini's grounding sources (resolved from Google's redirect links to the real page). The same URL twice in one answer counts once. Links that only appear inside the answer text are not counted.

How a source type is assigned

By the URL, with a written rulebook: known lists of directories, review sites, marketplaces, news outlets, forums, video and social sites; “best”, “top N”, “alternatives” and “vs” in the page path for listicles; docs subdomains and paths for docs. I read the most-cited domains the rules left untyped and added the directories, job boards, local news sites and research papers among them by hand. Anything left over counts as a company or vendor site, so that bucket also holds some independent blogs.

Named vs cited

For each cited company site, I took its domain name without the ending (“acmeair” for acmeair.com) and checked whether it appears in the answer text, ignoring spaces, case and punctuation, after removing the links. This used the full answer text for 2,950 answers; answers where only a snippet was stored were left out of this one measure.

Run to run

Cited domains: every question and engine with three fresh runs in the same probe (970 sets) was checked for an identical set of cited domains. Named businesses: on four local-service leaderboards (360 answers, probed on 2026-09-26 and 2026-09-28), a list of the business names in the answers was built first, then each run was checked for which of them it named.

Privacy

The questions were written for businesses I work with and for prospects. This page reports categories and domains only. Our clients' and prospects' own sites (17 domains, 594 citations) are counted in every total but left out of the domain table. Travel sites cited only for the travel questions are left out of the table too.

Limitations

What this study cannot tell you

  • The questions are not a random sample. They were chosen because our clients, prospects and category scans care about them. Local services (1,032 answers) and travel (757 answers) are a large share, and that shapes the totals and the top-domain list.
  • APIs are not the apps. The ChatGPT, Perplexity and Gemini apps can use other models, other search settings and your history. Treat these numbers as what the engines read, not a screenshot of what you will see.
  • One week, one set of models. Engines change their search stacks often. A re-run next quarter may differ, which is why the script is saved and dated.
  • Source types come from rules, not a person reading every page. Some pages will be in the wrong bucket, mostly between company sites, blogs and listicles.
  • The named check is strict. It looks for the domain name in the answer, so a business whose name differs from its domain is missed by this check.
  • The name check across runs covers local services only, the four categories where we had a verified list of the businesses named.
  • Gemini has fewer answers (733) than ChatGPT (1,206) and Perplexity (1,179), because it was not switched on for every question set that week.

Check your own business

The AI Answer Check runs the same probe on your buyers' questions: 12 to 15 questions, three runs each on ChatGPT, Perplexity and Gemini, with who gets named instead and which sites they read. It costs $299. For how to act on it, see how to get your business mentioned in ChatGPT and how to get cited by Perplexity.