All categories

AI voice agents

AI toolsAugust 2026 edition

Which questions do you win vs lose?

Get your brand’s prompt-level results

Recommendation share

How often each brand is named, across unbranded questions and all five engines

BRANDS

SHAREShare of answers naming this brand, averaged across the five engines

ChatGPTChatGPTShare of ChatGPT answers that name this brand
ClaudeClaudeShare of Claude answers that name this brand
GeminiGeminiShare of Gemini answers that name this brand
PerplexityPerplexityShare of Perplexity answers that name this brand
Google AI OverviewsGoogle AI OverviewsShare of Google AI Overviews answers that name this brand
Retell AI
22.2
89.8
54.6
26.9
88.6
Vapi
15.7
86.1
63.0
7.4
18.1
Bland AI
13.0
63.0
62.0
14.8
22.9
4
PolyAI
35.2
44.4
33.3
14.8
32.4
5
Synthflow
4.6
41.7
39.8
9.3
29.5
6
ElevenLabs
13.0
49.1
20.4
1.9
7.6
7
Sierra
0.9
22.2
1.9
0.0
0.0
8
Deepgram
3.7
13.9
4.6
0.0
0.0
How we measurePowered by daydream

Every question changed this edition. The previous editions asked a set we wrote for the category; this one asks questions derived from what buyers actually searched or posted, so a month-over-month figure would compare two different questions and read as movement. Comparison returns next edition, when there are two mined runs to set against each other.

Framing sensitivity

The same question asked at each company size

Share of answers naming the brand

SMB50-person startupENTERPRISE2,000-person enterprisePolyAI11.1%85.4%Synthflow40.0%10.1%Retell AI65.6%52.8%ElevenLabs26.1%14.0%Bland AI32.2%42.1%Sierra6.1%9.0%Vapi42.2%34.8%Deepgram2.2%3.4%

BRANDS

SMBShare of answers naming this brand when the question is asked at this company size50-person startup

ENTERPRISEShare of answers naming this brand when the question is asked at this company size2,000-person enterprise

SPREADThe gap between the highest and lowest column in this row, in percentage points

PolyAI
11.1%
85.4%
74.3%
Synthflow
40.0%
10.1%
29.9%
Retell AI
65.6%
52.8%
12.7%
ElevenLabs
26.1%
14.0%
12.1%
Bland AI
32.2%
42.1%
9.9%
Vapi
42.2%
34.8%
7.4%
Sierra
6.1%
9.0%
2.9%
Deepgram
2.2%
3.4%
1.1%
How we measurePowered by daydream

Model divergence

The engines that name a brand least and most often, and the gap between them

BRANDS

LOWEST ENGINEThe engine that names this brand least often, and its share

HIGHEST ENGINEThe engine that names this brand most often, and its share

SPREADThe gap between the highest and lowest column in this row, in percentage points

Vapi
7.4PerplexityPerplexity
86.1ClaudeClaude
78.7
Retell AI
22.2ChatGPTChatGPT
89.8ClaudeClaude
67.6
Bland AI
13.0ChatGPTChatGPT
63.0ClaudeClaude
50.0
ElevenLabs
1.9PerplexityPerplexity
49.1ClaudeClaude
47.2
Synthflow
4.6ChatGPTChatGPT
41.7ClaudeClaude
37.0
PolyAI
14.8PerplexityPerplexity
44.4ClaudeClaude
29.6
Sierra
0.0PerplexityPerplexity
22.2ClaudeClaude
22.2
Deepgram
0.0PerplexityPerplexity
13.9ClaudeClaude
13.9
How we measurePowered by daydream

Citation trail

The sites the engines linked to when they answered

SOURCES

REACHShare of all answers that cited this source at least once% of all answers

LINKSEvery link to this source, counted across all answerstotal

DEPTHLinks to this source per answer that cited itlinks/answer

retellai.com
961
2.1
cloudtalk.io
433
1.6
bland.ai
219
1.3
getnextphone.com
222
1.6
youtube.com
248
1.9
reddit.com
135
1.1
vellum.ai
140
1.2
lumay.ai
121
1.3
getvoip.com
95
1.1
withallo.com
124
1.4
synthflow.ai
98
1.1
upfirst.ai
145
1.7
ringcentral.com
96
1.2
trillet.ai
118
1.5
lindy.ai
88
1.1
zeeg.me
132
1.8
smash.vc
74
1.2
happyrobot.ai
74
1.2
telnyx.com
75
1.2
dapta.ai
68
1.2
How we measurePowered by daydream

Which questions do you win vs lose?

Get your brand’s prompt-level results

How we measure this3 of 12 questions published

We write a fixed set of category questions, none of which names a brand, and count how often each brand comes up. A brand is present in an answer only when one of its names literally appears in the text, because a string match cannot hallucinate.

Every question starts from one somebody already asked. Some come from search demand, where the figure is how many people typed that phrasing in a month. The rest come from a public forum post, and we link to the page it was written on. We rewrite each one into a plain question, because a search string is a fragment and a forum post is written the way people type, and neither is a fair thing to ask an engine. We change the wording and never the subject, and the original is published beside every question we disclose. The set is frozen before a single answer is collected, so every engine and every edition is asked exactly the same thing.

12questions scored
3buyer framings each
36prompts run
5engines

The 3 we publish, of 12

Chosen to span both where the questions come from and how they are shaped, so the sample describes the battery rather than one corner of it. That is 25% of the questions the ranking comes from.

  • What is the best AI receptionist appas asked: best ai receptionist app20/mocategoryawareness
  • What is the best AI phone receptionistas asked: best ai phone receptionist50/mocategoryawareness
  • What is the best AI voice assistantas asked: best ai voice assistant250/mocategoryawareness

Each is asked 3 ways

  • for a 50-person startup50-person startup
  • for a 2,000-person enterprise with high call volumes2,000-person enterprise
  • at the lowest costCheapest option

Why the rest stays private

A published battery invites brands to write pages against the exact wording, at which point a score moves without the brand’s actual standing moving and the measurement stops describing anything. Benchmark suites keep a held-out set for the same reason. The method is public so it can be judged, a quarter of the questions are public so it can be checked, and the rest stays private so the numbers stay worth checking.

What the numbers are, and aren’t

  • The August 2026 edition is one dated run, not a rolling average.
  • Share is averaged across the engines rather than pooled, so an engine that returned fewer answers cannot look like a brand losing ground.
  • A move is marked only when a two-proportion test puts it outside what the sample size can explain. A few points between neighbouring brands is a tie.
  • Rankings reflect how often AI engines name a brand. They are not endorsements by daydream.
  • 16 further questions name competitors directly. None of them feeds the ranking, because a question that already names the contenders cannot measure who gets recommended. They do count towards the citation trail, which is measured over every answer we collected rather than the unbranded ones alone.