CiteWorks Studio

How AI Search Is Recommending Live Chat Software: Monthly Trends

Mark HuntleyBy Mark HuntleyFounder and CEO
10 minutes read

Key Takeaways

  • Tidio remained the category leader in September 2026 with 68.9% valid recommendation coverage, while tawk.to held second at 61.3%.
  • LiveChat was the only significant riser versus July, increasing from 11.2% to 23.3% after an August spike to 40.2% and a September pullback.
  • Intercom's overall coverage was stable at 46.3%, but its top-3 rate rose 10.3 points and its rank-one rate rose 5.5 points from baseline.
  • Zendesk Chat declined from 45.6% to 42.6% as presence fell, while the benchmark remained concentrated on discovery-stage prompts rather than pricing or comparison queries.

Executive Summary

Tidio remains the category leader in Live Chat Software, holding 68.9% valid recommendation coverage in September 2026 against 67.9% in July 2026, a 1.0-point gain the benchmark classifies as stable. tawk.to holds second position at 61.3%, down from 63.0%, keeping the gap to the leader at 7.6 points.

The one brand the benchmark flags as a significant riser this period is LiveChat (Text S.A., formerly LiveChat Software S.A.), whose valid recommendation coverage rose from 11.2% in July 2026 to 23.3% in September 2026, a 12.1-point gain. That movement came after an August spike to 40.2%, followed by a 16.9-point pullback in September. Despite the retreat, LiveChat remains well above its July baseline. The gap between LiveChat and Zendesk Chat narrowed from 34.4 points in July 2026 to 19.3 points in September 2026.

Every other tracked brand moved within the range the benchmark classifies as normal month-to-month variation. Zendesk Chat posted the largest decline from baseline, down 3.0 points from 45.6% to 42.6%, while Intercom made the strongest steady upward move, rising 1.0 point to 46.3%. Intercom also recorded the category's notable placement gains, with its top-3 rate rising 10.3 points and rank-one rate rising 5.5 points from baseline.

Each monthly run begins with 800 prompt-surface observations (607 unique questions in July 2026, 626 in August 2026, 580 in September 2026) across the benchmark's defined AI/search surface universe. All 800 mentioned a tracked brand or competitor in each month; 446 were relevant and 354 were irrelevant in July 2026, 478 were relevant and 322 were irrelevant in August 2026, and 520 were relevant and 280 were irrelevant in September 2026. The public metrics use the 349 qualified observations in July 2026, 371 in August 2026, and 408 in September 2026 that survive both qualification stages.

AI recommendation trend

valid recommendation coverage, Jul 2026 to Sep 2026

0%20%40%60%80%Jul 2026Aug 2026Sep 2026
  • Tidio68.9%
  • tawk.to61.3%
  • Intercom46.3%
  • Zendesk Chat42.6%
  • HubSpot Live Chat40.4%
  • Crisp36.5%
  • LiveChat (Text S.A. (formerly LiveChat Software S.A.)23.3%
  • Drift6.6%
  • Olark3.2%
  • LivePerson0.5%

Key Findings

Signal

September 2026 finding

Category leader

Tidio leads at 68.9% valid recommendation coverage, stable versus 67.9% in July 2026

Largest riser

LiveChat (Text S.A., formerly LiveChat Software S.A.) rose 12.1 points, from 11.2% to 23.3% — the only brand classified as a significant riser this period

Largest decliner

Zendesk Chat fell 3.0 points, from 45.6% to 42.6%, within the benchmark's stable range

Significant placement gains

Intercom's top-3 rate rose 10.3 points and rank-one rate rose 5.5 points from baseline

LiveChat pullback

LiveChat declined 16.9 points from August 2026's 40.2% to 23.3% in September 2026

Stable leadership

Tidio's 68.9% and tawk.to's 61.3% keep the top two positions unchanged

Benchmark Context

The report separates the raw collection universe from the qualified analysis set. Brand-level recommendation percentages are calculated within the qualified benchmark set.

Research stage

Jul 2026

Sep 2026

What it represents

Source prompt-surface observations collected

800

800

Total prompts run across the AI/search surface universe

Unique questions

607

580

Distinct question phrasings in the run

Brand / competitor mentions

800

800

Prompts that mentioned a tracked brand or competitor

Relevant prompts

446

520

Prompts relevant to the live chat software category

Irrelevant prompts

354

280

Prompts outside the category scope

Qualified benchmark observations

349

408

Public denominator for brand-level recommendation rates

Qualified surface breadth

6

6

AI surface families with at least one qualified observation (ChatGPT, Copilot, Gemini, Perplexity, AI Overviews, AI Mode)

These stages define the denominator used for every brand-level percentage that follows; the qualified benchmark observations row is the base for all coverage, presence, and placement rates in this report. August 2026 fell between these two months with 371 qualified observations.

Benchmark-Level Metrics

Metric

Jul 2026

Sep 2026

Change

Qualified observations

349

408

+59

Companies tracked

10

10

No change

Recommendation-shaped answer share

53.6%

47.8%

Down 5.8 points

Valid recommendation shortlist share

75.9%

77.9%

Up 2.0 points

Category leader by coverage

Tidio (67.9%)

Tidio (68.9%)

Stable

The qualified observation count grew steadily across the three months, from 349 to 371 to 408, while the recommendation-shaped answer share followed a different path: 53.6% in July, down to 46.6% in August, then recovering to 47.8% in September. The valid recommendation shortlist share rose to 77.9%, meaning a larger proportion of the qualified set contained usable recommendation signals than at baseline.

AI Recommendation Trend

Leadership holds while the mid-field recalibrates after a spike

The top of the category held firm in September 2026, with Tidio and tawk.to retaining their positions. The more notable story is the mid-field: LiveChat's August surge to 40.2% did not hold, settling back to 23.3% in September, which reshaped the competitive spacing around the third-through-sixth positions.

Brand

Jul 2026

Sep 2026

Movement

Sep 2026 rank

Tidio

67.9%

68.9%

Up 1.0 points

1st

tawk.to

63.0%

61.3%

Down 1.7 points

2nd

Intercom

45.3%

46.3%

Up 1.0 points

3rd

Zendesk Chat

45.6%

42.6%

Down 3.0 points

4th

HubSpot Live Chat

42.4%

40.4%

Down 2.0 points

5th

Crisp

33.2%

36.5%

Up 3.3 points

6th

LiveChat (Text S.A., formerly LiveChat Software S.A.)

11.2%

23.3%

Up 12.1 points

7th

Drift

6.6%

6.6%

No change

8th

Olark

6.0%

3.2%

Down 2.8 points

9th

LivePerson

1.1%

0.5%

Down 0.6 points

10th

The only movement the benchmark's series-based check classifies as a significant riser this period is LiveChat's 12.1-point coverage gain from baseline; every other brand's movement falls within its normal range. The category-level change in September comes from the combination of several smaller movements plus LiveChat's retreat from its August spike rather than from any single unstable shift.

What Changed This Month

LiveChat (Text S.A., formerly LiveChat Software S.A.): Significant Riser With a Two-Month Story

LiveChat's valid recommendation coverage rose from 11.2% in July 2026 to 23.3% in September 2026, a 12.1-point gain the benchmark's series-based check classifies as significant. However, the path was not linear: the brand spiked to 40.2% in August 2026 before pulling back 16.9 points in September. The September level still stands well above the July baseline, which is why the brand is classified as a significant riser in the cumulative series despite the month-over-month decline.

The supporting metrics moved with the baseline-to-current story. Raw mention presence rose from 18.1% in July to 34.3% in September, a significant 16.2-point gain. The recommended top-3 rate climbed from 7.7% to 20.3%, up 12.6 points and significant, while the rank-one rate rose from 4.3% to 12.5%, up 8.2 points and significant. In absolute terms, valid recommendations grew from 39 in July 2026 to 95 in September 2026, with rank-one placements rising from 15 to 51.

LiveChat's presence rate nearly doubled from 18.1% in July to 34.3% in September, while its top-3 and rank-one rates both stayed far above baseline even after the August peak receded. The benchmark records the size and direction of this change; it cannot establish what caused the August spike or the September pullback.

Highest-priority diagnostic: Which prompt categories drove the September retention above baseline, and did the August spike concentrate on particular surfaces that then normalized?

Intercom: Placement Gains While Coverage Holds

Intercom's valid recommendation coverage moved from 45.3% in July 2026 to 46.3% in September 2026, up 1.0 point — within the benchmark's stable range for the brand.

The more notable signal is in placement quality. Intercom's top-3 rate rose from 22.1% to 32.4%, a significant 10.3-point gain, with top-3 placements increasing from 77 to 132. The rank-one rate rose from 6.3% to 11.8%, a significant 5.5-point gain, with rank-one placements rising from 22 to 48. Raw mention presence moved in the opposite direction, from 64.8% in July 2026 to 62.0% in September 2026.

Intercom is winning top-3 and rank-one positions in a larger share of the answers where it appears, even as its raw presence edges down — a placement gain that outpaces its visibility trend.

Highest-priority diagnostic: Which prompt types are yielding the additional top-3 and rank-one placements, and which brands are being displaced in those positions?

Zendesk Chat: Steady Declines From Baseline

Zendesk Chat's valid recommendation coverage fell from 45.6% in July 2026 to 42.6% in September 2026, down 3.0 points. That is a two-month decline, with the brand moving from 45.6% to 45.0% in August and then to 42.6% in September. The brand remains 4th in the category, but the gap to 3rd-place Intercom widened.

Raw mention presence fell from 61.6% in July 2026 to 56.1% in September 2026, down 5.5 points. The top-3 rate moved in the opposite direction, from 17.5% to 19.9%, while the rank-one rate held nearly flat at 3.2% versus 3.1%. In absolute terms, valid recommendations rose from 159 in July 2026 to 174 in September 2026 as the qualified observation count grew.

Zendesk Chat's coverage decline comes alongside a falling presence rate, meaning the brand appears in fewer answers overall, even though its placement within those answers improved slightly.

Highest-priority diagnostic: Which prompts no longer surface Zendesk Chat, and which brands are taking its place in those answers?

Olark: Continued Marginal Presence

Olark's valid recommendation coverage fell from 6.0% in July 2026 to 3.2% in September 2026, a 2.8-point decline that falls within the benchmark's stable range but continues a pattern of marginal presence for the brand.

Raw mention presence fell from 8.6% to 4.7%, a significant 3.9-point decline. The top-3 rate fell from 0.6% to 0.2%, with top-3 placements falling from 2 to 1. In absolute terms, valid recommendations fell from 21 in July 2026 to 13 in September 2026, and the brand's total mentions fell from 30 to 19 across the surface universe.

Olark remains visible in a narrow set of prompts, but within that set it is placed in recommendation lists and shortlists less often than at baseline.

Highest-priority diagnostic: Which prompts still mention Olark, and what attributes appear alongside the brand in those remaining contexts?

Buyer-Intent Interpretation

Buyer-intent cluster

What it captures

Strategic question

Brand Recommendation

Prompts seeking a recommended live chat software solution

Which brands earn the recommendation when a buyer asks for a single option?

Pricing & Value

Prompts comparing cost and value across options

What role does pricing information play in AI recommendations?

Multi-Brand Comparison

Prompts asking AI to weigh multiple options side by side

Which brands are surfaced in head-to-head comparisons, and who wins?

In September 2026, all 408 qualified observations fell into the discovery and consideration cluster. The benchmark run captured no qualified pricing or value observations, and no multi-brand comparison observations. This means the public metrics describe which brands AI systems recommend during initial discovery, but they cannot yet describe how those systems handle price sensitivity or direct competitive comparison.

The practical implication is that the benchmark measures first-position visibility and recommendation presence, not the full buyer journey. A brand could perform strongly in discovery prompts while faring differently in later-stage comparison questions. The current data cannot distinguish between those scenarios, which is exactly the kind of question a company-level analysis can answer.

Brand Opportunity Summary

Brand

Sep 2026 coverage

Current signal

Highest-priority diagnostic

Tidio

68.9%

Category leader, stable

Which surfaces and prompts sustain the leadership position?

tawk.to

61.3%

Second position, stable

Where is the remaining gap to the leader concentrated?

Intercom

46.3%

Top-3 and rank-one rates rising

Which prompt types drive the placement gains?

Zendesk Chat

42.6%

Coverage declining from baseline

Which prompts no longer surface the brand?

HubSpot Live Chat

40.4%

Stable, modest declines

Which brands are taking share in its strong prompts?

Crisp

36.5%

Coverage rising, top-3 rate gaining

Which evidence sources support the higher placement?

LiveChat (Text S.A., formerly LiveChat Software S.A.)

23.3%

Significant riser, post-spike level

Which prompt categories sustained the September level?

Drift

6.6%

Stable, low coverage

Which prompts still surface Drift at all?

Olark

3.2%

Marginal, declining presence

What attributes remain associated with Olark?

LivePerson

0.5%

Minimal presence, declining

Is the brand present in any meaningful prompt cluster?

The benchmark identifies where attention is warranted; a company-level analysis is needed to explain why.

Evidence Behind the Benchmark

The aggregate metrics are built from prompt-level observations (query, surface, recommendation outcome, rank, sentiment, and citations where exposed). Company-level analysis can go deeper into prompt, competitor, surface, and evidence patterns. Source presence is not automatically treated as proof of causation.

About This Benchmark

This report is part of the AI Industry Market Discovery research program from CiteWorks Studio.

Report-Specific Interpretation Notes

  • Small-count movements: brands with low absolute counts (LivePerson at 2 valid recommendations in September 2026, Olark at 13) can show large percentage swings from small changes in raw numbers. Read their movements with appropriate caution.
  • Qualified denominator: all percentage figures use the 408 qualified observations in September 2026, not the 800 prompts collected. The funnel from collection to qualification is disclosed in Benchmark Context.
  • Directional analysis: month-over-month movement identifies changes worth investigating, but it does not by itself establish the cause of those changes. LiveChat's August spike and September pullback both require investigation.
  • LiveChat is the only brand the benchmark's series-based check classifies as a significant riser this period; its presence, top-3, and rank-one metrics all moved significantly above baseline even as the brand retreated from its August peak.

Next Step

The Public Benchmark Shows Where a Brand Is Winning or Losing. A Company-Level Audit Shows Why.

The aggregate percentages in this report describe which brands AI systems recommend, but they do not explain why those recommendations form. Beneath LiveChat's 12.1-point baseline gain are questions about which high-intent prompts it now wins, why its August spike did not hold, which brand takes the recommendation when it loses, what attributes AI systems associate with each option, and which external sources shape those answers. The same questions apply to every brand in the category, whether stable or moving.

A company-specific AI visibility audit maps those prompt, surface, competitor, ranking, sentiment, and evidence-source patterns into a prioritized visibility strategy. Rather than reacting to aggregate movements, the audit identifies the specific levers that influence where and how AI systems recommend a brand.

Request an AI visibility audit

/ Take the next step

Want to Understand Your AI Citation Footprint?

We start every engagement with a full audit of how AI systems reference your brand today.

Measurable, Repeatable Programme

Build a durable foundation of credible citations that compounds over time and continues to influence AI answers as new queries emerge

Citation Architecture Review

Identify which high-authority community sources are and aren't working in your favour across AI platforms.

AI Visibility Audit

Understand exactly how LLMs are referencing your brand today and which sources are shaping those answers.

/ Learn More

Understanding AI search visibility.

AI search experiences create answers by pulling information from many places online and summarizing it into a single response.

What Is AI Citation Intelligence?
AI citation intelligence is the process of measuring where AI platforms source their information and how frequently a brand is mentioned or referenced in AI-generated responses. Because LLMs synthesize across multiple sources, the sites and brands that appear repeatedly tend to influence how a topic or company is framed. This practice focuses on identifying which sources shape AI outputs and tracking brand visibility across different AI systems.
What Is Citation Architecture?
Citation architecture describes the set of sources that consistently inform how AI systems talk about a brand, product, or topic. LLMs draw from websites, articles, forums, and public discussion, and the sources they rely on most often become the backbone of their answers. Building strong citation architecture means ensuring that accurate, credible, high authority sources are the ones most likely to shape the way AI tools summarize and recommend a brand.
What Is Generative Engine Optimization?
Generative engine optimization (GEO) is the practice of improving the chances that AI systems use and cite your brand or content when generating answers. While traditional SEO is centered on ranking pages in search results, GEO focuses on how LLMs retrieve, interpret, and combine information when responding to a question. The objective is to strengthen the content and sources AI systems rely on, so your brand is treated as a trusted reference in AI responses.
What Is AI Share of Voice?
AI share of voice tracks how often a brand appears in AI-generated answers compared with competitors in the same category. It reflects visibility across AI platforms such as ChatGPT, Gemini, Claude, and Perplexity. Monitoring AI share of voice helps organizations see whether AI systems consistently include and recommend their brand for key queries or whether competitor brands are showing up more often.

About The Author

Mark Huntley

Mark Huntley

Founder and CEO

Mark Huntley, J.D. is founder of CiteWorks Studio, a strategic advisory focused on visibility, authority, and recommendation presence in AI-shaped search environments. His work centers on embedding-level GEO, vector optimization, and cosine gap engineering — helping brands align their digital presence with the retrieval systems that increasingly shape discovery, interpretation, and choice.

VIEW ALL CASE STUDIESREQUEST AN AI VISIBILITY AUDIT