CiteWorks Studio

How AI Search Is Recommending Customer Service Software: Monthly Trends

Mark HuntleyBy Mark HuntleyFounder and CEO
11 minutes read

Key Takeaways

  • Freshdesk led the category at 56.1% valid recommendation coverage, 5.5 points ahead of Zendesk Chat.
  • Zendesk Chat entered the tracked set at 50.6% coverage and posted the highest rank-one rate at 25.6%.
  • Intercom was the strongest riser among continuously tracked brands, gaining 12.3 points to reach 44.7% coverage.
  • September brand changes affected comparisons: Zendesk, HubSpot Service Hub, and Zoho Desk were replaced by Zendesk Chat, HubSpot Live Chat, and Zoho Inventory.

Executive Summary

The tracked brand set in the customer service software benchmark changed materially in September 2026. Three full-suite platforms — Zendesk, HubSpot Service Hub, and Zoho Desk — were replaced in the tracked comparison by Zendesk Chat, HubSpot Live Chat, and Zoho Inventory. That is a change in which brands were tracked, not a measured decline in AI recommendation standing for the brands that were removed.

Freshdesk retains the category lead at 56.1% valid recommendation coverage, up from 52.6% in July 2026 and 49.1% in August 2026, and now sits 5.5 points ahead of second-place Zendesk Chat, which entered tracking this month at 50.6% coverage.

The strongest upward mover was Zendesk Chat, which posted 50.6% valid recommendation coverage in its first tracked month, reaching the top rank in 92 of 360 qualified observations (25.6% rank-one rate). Intercom was the strongest riser among brands tracked in both July and September, climbing 12.3 points from 32.4% to 44.7%, a significant move driven by a 19.9-point gain in raw mention presence to 69.4%.

Front was the sharpest decliner among continuously tracked brands, down 6.6 points from 17.7% in July to 11.1% in September, a significant move on the baseline-to-current comparison. Among the seven brands tracked continuously since July, six gained coverage and only Front declined.

Each monthly run begins with a full prompt-surface collection before qualification. In July 2026, the run began with 800 prompt-surface observations (484 unique questions), of which 800 mentioned a tracked brand or competitor; 709 were relevant and 91 were irrelevant, leaving 481 qualified observations. In September 2026, the run began with 800 prompt-surface observations (605 unique questions), of which 800 mentioned a tracked brand or competitor; 563 were relevant and 237 were irrelevant, leaving 360 qualified observations reported below.

AI recommendation trend

valid recommendation coverage, Jul 2026 to Sep 2026

0%20%40%60%80%Jul 2026Aug 2026Sep 2026
  • Freshdesk56.1%
  • Zendesk Chat50.6%
  • Intercom44.7%
  • Help Scout38.1%
  • Salesforce Service Cloud28.3%
  • HubSpot Live Chat25.3%
  • Gorgias25.0%
  • Front11.1%
  • Zoho Inventory2.2%
  • Gladly0.6%
  • HubSpot Service Hub0.0%
  • Zendesk0.0%
  • Zoho Desk0.0%

Key Findings

Signal

September 2026 finding

Category leader

Freshdesk leads at 56.1% valid recommendation coverage (202 of 360 observations)

Leader gap

Freshdesk leads Zendesk Chat by 5.5 points, with Zendesk Chat at 50.6% coverage in its first tracked month

Largest riser

Zendesk Chat enters at 50.6% coverage with a 25.6% rank-one rate, the highest in the category

Strongest riser among continuous brands

Intercom up 12.3 points from 32.4% (July) to 44.7% valid recommendation coverage, a significant move

Largest decliner among continuous brands

Front down 6.6 points from 17.7% (July) to 11.1% valid recommendation coverage, a significant move

Brand set rotation

Zendesk, HubSpot Service Hub, and Zoho Desk dropped from tracking; Zendesk Chat, HubSpot Live Chat, and Zoho Inventory entered

Benchmark Context

The report separates the raw collection universe from the qualified analysis set. Brand-level recommendation percentages are calculated within the qualified benchmark set.

Research stage

Jul 2026

Sep 2026

What it represents

Source prompt-surface observations collected

800

800

Raw prompt-surface observations across the defined AI/search surface universe

Unique questions

484

605

Distinct questions after de-duplication

Brand / competitor mentions

800

800

Prompts mentioning a tracked brand or competitor

Relevant prompts

709

563

Prompts relevant to the category

Irrelevant prompts

91

237

Prompts set aside as irrelevant

Qualified benchmark observations

481

360

Public denominator after both qualification stages

Qualified surface breadth

6

6

Canonical AI surface families with at least one qualified observation

The funnel above shows how the raw collection narrows to the qualified analysis set used for every brand-level percentage below.

Benchmark-Level Metrics

Metric

Jul 2026

Sep 2026

Change

Qualified observations

481

360

Down 121

Companies tracked

10

10

Flat

Recommendation-shaped answer share

34.1%

31.1%

Down 3.0 points

Valid recommendation shortlist share

62.8%

62.5%

Down 0.3 points

Category leader by coverage

Zendesk

Freshdesk

Changed

The qualified set narrowed from 481 observations in July 2026 to 360 in September 2026, with August 2026 at 342 as an intermediate month. The tracked brand composition also changed: each month tracks 10 companies, but three of the tracked identities changed — Zendesk, HubSpot Service Hub, and Zoho Desk were replaced by Zendesk Chat, HubSpot Live Chat, and Zoho Inventory — while the other seven brands were tracked continuously. The recommendation-shaped answer share fell from 34.1% to 31.1%, while the valid recommendation shortlist share held nearly flat at 62.5%.

AI Recommendation Trend

Chat-Specific Offerings Reshape the Leadership Table

Brand

Jul 2026

Sep 2026

Movement

Sep 2026 rank

Freshdesk

52.6%

56.1%

Up 3.5 points

1st

Zendesk Chat

0.0%

50.6%

Up 50.6 points

2nd

Intercom

32.4%

44.7%

Up 12.3 points

3rd

Help Scout

35.3%

38.1%

Up 2.8 points

4th

Salesforce Service Cloud

26.2%

28.3%

Up 2.1 points

5th

HubSpot Live Chat

0.0%

25.3%

Up 25.3 points

6th

Gorgias

20.2%

25.0%

Up 4.8 points

7th

Front

17.7%

11.1%

Down 6.6 points

8th

Zoho Inventory

0.0%

2.2%

Up 2.2 points

9th

Gladly

0.2%

0.6%

Up 0.4 points

10th

HubSpot Service Hub

38.5%

0.0%

Down 38.5 points

11th

Zendesk

61.1%

0.0%

Down 61.1 points

12th

Zoho Desk

42.4%

0.0%

Down 42.4 points

13th

The zero-percent rows for Zendesk, HubSpot Service Hub, and Zoho Desk reflect their removal from the September tracked set, not a measured drop in AI recommendation standing; their replacements, Zendesk Chat, HubSpot Live Chat, and Zoho Inventory, entered with 50.6%, 25.3%, and 2.2% coverage respectively. Among the seven brands tracked continuously since July, six gained coverage and only Front declined.

What Changed This Month

Freshdesk

Freshdesk's valid recommendation coverage rose 3.5 points from 52.6% in July 2026 to 56.1% in September 2026, a move within normal month-to-month variation but enough to keep the brand in the category lead. The brand was the most recommended in the set with 202 valid recommendations out of 360 qualified observations.

Mention presence rose 5.0 points to 85.3%, while top-three rate slipped 5.2 points from 40.8% to 35.6%. Rank-one rate rose slightly to 6.9% (25 of 360 observations) from 6.2% in July, and net sentiment stayed flat at 0.8.

Freshdesk is the only brand with both leading coverage and near-universal presence; its strength lies in breadth across AI answers.

Highest-priority diagnostic: which prompt themes are driving the coverage gain, and whether the top-three rate decline signals weaker positioning in shortlist responses.

Zendesk Chat

Zendesk Chat entered the tracked set in September 2026 with 50.6% valid recommendation coverage, the second-highest in the category. The brand appeared in 73.6% of qualified observations with 182 valid recommendations.

The brand posted a 35.0% top-three rate and a 25.6% rank-one rate, the highest rank-one share in the category, with 92 of 360 observations placing it first.

The entry of Zendesk Chat alongside Zendesk's removal from the tracked set means the full-suite platform and its chat product are not being compared in the same table; the rank-one strength is notable within the current tracked universe.

Highest-priority diagnostic: which prompt types are driving the rank-one placements, and whether they overlap with the prompts where Zendesk previously appeared.

Intercom

Intercom rose 12.3 points from 32.4% in July 2026 to 44.7% in September 2026, a significant move and the strongest among brands tracked continuously across the quarter. Valid recommendations reached 161 of 360 observations, up from 156 of 481 in July.

Rank-one placement rose 3.3 points from 2.3% to 5.6% (20 of 360 observations), and top-three rate climbed 8.4 points to 21.7% (78 of 360 observations). Net sentiment held at 0.8.

The gain is driven by a 19.9-point presence increase to 69.4%, meaning more AI answers now name Intercom and more of those answers place it in a top-three position.

Highest-priority diagnostic: which AI surfaces or prompt categories drove the presence jump, and whether the top-three gains are concentrated in shortlist-style answers.

Front

Front declined 6.6 points from 17.7% in July 2026 to 11.1% in September 2026, a significant move and the sharpest among continuous brands. Valid recommendations fell from 85 of 481 observations in July to 40 of 360 in September.

The decline bottomed in August at 5.9%, then Front recovered 5.2 points in September, a significant prior-month move; even with that recovery, the brand remains well below its July level. Top-three rate fell 4.1 points from 6.9% to 2.8% (10 of 360 observations), with rank-one rate flat at 0.6% (2 observations).

Front's mention presence was 21.1% in September, close to its July level of 22.7%, meaning the brand appears in answers more often than it is recommended within them.

Highest-priority diagnostic: what AI systems say when Front is mentioned but not recommended, and which prompt themes formerly produced Front recommendations.

Help Scout

Help Scout rose 2.8 points from 35.3% in July 2026 to 38.1% in September 2026, within normal month-to-month variation, with 137 valid recommendations out of 360 observations.

The brand had dropped to 22.8% in August 2026, then recovered 15.3 points in September, a significant prior-month move. Raw mention presence rose 11.7 points to 58.1%, a significant increase.

Help Scout's recovery is broad-based on presence, but top-three rate fell 4.9 points to 9.4%, meaning the brand is more visible yet placed lower when recommended.

Highest-priority diagnostic: what drove the August decline and the September recovery, and whether the top-three rate erosion reflects a positioning shift.

Gorgias

Gorgias rose 4.8 points from 20.2% in July 2026 to 25.0% in September 2026, within normal variation, with 90 valid recommendations out of 360 observations.

Raw mention presence rose 7.2 points to 33.6% (121 mentions), a significant gain, and top-three rate climbed 3.2 points to 6.1% (22 observations). Rank-one rate nearly tripled from 0.2% to 0.6% (2 observations).

The coverage gain was driven by a sharp September rebound: Gorgias had fallen to 14.9% in August, then rose 10.1 points in September, a significant prior-month move.

Highest-priority diagnostic: which prompt themes or surfaces drove the September recovery, and whether the gains extend beyond ecommerce-focused queries.

Salesforce Service Cloud

Salesforce Service Cloud rose 2.1 points from 26.2% in July 2026 to 28.3% in September 2026, within normal month-to-month variation, with 102 valid recommendations out of 360 observations.

The brand dropped to 19.6% in August 2026, then recovered 8.7 points in September, a significant prior-month move. Net sentiment improved from 0.7 to 0.8.

Rank-one rate fell 1.3 points from 2.7% to 1.4% (5 of 360 observations), even as overall coverage recovered, indicating a position-quality challenge.

Highest-priority diagnostic: which prompts recovered in September, and why rank-one placements did not recover at the same pace.

HubSpot Live Chat

HubSpot Live Chat entered the tracked set in September 2026 with 25.3% valid recommendation coverage, with 91 valid recommendations out of 360 observations. The brand appeared in 35.6% of qualified observations.

Top-three rate reached 6.4% (23 of 360 observations) with rank-one at 1.1% (4 observations). Net sentiment was 0.8.

HubSpot Live Chat's entry replaces HubSpot Service Hub in the tracked set, meaning the two HubSpot products are not directly compared in the same table.

Highest-priority diagnostic: which prompt categories drive the brand's recommendations, and how the live chat focus maps to buyer-intent queries.

Zoho Inventory

Zoho Inventory entered tracking in September 2026 with 2.2% valid recommendation coverage, a small but significant move given its tiny base. Valid recommendations reached 8 of 360 observations.

The brand held 2.8% raw mention presence, with one rank-one placement and one top-three placement. Net sentiment was 0.8.

Zoho Inventory is adjacent to the customer service category rather than a direct service platform, so its small coverage base is expected for the prompt universe.

Highest-priority diagnostic: whether the brand appears in customer-service-adjacent prompts or only in inventory-focused queries.

Gladly

Gladly rose 0.4 points from 0.2% in July 2026 to 0.6% in September 2026, within normal month-to-month variation for its very small base. Valid recommendations totaled 2 of 360 observations.

Mention presence rose 0.9 points to 1.9% (7 mentions), and the brand earned its first top-three placement at 0.3% (1 of 360 observations). Net sentiment improved from 0.4 to 0.6.

Gladly's movement is driven by a handful of observations; percentage changes on such a small denominator should be read cautiously.

Highest-priority diagnostic: whether the gains reflect durable presence improvements or small-count variation.

Buyer-Intent Interpretation

Buyer-intent cluster

What it captures

Strategic question

Brand Recommendation

Prompts asking which customer service software to use or choose

Which brand does the AI actually recommend, and in what position?

Pricing & Value

Prompts asking about cost, pricing tiers, or value for money

How do AI systems discuss price and value, and which brands benefit?

Multi-Brand Comparison

Prompts asking to compare two or more brands head to head

Which brand wins the comparison, and on what attributes?

In September 2026, all 360 qualified observations fell into the Brand Recommendation cluster. The response mix included recommendation shortlists, comparison analyses, factual answers, and ranked lists, but none of those responses were classified into dedicated pricing or multi-brand comparison clusters.

The public benchmark therefore answers "which brand does AI recommend" but cannot yet answer "how does AI discuss price and value" or "which brand wins a head-to-head comparison." Those commercial questions remain open for company-level analysis.

Brand Opportunity Summary

Brand

Sep 2026 coverage

Current signal

Highest-priority diagnostic

Freshdesk

56.1%

Category leader with broad presence

Which prompt themes drive the lead, and why top-three rate slipped

Zendesk Chat

50.6%

New entrant with strongest rank-one rate

Which prompt types drive rank-one placements

Intercom

44.7%

Strongest riser among continuous brands

Which surfaces or prompts drove the presence jump

Help Scout

38.1%

Recovered presence with weaker top-three rate

What drove the August decline and September recovery

Salesforce Service Cloud

28.3%

Recovered coverage with flat rank-one

Which prompts recovered, and why rank-one lagged

HubSpot Live Chat

25.3%

New entrant with moderate coverage

Which buyer-intent prompts generate recommendations

Gorgias

25.0%

Rebound after August dip

Which surfaces drove the recovery beyond ecommerce

Front

11.1%

Coverage decline despite stable presence

What AI says when Front is mentioned but not recommended

Zoho Inventory

2.2%

Small adjacent-category presence

Whether recommendations are service-adjacent or inventory-only

Gladly

0.6%

Tiny base with first top-three placement

Whether gains reflect durable presence or small-count noise

HubSpot Service Hub

0.0%

Dropped from September tracking

Not applicable in current tracked set

Zendesk

0.0%

Dropped from September tracking

Not applicable in current tracked set

Zoho Desk

0.0%

Dropped from September tracking

Not applicable in current tracked set

The benchmark identifies where attention is warranted across the tracked set; a company-level analysis is needed to explain why specific brands gain or lose recommendation coverage.

Evidence Behind the Benchmark

The aggregate metrics are built from prompt-level observations (query, surface, recommendation outcome, rank, sentiment, and citations where exposed). Company-level analysis can go deeper into prompt, competitor, surface, and evidence patterns. Source presence is not automatically treated as proof of causation.

About This Benchmark

This report is part of the LLM Authority Index AI Market Discovery research program.

Report-Specific Interpretation Notes

  • Small-count movement: Gladly's coverage of 0.6% is based on 2 valid recommendations out of 360 qualified observations; percentage movement on such a small base should be interpreted with caution.
  • Brand-set rotation: September 2026 replaced three tracked brands (Zendesk, HubSpot Service Hub, Zoho Desk) with chat- and inventory-focused offerings (Zendesk Chat, HubSpot Live Chat, Zoho Inventory); coverage of 0.0% for the departed brands reflects the tracking change, not a measured decline in AI recommendation standing.
  • Qualified denominator: all percentages are calculated against the qualified set (360 observations in September 2026, versus 481 in July 2026), not the raw prompt collection.
  • Directional analysis: month-over-month movement identifies changes worth investigating; it does not by itself establish the cause of those changes.

Next Step

The Public Benchmark Shows Where a Brand Is Winning or Losing. A Company-Level Audit Shows Why.

The aggregate percentages in this report leave the important questions open. Which high-intent prompts is a brand actually winning? When a brand loses the recommendation, which competitor takes the slot? What attributes do AI systems associate with each option, and which external sources shape those answers? A category-wide benchmark cannot answer those questions from percentages alone.

A company-specific AI visibility audit maps those prompt, surface, competitor, ranking, sentiment, and evidence-source patterns into a prioritized visibility strategy. It identifies the specific prompts where a brand is vulnerable, the evidence sources that drive AI recommendations, and the surfaces where presence does not convert to placement.

Request an AI visibility audit

/ Take the next step

Want to Understand Your AI Citation Footprint?

We start every engagement with a full audit of how AI systems reference your brand today.

Measurable, Repeatable Programme

Build a durable foundation of credible citations that compounds over time and continues to influence AI answers as new queries emerge

Citation Architecture Review

Identify which high-authority community sources are and aren't working in your favour across AI platforms.

AI Visibility Audit

Understand exactly how LLMs are referencing your brand today and which sources are shaping those answers.

/ Learn More

Understanding AI search visibility.

AI search experiences create answers by pulling information from many places online and summarizing it into a single response.

What Is AI Citation Intelligence?
AI citation intelligence is the process of measuring where AI platforms source their information and how frequently a brand is mentioned or referenced in AI-generated responses. Because LLMs synthesize across multiple sources, the sites and brands that appear repeatedly tend to influence how a topic or company is framed. This practice focuses on identifying which sources shape AI outputs and tracking brand visibility across different AI systems.
What Is Citation Architecture?
Citation architecture describes the set of sources that consistently inform how AI systems talk about a brand, product, or topic. LLMs draw from websites, articles, forums, and public discussion, and the sources they rely on most often become the backbone of their answers. Building strong citation architecture means ensuring that accurate, credible, high authority sources are the ones most likely to shape the way AI tools summarize and recommend a brand.
What Is Generative Engine Optimization?
Generative engine optimization (GEO) is the practice of improving the chances that AI systems use and cite your brand or content when generating answers. While traditional SEO is centered on ranking pages in search results, GEO focuses on how LLMs retrieve, interpret, and combine information when responding to a question. The objective is to strengthen the content and sources AI systems rely on, so your brand is treated as a trusted reference in AI responses.
What Is AI Share of Voice?
AI share of voice tracks how often a brand appears in AI-generated answers compared with competitors in the same category. It reflects visibility across AI platforms such as ChatGPT, Gemini, Claude, and Perplexity. Monitoring AI share of voice helps organizations see whether AI systems consistently include and recommend their brand for key queries or whether competitor brands are showing up more often.

About The Author

Mark Huntley

Mark Huntley

Founder and CEO

Mark Huntley, J.D. is founder of CiteWorks Studio, a strategic advisory focused on visibility, authority, and recommendation presence in AI-shaped search environments. His work centers on embedding-level GEO, vector optimization, and cosine gap engineering — helping brands align their digital presence with the retrieval systems that increasingly shape discovery, interpretation, and choice.

VIEW ALL CASE STUDIESREQUEST AN AI VISIBILITY AUDIT