Eleven of fourteen models named ChurnZero first on the direct prompt; one named Planhat. Both were named by all fourteen models and ChurnZero carries 64 labels and Planhat 40, so the shares are not directly comparable.
Named in three categories this edition.
Named in two categories this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all fourteen models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the customer success platforms page.
Across every category in the October 2026 Edition, ChurnZero and Planhat were named in the same answer 133 times, of the 227 answers naming ChurnZero and the 164 naming Planhat. In those answers Planhat took the first choice four times and ChurnZero eleven.
| Model | Direct | Paraphrase | Comparative | Budget-constrained | Scale-constrained | Negative |
|---|---|---|---|---|---|---|
| Claude Haiku 4.5 | ||||||
| GPT-5.4 mini | ||||||
| Gemini 3.5 Flash | ||||||
| Perplexity Sonar | ||||||
| Grok 4.1 Fast | ||||||
| Mistral Small | ||||||
| DeepSeek V4 Flash | ||||||
| Llama 4 Maverick | ||||||
| Qwen 3.7 Flash | ||||||
| Kimi K2 | ||||||
| GLM 4.7 FlashX | ||||||
| MiniMax M2.5 | ||||||
| GPT-6 Luna | ||||||
| Muse Glimmer 30B |
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of seven in this category shown.
“Avoid enterprise platforms like Gainsight or ChurnZero initially” Kimi K2 · budget prompt · hard negative
“Pricing Escalation... Declining Support Quality... Clunky UI” Kimi K2 · negative prompt · hard negative
“Don't assume it will be plug-and-play... admins handling implementation, data, and advanced features may need extra time” GPT-6 Luna · negative prompt · soft negative
“I'd start with ChurnZero. It's usually the best balance of speed, usability, and core retention functionality for mid-market teams.” GPT-5.4 mini · paraphrase prompt · first choice
“ChurnZero is the common first choice for health scoring, real-time usage-based risk alerts and automated retention playbooks” Muse Glimmer 30B · paraphrase prompt · first choice
“ChurnZero – the strongest mid-market generalist. Repeatedly called out as the strongest mid-market option in the category.” Muse Glimmer 30B · direct prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of seven in this category shown.
“What to Skip for Mid-Market - Gainsight and Planhat — powerful, but built for enterprise” DeepSeek V4 Flash · direct prompt · hard negative
“Enterprise giants like Gainsight, ChurnZero, or Planhat are highly capable but typically start at $15,000 to $50,000+ per year” Gemini 3.5 Flash · budget prompt · soft negative
“Some platforms use account-based pricing, which can become very expensive if churned accounts still count toward your total.” Mistral Small · negative prompt · soft negative
“best all-around customer success platform is usually ChurnZero or Planhat” GPT-5.4 mini · direct prompt · first choice
“I'd shortlist Vitally, ChurnZero, and Planhat first” DeepSeek V4 Flash · scale prompt · first choice
“Planhat if you want CS, renewals, customer projects, and possibly sales workflows brought together” GPT-6 Luna · paraphrase prompt · alternative
Comparisons are drawn for the top eight products in each category, each against each. The output is the models' output; nothing here is a recommendation by the index.