Five of fourteen models named CallTrackingMetrics first on the direct prompt; one named WhatConverts. CallTrackingMetrics was named by fourteen of the fourteen models and WhatConverts by thirteen and CallTrackingMetrics carries 49 labels and WhatConverts 40, so the shares are not directly comparable.
Named in one category this edition.
Named in two categories this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all fourteen models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the call tracking platforms page.
Across every category in the October 2026 Edition, CallTrackingMetrics and WhatConverts were named in the same answer sixty-three times, of the 138 answers naming CallTrackingMetrics and the 101 naming WhatConverts. In those answers WhatConverts took the first choice eight times and CallTrackingMetrics five.
| Model | Direct | Paraphrase | Comparative | Budget-constrained | Scale-constrained | Negative |
|---|---|---|---|---|---|---|
| Claude Haiku 4.5 | ||||||
| GPT-5.4 mini | ||||||
| Gemini 3.5 Flash | ||||||
| Perplexity Sonar | ||||||
| Grok 4.1 Fast | ||||||
| Mistral Small | ||||||
| DeepSeek V4 Flash | ||||||
| Llama 4 Maverick | ||||||
| Qwen 3.7 Flash | ||||||
| Kimi K2 | ||||||
| GLM 4.7 FlashX | ||||||
| MiniMax M2.5 | ||||||
| GPT-6 Luna | ||||||
| Muse Glimmer 30B |
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of eight in this category shown.
“High risk of billing surprises and support frustration... Powerful but risky. Only consider if you have a dedicated project manager and a legal review of their contract.” DeepSeek V4 Flash · negative prompt · hard negative
“you should approach with caution or consider avoiding due to common complaints and limitations” Mistral Small · negative prompt · hard negative
“Platforms like CallTrackingMetrics (CTM) and Invoca are incredibly powerful, but they are built for enterprise call centers... avoid these enterprise-level tools” Gemini 3.5 Flash · negative prompt · soft negative
“CallTrackingMetrics (CTM) is recommended for mid-market to enterprise companies with contact centers, high call volume, or complex routing” Claude Haiku 4.5 · direct prompt · first choice
“best for mid-market agencies and contact centers ... sits between CallRail's SMB focus and Invoca's enterprise depth” Muse Glimmer 30B · direct prompt · first choice
“Platforms like Invoca and CallTrackingMetrics are built specifically for enterprise scale” Kimi K2 · scale prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of four in this category shown.
“I'd recommend starting with WhatConverts at $30/month or Nimbata at $35/month” Kimi K2 · budget prompt · first choice
“Small businesses: CallRail or WhatConverts (good balance of features and cost)” MiniMax M2.5 · comparative prompt · first choice
“WhatConverts — simpler pricing, strong multi-channel attribution” Kimi K2 · negative prompt · first choice
Comparisons are drawn for the top eight products in each category, each against each. The output is the models' output; nothing here is a recommendation by the index.