Every product with at least 10 labels here, on both axes. The 30% line names a quadrant, not the verdict above: that one needs more than 40%.
“High BBB complaint volumes (Podium: D- rating, 200+ in 3 years) ... billing disputes, hard cancellations” Grok 4.1 Fast, negative prompt
“What to Avoid on a Budget - Birdeye ($299+/month) ... overkill for limited budgets” Kimi K2, budget prompt
“Frequent Trustpilot accusations of scams, undelivered services ... Users report it as "flawed" for reputation tracking.” Grok 4.1 Fast, negative prompt
“be wary of their premium packages, and treat the platform as just one of many signals rather than your primary focus” Gemini 3.5 Flash, negative prompt
22 of 72 answers in this category came back with a source list, from 4 of 6 models: citations where the model returns them, or the search results it consulted. 6 of those lists are Google grounding redirects that name no site and are left out of the counts. 206 links across 89 sites, every framing counted. Ranked by the number of answers carrying the site or page. The other 2 models searched through a gateway that returned no source list on this edition.
The ten pages named in the most answers, by full address. A page here is one the models returned with a recommendation, not one the index endorses.
| Prompt | Model | Version string | Time (UTC) | Searched | Sources | Latency |
|---|
Every judgment call made between the raw labels and the numbers above, listed so it is visible and reversible.