Two of twelve models named Artisan first on the direct prompt; one named Reply.io. Artisan was named by twelve of the twelve models and Reply.io by eleven and Artisan carries 36 labels and Reply.io 18, so the shares are not directly comparable.
Named in one category this edition.
San Jose, United States, founded 2014. Named in three categories this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all twelve models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the AI SDR agents page.
Across every category in the September 2026 Edition, Artisan and Reply.io were named in the same answer twenty-three times, of the 102 answers naming Artisan and the 94 naming Reply.io. In those answers Reply.io took the first choice two times and Artisan three.
| Model | Direct | Paraphrase | Comparative | Budget-constrained | Scale-constrained | Negative |
|---|---|---|---|---|---|---|
| Claude Haiku 4.5 | ||||||
| GPT-5.4 mini | ||||||
| Gemini 3.5 Flash | ||||||
| Perplexity Sonar | ||||||
| Grok 4.1 Fast | ||||||
| Mistral Small | ||||||
| DeepSeek V4 Flash | ||||||
| Llama 4 Maverick | ||||||
| Qwen 3.7 Flash | ||||||
| Kimi K2 | ||||||
| GLM 4.7 FlashX | ||||||
| MiniMax M2.5 |
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of eight in this category shown.
“Avoid fully autonomous AI SDRs with annual lock-in contracts, no trials, and documented quality problems (Artisan, 11x.ai especially)” DeepSeek V4 Flash · negative prompt · hard negative
“Artisan - LinkedIn restricted their automated outreach in early 2026... G2 rating of 3.8/5 (lowest among reviewed platforms)” MiniMax M2.5 · negative prompt · hard negative
“Artisan was temporarily banned by LinkedIn in late 2025/early 2026 due to platform compliance and data-scraping violations.” Gemini 3.5 Flash · negative prompt · hard negative
“unless your team specifically wants an AI SDR that books meetings more autonomously, in which case Artisan is the better fit” Perplexity Sonar · paraphrase prompt · first choice
“Artisan"s AI agent, Ava, is highly regarded for mid-market companies that want a fully autonomous out-of-the-box solution.” Gemini 3.5 Flash · direct prompt · first choice
“or Artisan/Ava if you want more autonomous outbound prospecting with built-in contact data.” Claude Haiku 4.5 · paraphrase prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of three in this category shown.
“though the cited evaluation ranks it far below the leaders on total capability” Perplexity Sonar · direct prompt · soft negative
“Best Overall for Mid-Market ... Start with Reply.io (Jason AI) or Amplemarket Duo for mid-market B2B.” Kimi K2 · direct prompt · first choice
“Reply.io's paid plans start at $59/month for email sequences with Jason AI available on higher tiers, making it the most accessible entry point” Claude Haiku 4.5 · budget prompt · alternative
Comparisons are drawn for the top eight products in each category, each against each. The output is the models' output; nothing here is a recommendation by the index.