Two of twelve models named StackAdapt first on the direct prompt; one named YouTube Ads. Both were named by all twelve models and StackAdapt carries 31 labels and YouTube Ads 41, so the shares are not directly comparable.
Named in nine categories this edition.
Named in two categories this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all twelve models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the video advertising page.
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
No label in this category carried a quote.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of eight in this category shown.
“Brand Safety: Content classification issues, especially in sensitive categories” GLM 4.7 FlashX · negative prompt · hard negative
“recommends avoiding influencer and UGC channels on YouTube due to brand safety risks... brands have been advised to be cautious when advertising on the platform” Llama 4 Maverick · negative prompt · soft negative
“YouTube (when using GVP/non-premium inventory) ... While YouTube itself has better controls, the Video Partners network that runs alongside it is the major risk” Mistral Small · negative prompt · soft negative
“The Overall Winner: YouTube Ads (via Google Ads)... widely considered the best video advertising platform for companies on a tight budget” Gemini 3.5 Flash · budget prompt · first choice
“YouTube/Google Ads | Brand awareness, wide reach | Largest global reach, diverse ad formats, strong targeting across Google network” MiniMax M2.5 · comparative prompt · first choice
“Best for: Everything — awareness, consideration, and conversion on a single buy. Performance brands use it as their video anchor.” DeepSeek V4 Flash · comparative prompt · first choice
Comparisons are drawn for the top three products in each category. The output is the models' output; nothing here is a recommendation by the index.