Three of fourteen models named Microsoft Defender for Office 365 first on the direct prompt; one named Barracuda Email Protection. Both were named by all fourteen models and Microsoft Defender for Office 365 carries 51 labels and Barracuda Email Protection 31, so the shares are not directly comparable.
Named in three categories this edition.
Named in three categories this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all fourteen models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the email security page.
Across every category in the October 2026 Edition, Microsoft Defender for Office 365 and Barracuda Email Protection were named in the same answer sixty-three times, of the 176 answers naming Microsoft Defender for Office 365 and the 94 naming Barracuda Email Protection. In those answers Barracuda Email Protection took the first choice four times and Microsoft Defender for Office 365 twenty-one.
| Model | Direct | Paraphrase | Comparative | Budget-constrained | Scale-constrained | Negative |
|---|---|---|---|---|---|---|
| Claude Haiku 4.5 | ||||||
| GPT-5.4 mini | ||||||
| Gemini 3.5 Flash | ||||||
| Perplexity Sonar | ||||||
| Grok 4.1 Fast | ||||||
| Mistral Small | ||||||
| DeepSeek V4 Flash | ||||||
| Llama 4 Maverick | ||||||
| Qwen 3.7 Flash | ||||||
| Kimi K2 | ||||||
| GLM 4.7 FlashX | ||||||
| MiniMax M2.5 | ||||||
| GPT-6 Luna | ||||||
| Muse Glimmer 30B |
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of seven in this category shown.
“Microsoft's published comparison, for example, uses different "missed threat" criteria for Defender than for competing gateways, so it shouldn't be treated as a neutral head-to-head ranking.” GPT-6 Luna · negative prompt · soft negative
“strong execution but limited vision for evolving threats. Complaints include high false positives, poor internal email scanning” Grok 4.1 Fast · negative prompt · soft negative
“relying *solely* on native cloud email security is a risk for organizations with high threat profiles” Gemini 3.5 Flash · negative prompt · soft negative
“Microsoft Defender for Office 365 is the default choice for most Microsoft-365 shops—it's deeply integrated, cost-effective, and provides a solid baseline.” GLM 4.7 FlashX · comparative prompt · first choice
“Start with what you already have — if you're on Microsoft 365 E5, check if Defender for O365 Plan 2 is already included before paying for a new platform.” DeepSeek V4 Flash · scale prompt · first choice
“Microsoft Defender for Office 365 is the most pragmatic starting point—low cost, deep integration, and solid baseline protection.” GLM 4.7 FlashX · direct prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of seven in this category shown.
“Barracuda Email Protection (High Caution / Avoid)” GLM 4.7 FlashX · negative prompt · hard negative
“Barracuda disclosed an RCE vulnerability affecting its Email Security Gateway appliances... not, by itself, to reject all Barracuda products.” GPT-6 Luna · negative prompt · soft negative
“Caveat: Some users report significant price increases at renewal and forced migrations to bundled products” Kimi K2 · direct prompt · soft negative
“Barracuda and SpamTitan are consistently highlighted as the most affordable and accessible options for budget-conscious companies.” Claude Haiku 4.5 · budget prompt · first choice
“"the platform of choice for budget-conscious SMBs and lean IT departments"... Start with SpamTitan or Barracuda” MiniMax M2.5 · budget prompt · first choice
“I'd recommend Barracuda as the strongest all-around choice for phishing and email threat protection.” Perplexity Sonar · paraphrase prompt · first choice
Comparisons are drawn for the top eight products in each category, each against each. The output is the models' output; nothing here is a recommendation by the index.