Zero of twelve models named Wizer first on the direct prompt; four named Hoxhunt. Wizer was named by nine of the twelve models and Hoxhunt by twelve and Wizer carries 11 labels and Hoxhunt 41, so the shares are not directly comparable.
Named in one category this edition.
Named in one category this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all twelve models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the security awareness training page.
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of four in this category shown.
“If you want something incredibly easy to deploy, highly effective, and very budget-friendly, Wizer is a top recommendation.” Gemini 3.5 Flash · paraphrase prompt · first choice
“I'd start with Wizer's free tier if you truly have zero budget, or Curricula if you can spend a modest amount” DeepSeek V4 Flash · budget prompt · first choice
“Wizer is frequently cited as the best choice for companies that want to start without paying immediately” Qwen 3.7 Flash · budget prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of four in this category shown.
“Highest user rating (4.8/5 on G2, 3,685+ reviews), adaptive phishing simulations, strong engagement, and measurable behavior change” Mistral Small · comparative prompt · first choice
“Hoxhunt is the best default pick for a mid-market B2B company that wants modern, behavior-focused phishing awareness training” Perplexity Sonar · direct prompt · first choice
“Shortlist Hoxhunt instead if your leadership specifically cares about *measurable phishing-resistance and reduced human risk*” DeepSeek V4 Flash · direct prompt · first choice
Comparisons are drawn for the top three products in each category. The output is the models' output; nothing here is a recommendation by the index.