Seven of twelve models named Wiz first on the direct prompt; three named Orca Security. Both were named by all twelve models and Wiz carries 55 labels and Orca Security 40, so the shares are not directly comparable.
Named in six categories this edition.
Named in five categories this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all twelve models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the cloud-native application protection page.
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of three in this category shown.
“Orca/Wiz | $7K+/mo (small) | Agentless full CNAPP | Avoid for tight budgets” Grok 4.1 Fast · budget prompt · hard negative
“Extremely popular, pioneered the agentless graph-based approach, and offers great UX. It is incredibly easy to use but tends to be on the pricier end” Gemini 3.5 Flash · scale prompt · first choice
“Market leader by adoption ... Choose Wiz or Orca if you need fast, comprehensive visibility across multi-cloud with minimal operational overhead” Kimi K2 · comparative prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of three in this category shown.
“often out of reach because they typically require high annual contract minimums” Gemini 3.5 Flash · budget prompt · soft negative
“Very strong on data security posture management (DSPM) and attack path analysis, often highly competitive on pricing for mid-market companies.” Gemini 3.5 Flash · scale prompt · first choice
“Choose Wiz or Orca Security if you have a multi-cloud footprint, want immediate visibility without installing agents” Gemini 3.5 Flash · comparative prompt · first choice
Comparisons are drawn for the top three products in each category. The output is the models' output; nothing here is a recommendation by the index.