Seven of twelve models named LaunchDarkly first on the direct prompt; four named ConfigCat. Both were named by all twelve models and LaunchDarkly carries 61 labels and ConfigCat 36, so the shares are not directly comparable.
Named in one category this edition.
Named in one category this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all twelve models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the feature flags page.
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of eight in this category shown.
“Caution level: High (especially for new adopters or scale-ups)... Avoid if... Microservices, startups, regulated industries, cost-sensitive” DeepSeek V4 Flash · negative prompt · hard negative
“Avoid for tight budgets: LaunchDarkly/Split (free tiers cap quickly, then $10+/MAU or seats).” Grok 4.1 Fast · budget prompt · hard negative
“Avoid: LaunchDarkly... they are notoriously expensive for startups.” Qwen 3.7 Flash · budget prompt · hard negative
“LaunchDarkly: The market giant. Best-in-class for enterprise governance, automated APM rollbacks, and security—but typically the most expensive, using a strict MAU pricing model.” Gemini 3.5 Flash · scale prompt · first choice
“LaunchDarkly — The Enterprise Benchmark ... Choose LaunchDarkly if release governance, approvals, and auditability are your top priorities.” GLM 4.7 FlashX · comparative prompt · first choice
“The most complete feature management platform... LaunchDarkly suits enterprises needing governance and audit trails” Claude Haiku 4.5 · comparative prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of eight in this category shown.
“Platforms with significant limitations or red flags: - ConfigCat: ... its pricing model is criticized for being expensive as you scale” Mistral Small · negative prompt · hard negative
“Lighterweight option may lack advanced governance, stats, and enterprise features for regulated environments.” GLM 4.7 FlashX · negative prompt · soft negative
“ConfigCat - Operational Limitations - Limited A/B testing capabilities” Kimi K2 · negative prompt · soft negative
“the best default choice is usually ConfigCat if you want the simplest low-cost managed service” Perplexity Sonar · budget prompt · first choice
“The Best "All-Rounder": ConfigCat ... strikes the best balance between power and simplicity” Qwen 3.7 Flash · direct prompt · first choice
“ConfigCat is probably the best default choice for a mid-market B2B company.” GPT-5.4 mini · direct prompt · first choice
Comparisons are drawn for the top three products in each category. The output is the models' output; nothing here is a recommendation by the index.