Zero of fourteen models named AWS Cost Explorer first on the direct prompt; zero named Microsoft Cost Management. AWS Cost Explorer was named by thirteen of the fourteen models and Microsoft Cost Management by thirteen and AWS Cost Explorer carries 38 labels and Microsoft Cost Management 30, so the shares are not directly comparable.
Named in two categories this edition.
Named in two categories this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all fourteen models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the cloud cost management page.
Across every category in the October 2026 Edition, AWS Cost Explorer and Microsoft Cost Management were named in the same answer eighty-nine times, of the 124 answers naming AWS Cost Explorer and the 94 naming Microsoft Cost Management. In those answers Microsoft Cost Management took the first choice zero times and AWS Cost Explorer eight.
| Model | Direct | Paraphrase | Comparative | Budget-constrained | Scale-constrained | Negative |
|---|---|---|---|---|---|---|
| Claude Haiku 4.5 | ||||||
| GPT-5.4 mini | ||||||
| Gemini 3.5 Flash | ||||||
| Perplexity Sonar | ||||||
| Grok 4.1 Fast | ||||||
| Mistral Small | ||||||
| DeepSeek V4 Flash | ||||||
| Llama 4 Maverick | ||||||
| Qwen 3.7 Flash | ||||||
| Kimi K2 | ||||||
| GLM 4.7 FlashX | ||||||
| MiniMax M2.5 | ||||||
| GPT-6 Luna | ||||||
| Muse Glimmer 30B |
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of seven in this category shown.
“work well for single-cloud environments at small scale but require perfect tagging discipline and lack multi-cloud or Kubernetes cluster visibility” Muse Glimmer 30B · negative prompt · soft negative
“independent reviews highlight significant data lag (up to 24 hours) in usage reporting and shallow drill-down capabilities” Mistral Small · negative prompt · soft negative
“Use them to start, but avoid relying on them if you run multi-cloud or container-heavy infrastructures.” Gemini 3.5 Flash · negative prompt · soft negative
“Best for a limited budget: start with your cloud provider's built-in cost tools... AWS Cost Explorer, Microsoft Cost Management, or Google Cloud Cost Management.” GPT-6 Luna · budget prompt · first choice
“Native cloud provider tools are free or nearly free... Zero spend / < $2.5K/mo cloud bill: Native Cost Explorer + Infracost for prevention, plus Vantage Starter” Muse Glimmer 30B · budget prompt · first choice
“If you have < $1,000/month in cloud spend: Stick to Native Tools (AWS/Azure/GCP). Do not pay for a third-party tool yet.” Qwen 3.7 Flash · budget prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of six in this category shown.
“work well for single-cloud environments at small scale but require perfect tagging discipline and lack multi-cloud or Kubernetes cluster visibility” Muse Glimmer 30B · negative prompt · soft negative
“the overall experience can be inconsistent, and the tool may not provide the depth of analysis required for complex environments” Mistral Small · negative prompt · soft negative
“stick to native tools (AWS Cost Explorer/Azure Cost Management)... native tools cannot provide a unified view” Qwen 3.7 Flash · negative prompt · soft negative
“start with your cloud provider's built-in cost tools... AWS Cost Explorer, Microsoft Cost Management, or Google Cloud Cost Management” GPT-6 Luna · budget prompt · first choice
“Under $50K/month: Start with native tools (AWS Cost Explorer, Azure Cost Management) — free and immediate” Claude Haiku 4.5 · budget prompt · first choice
“starting with your cloud provider\u2019s native tools (AWS Cost Explorer, Azure Cost Management)” DeepSeek V4 Flash · negative prompt · first choice
Comparisons are drawn for the top eight products in each category, each against each. The output is the models' output; nothing here is a recommendation by the index.