One of fourteen models named IBM Apptio first on the direct prompt; one named ServiceNow IT Financial Management. IBM Apptio was named by ten of the fourteen models and ServiceNow IT Financial Management by twelve and IBM Apptio carries 29 labels and ServiceNow IT Financial Management 33, so the shares are not directly comparable.
Named in one category this edition.
Named in one category this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all fourteen models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the IT financial management page.
Across every category in the October 2026 Edition, IBM Apptio and ServiceNow IT Financial Management were named in the same answer fifty-three times, of the 82 answers naming IBM Apptio and the 97 naming ServiceNow IT Financial Management. In those answers ServiceNow IT Financial Management took the first choice one time and IBM Apptio twenty-six.
| Model | DirectIA | ParaphraseIA | ComparativeIA | Budget-constrainedIA | Scale-constrainedIA | NegativeIA |
|---|---|---|---|---|---|---|
| Claude Haiku 4.5 | ||||||
| GPT-5.4 mini | IA | IA | IA | IA | ||
| Gemini 3.5 Flash | IA | IA | IA | IA | IA | IA |
| Perplexity Sonar | ||||||
| Grok 4.1 Fast | IA | |||||
| Mistral Small | IA | |||||
| DeepSeek V4 Flash | IA | IA | IA | IA | IA | |
| Llama 4 Maverick | ||||||
| Qwen 3.7 Flash | ||||||
| Kimi K2 | IA | IA | ||||
| GLM 4.7 FlashX | IA | IA | ||||
| MiniMax M2.5 | IA | IA | ||||
| GPT-6 Luna | IA | IA | IA | IA | ||
| Muse Glimmer 30B | IA |
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of eight in this category shown.
“What to Avoid at Mid-Size Scale... priced from $155K–$200K+ annually — designed for large enterprises. Overkill” DeepSeek V4 Flash · paraphrase prompt · hard negative
“Avoid enterprise tools like IBM Apptio, CloudZero, or Cloudability on a limited budget” DeepSeek V4 Flash · budget prompt · hard negative
“enterprise giants like IBM Apptio or ServiceNow are the gold standards for Fortune 500 companies, they are often too complex, require massive consulting budgets” Gemini 3.5 Flash · paraphrase prompt · soft negative
“Apptio is the strongest fit when IT and Finance need one maintained IT cost model for planning, showback, and variance workflows.” Muse Glimmer 30B · comparative prompt · first choice
“Choose Apptio if: You need a full-featured, model-centric ITFM platform with deep budgeting, forecasting, and AI-driven insights” GLM 4.7 FlashX · comparative prompt · first choice
“IBM Apptio (Market Leader) - Focus: Full Technology Business Management (TBM) - the most comprehensive ITFM solution” Kimi K2 · comparative prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of eight in this category shown.
“So avoid ServiceNow ITFM as a standalone buy if you are not already committed to ServiceNow ITSM/CMDB” Muse Glimmer 30B · negative prompt · hard negative
“Powerful but expensive; requires broader ServiceNow investment” Kimi K2 · budget prompt · hard negative
“Large platforms like ServiceNow or other enterprise suites can be attractive if you already use the ecosystem, but they can become complex, expensive, and harder to tailor” GPT-5.4 mini · negative prompt · soft negative
“vendors and analysts point to ServiceNow-native ITFM with a managed-service provider such as Nicus/Proven Optics” Muse Glimmer 30B · direct prompt · first choice
“names ServiceNow as market leader by revenue in 2025” Muse Glimmer 30B · comparative prompt · first choice
“Integrated Suites (e.g., ServiceNow ITFM) | Teams already using ServiceNow for IT Support/Ticketing.” Qwen 3.7 Flash · scale prompt · alternative
Comparisons are drawn for the top eight products in each category, each against each. The output is the models' output; nothing here is a recommendation by the index.