AI Indexes
IT AI Index
October 2026 Edition · The permanent record of this edition. The unqualified address always carries the latest edition.
Index › IT operations and endpoint › October 2026 Edition

Application performance monitoring

Asked as “APM tool”, and as “application performance monitoring solution”, on behalf of a mid-market B2B company. 54 first choices recorded across the direct, paraphrase, budget and scale prompts, fourteen models each.
Standing · first-choice share
46%
Clear leader
46New Relic20Datadog APM09SigNoz24others

46% of first choices, clear leader.

Since September 2026▼−5Since September 2026: 54% → 49%, −5 points. Inside the 11-point floor: within noise. Read over the models both editions asked.New Relic held the lead, −5 points on 54%, inside the 11-point floor.

By buyer segment

The same question asked on behalf of a different buyer. Each standing is computed within its segment; they sit side by side and are never added together.

The standing

Share is the count of first choices across the direct, paraphrase, budget and scale prompts, over all fourteen models, for a mid-market B2B company. Ordered by share.
ProductFirst-choice shareNegative rateLabelsQuadrantSince September 2026
01New Relic46%12%65endorsed leader▼−5Since September 2026: 54% → 49%, −5 points. Inside the 11-point floor: within noise. Read over the models both editions asked.54% → 49%
02Datadog APM20%29%68criticized challenger▲+7Since September 2026: 11% → 18%, +7 points. Inside the 11-point floor: within noise. Read over the models both editions asked.11% → 18%
03SigNoz9%0%23accepted challenger▼−6Since September 2026: 13% → 7%, −6 points. Inside the 11-point floor: within noise. Read over the models both editions asked.13% → 7%
04Site24x76%0%16accepted challenger▲+4Since September 2026: 2% → 7%, +4 points. Inside the 11-point floor: within noise. Read over the models both editions asked.2% → 7%
05Elastic APM4%11%28accepted challenger▲+4Since September 2026: 0% → 4%, +4 points. Inside the 11-point floor: within noise. Read over the models both editions asked.0% → 4%
06IBM Instana4%17%18accepted challenger=heldSince September 2026: 2% → 2%, ±0 points. Inside the 11-point floor: within noise. Read over the models both editions asked.2% → 2%
07Grafana Cloud2%0%26accepted challenger▼−2Since September 2026: 4% → 2%, −2 points. Inside the 11-point floor: within noise. Read over the models both editions asked.4% → 2%
08Honeycomb2%0%10accepted challenger=heldSince September 2026: 2% → 2%, ±0 points. Inside the 11-point floor: within noise. Read over the models both editions asked.2% → 2%
09Dynatrace2%48%50criticized challenger▲+2Since September 2026: 0% → 2%, +2 points. Inside the 11-point floor: within noise. Read over the models both editions asked.0% → 2%
Show the three products at 0%, ordered by negative rate
11AppDynamics0%27%37criticized challenger=heldSince September 2026: 0% → 0%, ±0 points. Inside the 11-point floor: within noise. Read over the models both editions asked.0% → 0%
12Jaeger0%18%11accepted challenger=heldSince September 2026: 0% → 0%, ±0 points. Inside the 11-point floor: within noise. Read over the models both editions asked.0% → 0%
10Sentry0%0%13accepted challenger=heldSince September 2026: 0% → 0%, ±0 points. Inside the 11-point floor: within noise. Read over the models both editions asked.0% → 0%

The floor is 11 points of share, measured: how far the models move a leader on their own when the same questions are asked twice with nothing changed. A larger change is movement; a smaller one is noise, and both are shown. Movement is read over the twelve models both editions asked; GPT-6 Luna, Muse Glimmer 30B joined this edition and are in the standing but not yet in the comparison. How the floor is measured

Bars are the share of first choices, 0 to 100Every product with at least 10 labels here. Every product name links to its product page.
All twenty-eight head-to-head pages: the top eight products, each against each

Recommended versus criticized

Every product with at least 10 labels here, on both axes. The 30% line names a quadrant, not the verdict above: that one needs more than 40%.

Criticized challengerCriticized default
Negative label rate →
01
02
03
04
05
06
07
08
09
10
11
J12
Accepted challengerEndorsed leader
0%First-choice share → · lines at 30% share and 25% negative60%
Key
01New Relic46%
02Datadog APM20%
03SigNoz9%
04Site24x76%
05Elastic APM4%
06IBM Instana4%
07Grafana Cloud2%
08Honeycomb2%
09Dynatrace2%
10Sentry0%
11AppDynamics0%
12Jaeger0%

What they warned about

Seven of fourteen models held their first choice under the paraphrase. Claude Haiku 4.5, Grok 4.1 Fast, Mistral Small, Qwen 3.7 Flash, Kimi K2, MiniMax M2.5 and Muse Glimmer 30B changed. A high negative share on a product with few labels is a warning. A low share on a product with many labels is salience, not sentiment.
Dynatrace
48%
24 of 50 labels negative · 13 of 14 models · 5 hard negative
“Verdict: Avoid unless you are a large enterprise with a dedicated Site Reliability Engineering (SRE) team” Qwen 3.7 Flash, negative prompt
Datadog APM
29%
20 of 68 labels negative · 13 of 14 models · 4 hard negative
“tools and vendors you should be **cautious about** or potentially **avoid**... **1. Datadog** - Most Common Complaint ... **Unpredictable pricing**” Kimi K2, negative prompt
AppDynamics
27%
10 of 37 labels negative · 9 of 14 models · 3 hard negative
“Teams on perpetual licenses, Rapid Start SKUs and older agents are being actively sunset. If you are evaluating a "new" AppDynamics deployment in 2026 you are effectively buying into Splunk's portfolio roadmap” Muse Glimmer 30B, negative prompt
New Relic
12%
8 of 65 labels negative · 8 of 14 models · 1 hard negative
“**New Relic** if your team lacks a dedicated observability engineer.” DeepSeek V4 Flash, negative prompt

What they cite

Citations exist only for the models that return a source list: fourteen of the fourteen in this edition, and all six flagship models on the expanded tier.

Sites the answers cite

76 of 84 answers in this category came back with a source list, from 14 of 14 models: citations where the model returns them, or the search results it consulted. 964 links across 199 sites, every framing counted. Ranked by the number of answers carrying the site or page. 10 of the 252 answers across every segment cited this index's own page for the category; the method page measures whether that reading tilts an answer.

vendor site · Motadata44 answers · 45 citations · 11 models
vendor site · G234 answers · 58 citations · 13 models
vendor site · ManageEngine30 answers · 42 citations · 11 models
vendor site · New Relic28 answers · 44 citations · 12 models
vendor site · Middleware.io28 answers · 30 citations · 11 models
23 answers · 23 citations · 12 models
vendor site · G222 answers · 24 citations · 9 models
vendor site · Coralogix22 answers · 22 citations · 7 models
vendor site · Uptrace21 answers · 23 citations · 9 models
19 answers · 21 citations · 8 models
vendor site · Atatus18 answers · 29 citations · 10 models
vendor site · OpenObserve17 answers · 18 citations · 10 models

Pages the answers cite

The ten pages named in the most answers, by full address. A page here is one the models returned with a recommendation, not one the index endorses.

Search against answers

Each company's standing in the answers beside its site's footprint in Google search, one row a site: the products the models named on it with their shares, and the share they add up to; monthly searches on Google, and DataForSEO's estimate of AI search demand (modeled from search signals, directional, not a count of queries to any assistant), for the most-searched of the company's and its products' names (the name is in each row's hover text); estimated monthly organic visits to the site; and its best position in Google's top ten for “apm tool”, “apm tools”, “best apm tool”. US estimates from DataForSEO and Google's Ads Transparency Center. A small company's site, or a mid-sized company's site for its flagship, is marked company; a product on a large parent's site (Google, Microsoft) has no site figures. A column with no figures for this category is left out, and an empty cell means none were seen, not none exist. Two measurements side by side: neither is read as the cause of the other.
Company and productsShareOwn site citedName searches, GoogleAI search demand, est.Organic visitsPaid search a month, est.Google ads, last 30 daysBest Google position
New Relic newrelic.com46%62914,80084383,276$412894
Datadog datadoghq.com
Datadog APM 20%
20%18674,0003,541159,408 company$633120
SigNoz signoz.io9%2082,400527,24415
Site24x7 site24x7.com6%2141,30061174,542$522120
Elastic elastic.co
Elastic APM 4%, Elastic Observability 0%
4%10974,00034,672208,306 company105
IBM parent site
IBM Instana 4%
4%74201,00034,098
OpenObserve openobserve.ai4%1571,60094,82575
Grafana Cloud grafana.com
Grafana Cloud 2%, Grafana Stack 2%, Prometheus + Grafana 0%
4%1742,4004084,520116
Honeycomb honeycomb.io2%89135,00013,7326,60365
Dynatrace dynatrace.com2%70022,2001,10662,738$3,9411202

Google's top ten for the three searches hold 15 sites; 4 of them are among the sites the models cited here (manageengine.com, motadata.com, newrelic.com, openobserve.ai). The first result for “apm tool” is blazemeter.com.

How they answered

Six framings of the same buying question, each sent to every model in a fresh session with search on. One row per model, so a row shows whether it held its answer under rewording, what it named when cost was the constraint, and what it argued against. Computed from the raw judge labels.
ShowHide
ModelDirect“What is the best APM tool for a mid-market B2B company?”Paraphrase“Which application performance monitoring solution would you recommend to a mid-sized B2B company?”Comparative“What are the top APM tools and how do they differ?”Budget-constrained“What is the best APM tool for a company with a limited budget?”Scale-constrained“We are a 500 person company evaluating a APM tool. What should we look at?”Negative“Which APM tools should I avoid or be cautious about?”
Claude Haiku 4.5Datadog APM
Three alternativesNew Relic, Site24x7, SolarWinds Observability
New RelicChanged
Three alternativesDatadog APM, Elastic APM, Site24x7
Datadog APM, Dynatrace
Four alternativesElastic APM, Grafana Cloud, New Relic, SigNoz
New Relic
Five alternativesElastic APM, Grafana Cloud, Prometheus, Sentry, SigNoz
no first choiceagainst: Baselime, Datadog APM, Dynatrace, Elastic APM, New Relic, Odigos, Splunk Observability Cloud
GPT-5.4 miniNew Relic
One alternativeDatadog APM
against: Dynatrace
New RelicHeld
Three alternativesAppDynamics, Datadog APM, Dynatrace
no first choice
Five alternativesAppDynamics, Datadog APM, Dynatrace, Elastic APM, New Relic
Elastic APM
Two alternativesGrafana Cloud, New Relic
against: Datadog APM
no first choicenothing named
Gemini 3.5 FlashNew Relic
Two alternativesGrafana Cloud, Sentry
against: Datadog APM
New RelicHeld
Three alternativesDatadog APM, Honeycomb, Sentry
Datadog APM
Five alternativesAppDynamics, Dynatrace, Elastic APM, IBM Instana, New Relic
Grafana Cloud, New Relic, SigNoz
Four alternativesOpenObserve, Sentry, Site24x7, Uptrace
against: Datadog APM, Dynatrace
Honeycomb
Three alternativesGrafana Cloud, New Relic, SigNoz
against: Datadog APM
against: AppDynamics, Datadog APM, Dynatrace, Grafana Loki, Jaeger, Prometheus
Perplexity SonarDatadog APM
Three alternativesAppDynamics, Dynatrace, ManageEngine Applications Manager
Datadog APMHeld
Three alternativesDynatrace, ManageEngine Applications Manager, Site24x7
no first choiceSigNoz
Three alternativesAtatus, CubeAPM, New Relic
no first choicenothing named
Grok 4.1 FastDatadog APM
Two alternativesIBM Instana, New Relic
against: Dynatrace
New RelicChanged
Two alternativesDatadog APM, Site24x7
against: Dynatrace
no first choice
Five alternativesAppDynamics, Datadog APM, Dynatrace, Elastic APM, New Relic
New Relic
Five alternativesGrafana Cloud, Prometheus + Grafana, Sentry, SigNoz, Uptrace
no first choiceagainst: AppDynamics, Datadog APM, Dynatrace, New Relic
Mistral SmallIBM Instana
Three alternativesAppDynamics, Site24x7, Stackify Retrace
no first choiceChangedDatadog APM
Four alternativesAppDynamics, Dynatrace, Elastic APM, New Relic
OpenObserve
Two alternativesElastic APM, Grafana Stack
no first choiceagainst: AppDynamics, Datadog APM, Dynatrace, Elastic APM, IBM Instana, New Relic, Splunk Observability Cloud
DeepSeek V4 FlashNew Relic
Three alternativesDatadog APM, IBM Instana, Site24x7
against: Dynatrace
New RelicHeld
Two alternativesDatadog APM, Grafana Cloud
against: AppDynamics, Dynatrace
no first choiceNew Relic
Three alternativesHoneycomb, Sentry, SigNoz
no first choiceagainst: AppDynamics, Datadog APM, Dynatrace, IBM Instana, New Relic
Llama 4 Maverickno first choice
Two alternativesOpManager Nexus, Stackify Retrace
no first choiceHeldno first choiceElastic APM, Grafana Stack, OpenObserveno first choiceagainst: AppDynamics, Datadog APM, Dynatrace, New Relic
Qwen 3.7 FlashNew Relic
One alternativeSolarWinds Observability
against: Datadog APM
Datadog APMChanged
Three alternativesElasticsearch, Grafana Cloud, SigNoz
against: Dynatrace, IBM AppDynamics
Datadog APM
Six alternativesAppDynamics, Dynatrace, Elastic APM, Elastic Observability, Honeycomb, New Relic
New Relic
Five alternativesGrafana Cloud, Grafana Loki, ManageEngine OpManager Nexus, Prometheus, SigNoz
no first choiceagainst: AppDynamics, Broadcom DX APM, Datadog APM, Dynatrace, IBM APM, Oracle DBCA, Splunk
Kimi K2New Relic
Three alternativesDatadog APM, IBM Instana, Site24x7
against: Dynatrace
New Relic, Site24x7Changedagainst: Datadog APM, Dynatraceno first choiceSigNoz
Three alternativesAtatus, Jaeger, Zipkin
Datadog APM, Dynatrace, New Relicagainst: Datadog APM, Dynatrace, New Relic
GLM 4.7 FlashXSite24x7
Two alternativesDatadog APM, IBM Instana
Site24x7Held
One alternativeManageEngine Applications Manager
against: Datadog APM, New Relic
no first choiceNew Relic
Nine alternativesAtatus, Elastic APM, Grafana Cloud, Highlight.io, Scout APM, Scouter, Sentry, SigNoz, Uptrace
no first choiceagainst: AppDynamics, Datadog APM, Dynatrace, Elastic APM, Jaeger, Zipkin
MiniMax M2.5New Relic
One alternativeDatadog APM
against: Dynatrace
Datadog APM, New RelicChanged
One alternativeSite24x7
against: AppDynamics, Dynatrace
no first choiceNew Relic
Five alternativesAzure Application Insights, Elastic APM, Google Cloud Operations, Sentry, SigNoz
Datadog APM, New Relic
Two alternativesJaeger, Prometheus
against: Datadog APM, Dynatrace, IBM Instana
GPT-6 LunaDatadog APM
One alternativeNew Relic
Datadog APMHeld
Two alternativesGrafana Cloud, New Relic
no first choiceSigNoz
Two alternativesGrafana Cloud, Self-hosted SigNoz
against: Datadog APM
no first choiceagainst: Datadog APM, Dynatrace
Muse Glimmer 30BDatadog APM, IBM Instana, New Relic
Two alternativesAppDynamics, Site24x7
New RelicChanged
Two alternativesDatadog APM, IBM Instana
against: Dynatrace
Datadog APM
Two alternativesDynatrace, New Relic
New Relic, SigNoz
Five alternativesAtatus, Elastic APM, Grafana Cloud, Prometheus + Grafana, SolarWinds AppOptics
no first choiceagainst: AppDynamics, Datadog APM, Dynatrace, New Relic
Bold is the first choiceAlternatives are counted; the count opens them.What the answer argued against

The record

One row per call: the version string exactly as returned, whether the model searched, sources cited, and latency. Full answer text is in the free responses file. Download the record
Eighty-four rows: every prompt, every model, every answer.
PromptModelVersion stringTime (UTC)SearchedSourcesLatency
Direct recommendationClaude Haiku 4.5claude-haiku-4-5-202510012026-10-01 10:43yes1510 s
Direct recommendationGPT-5.4 minigpt-5.4-mini-2026-03-172026-10-01 11:11yes46 s
Direct recommendationGemini 3.5 Flashgemini-3.5-flash2026-10-01 13:46yes1828 s
Direct recommendationPerplexity Sonarsonar2026-10-01 10:55yes183 s
Direct recommendationGrok 4.1 Fastspacexai/grok-4.1-fast-non-reasoning via vertex2026-10-01 08:54yes187 s
Direct recommendationMistral Smallmistral/mistral-small via mistral2026-10-01 12:09yes115 s
Direct recommendationDeepSeek V4 Flashdeepseek/deepseek-v4-flash via deepinfra2026-10-01 12:32yes2231 s
Direct recommendationLlama 4 Maverickmeta/llama-4-maverick via bedrock2026-10-01 10:52yes52 s
Direct recommendationQwen 3.7 Flashalibaba/qwen3.7-flash via alibaba2026-10-01 12:15yes950 s
Direct recommendationKimi K2moonshotai/kimi-k2 via novita2026-10-01 11:37yes2019 s
Direct recommendationGLM 4.7 FlashXzai/glm-4.7-flashx via zai2026-10-01 08:09yes1756 s
Direct recommendationMiniMax M2.5minimax/minimax-m2.5 via minimax2026-10-01 09:46yes927 s
Direct recommendationGPT-6 Lunagpt-6-luna2026-10-01 12:48yes212 s
Direct recommendationMuse Glimmer 30Bmeta/muse-glimmer-30b via togetherai2026-10-01 11:07yes1116 s
ParaphraseClaude Haiku 4.5claude-haiku-4-5-202510012026-10-01 09:07yes189 s
ParaphraseGPT-5.4 minigpt-5.4-mini-2026-03-172026-10-01 12:51yes45 s
ParaphraseGemini 3.5 Flashgemini-3.5-flash2026-10-01 13:51no014 s
ParaphrasePerplexity Sonarsonar2026-10-01 10:10yes153 s
ParaphraseGrok 4.1 Fastspacexai/grok-4.1-fast-non-reasoning via vertex2026-10-01 08:58yes206 s
ParaphraseMistral Smallmistral/mistral-small via mistral2026-10-01 13:04no02 s
ParaphraseDeepSeek V4 Flashdeepseek/deepseek-v4-flash via deepinfra2026-10-01 08:39yes2424 s
ParaphraseLlama 4 Maverickmeta/llama-4-maverick via bedrock2026-10-01 08:01yes52 s
ParaphraseQwen 3.7 Flashalibaba/qwen3.7-flash via alibaba2026-10-01 12:28no032 s
ParaphraseKimi K2moonshotai/kimi-k2 via novita2026-10-01 08:46yes1329 s
ParaphraseGLM 4.7 FlashXzai/glm-4.7-flashx via zai2026-10-01 10:40yes2026 s
ParaphraseMiniMax M2.5minimax/minimax-m2.5 via minimax2026-10-01 07:56yes536 s
ParaphraseGPT-6 Lunagpt-6-luna2026-10-01 12:14yes313 s
ParaphraseMuse Glimmer 30Bmeta/muse-glimmer-30b via togetherai2026-10-01 10:48yes1519 s
ComparativeClaude Haiku 4.5claude-haiku-4-5-202510012026-10-01 13:24yes1010 s
ComparativeGPT-5.4 minigpt-5.4-mini-2026-03-172026-10-01 08:24yes76 s
ComparativeGemini 3.5 Flashgemini-3.5-flash2026-10-01 12:42yes1429 s
ComparativePerplexity Sonarsonar2026-10-01 09:38yes186 s
ComparativeGrok 4.1 Fastspacexai/grok-4.1-fast-non-reasoning via vertex2026-10-01 08:11yes249 s
ComparativeMistral Smallmistral/mistral-small via mistral2026-10-01 07:56yes87 s
ComparativeDeepSeek V4 Flashdeepseek/deepseek-v4-flash via deepinfra2026-10-01 09:38yes2229 s
ComparativeLlama 4 Maverickmeta/llama-4-maverick via bedrock2026-10-01 11:58yes51 s
ComparativeQwen 3.7 Flashalibaba/qwen3.7-flash via alibaba2026-10-01 11:32yes931 s
ComparativeKimi K2moonshotai/kimi-k2 via novita2026-10-01 07:28yes1527 s
ComparativeGLM 4.7 FlashXzai/glm-4.7-flashx via zai2026-10-01 08:29yes1327 s
ComparativeMiniMax M2.5minimax/minimax-m2.5 via minimax2026-10-01 10:45yes512 s
ComparativeGPT-6 Lunagpt-6-luna2026-10-01 09:27yes824 s
ComparativeMuse Glimmer 30Bmeta/muse-glimmer-30b via togetherai2026-10-01 10:16yes2245 s
Budget constrainedClaude Haiku 4.5claude-haiku-4-5-202510012026-10-01 12:28yes97 s
Budget constrainedGPT-5.4 minigpt-5.4-mini-2026-03-172026-10-01 13:50yes45 s
Budget constrainedGemini 3.5 Flashgemini-3.5-flash2026-10-01 13:04yes2026 s
Budget constrainedPerplexity Sonarsonar2026-10-01 09:11yes182 s
Budget constrainedGrok 4.1 Fastspacexai/grok-4.1-fast-non-reasoning via vertex2026-10-01 11:27yes1112 s
Budget constrainedMistral Smallmistral/mistral-small via mistral2026-10-01 07:49yes52 s
Budget constrainedDeepSeek V4 Flashdeepseek/deepseek-v4-flash via deepinfra2026-10-01 12:33yes2233 s
Budget constrainedLlama 4 Maverickmeta/llama-4-maverick via bedrock2026-10-01 13:11yes52 s
Budget constrainedQwen 3.7 Flashalibaba/qwen3.7-flash via alibaba2026-10-01 13:16yes533 s
Budget constrainedKimi K2moonshotai/kimi-k2 via novita2026-10-01 12:28yes1218 s
Budget constrainedGLM 4.7 FlashXzai/glm-4.7-flashx via zai2026-10-01 08:20yes1324 s
Budget constrainedMiniMax M2.5minimax/minimax-m2.5 via minimax2026-10-01 09:26yes959 s
Budget constrainedGPT-6 Lunagpt-6-luna2026-10-01 11:55yes38 s
Budget constrainedMuse Glimmer 30Bmeta/muse-glimmer-30b via togetherai2026-10-01 09:57yes1424 s
Scale constrainedClaude Haiku 4.5claude-haiku-4-5-202510012026-10-01 09:00no05 s
Scale constrainedGPT-5.4 minigpt-5.4-mini-2026-03-172026-10-01 10:34no08 s
Scale constrainedGemini 3.5 Flashgemini-3.5-flash2026-10-01 11:29yes828 s
Scale constrainedPerplexity Sonarsonar2026-10-01 10:47yes205 s
Scale constrainedGrok 4.1 Fastspacexai/grok-4.1-fast-non-reasoning via vertex2026-10-01 11:50yes148 s
Scale constrainedMistral Smallmistral/mistral-small via mistral2026-10-01 09:00no08 s
Scale constrainedDeepSeek V4 Flashdeepseek/deepseek-v4-flash via deepinfra2026-10-01 10:52yes1828 s
Scale constrainedLlama 4 Maverickmeta/llama-4-maverick via bedrock2026-10-01 11:22yes52 s
Scale constrainedQwen 3.7 Flashalibaba/qwen3.7-flash via alibaba2026-10-01 08:53no031 s
Scale constrainedKimi K2moonshotai/kimi-k2 via novita2026-10-01 07:26yes1318 s
Scale constrainedGLM 4.7 FlashXzai/glm-4.7-flashx via zai2026-10-01 12:59no012 s
Scale constrainedMiniMax M2.5minimax/minimax-m2.5 via minimax2026-10-01 12:14yes518 s
Scale constrainedGPT-6 Lunagpt-6-luna2026-10-01 07:50yes229 s
Scale constrainedMuse Glimmer 30Bmeta/muse-glimmer-30b via togetherai2026-10-01 11:51yes1125 s
Negative framingClaude Haiku 4.5claude-haiku-4-5-202510012026-10-01 11:21yes179 s
Negative framingGPT-5.4 minigpt-5.4-mini-2026-03-172026-10-01 10:49yes66 s
Negative framingGemini 3.5 Flashgemini-3.5-flash2026-10-01 11:39yes1922 s
Negative framingPerplexity Sonarsonar2026-10-01 12:27yes205 s
Negative framingGrok 4.1 Fastspacexai/grok-4.1-fast-non-reasoning via vertex2026-10-01 11:37yes248 s
Negative framingMistral Smallmistral/mistral-small via mistral2026-10-01 09:42yes57 s
Negative framingDeepSeek V4 Flashdeepseek/deepseek-v4-flash via deepinfra2026-10-01 11:50yes2434 s
Negative framingLlama 4 Maverickmeta/llama-4-maverick via bedrock2026-10-01 09:47yes52 s
Negative framingQwen 3.7 Flashalibaba/qwen3.7-flash via alibaba2026-10-01 10:48yes1847 s
Negative framingKimi K2moonshotai/kimi-k2 via novita2026-10-01 09:01yes1516 s
Negative framingGLM 4.7 FlashXzai/glm-4.7-flashx via zai2026-10-01 09:57yes2030 s
Negative framingMiniMax M2.5minimax/minimax-m2.5 via minimax2026-10-01 11:16yes1728 s
Negative framingGPT-6 Lunagpt-6-luna2026-10-01 08:30yes316 s
Negative framingMuse Glimmer 30Bmeta/muse-glimmer-30b via togetherai2026-10-01 09:49yes2436 s

Noise floor in this category

Flips between the edition run and its calibration repeat. Six prompts per model is a small sample; the index-wide floor is the number to trust.

ShowHide
Claude Haiku 4.5
5 of 5 flipped
GPT-5.4 mini
3 of 4 flipped
Gemini 3.5 Flash
3 of 5 flipped
Perplexity Sonar
3 of 4 flipped
Grok 4.1 Fast
2 of 4 flipped
Mistral Small
1 of 3 flipped
DeepSeek V4 Flash
3 of 5 flipped
Llama 4 Maverick
1 of 1 flipped
Qwen 3.7 Flash
3 of 6 flipped
Kimi K2
3 of 4 flipped
GLM 4.7 FlashX
3 of 3 flipped
MiniMax M2.5
3 of 4 flipped
GPT-6 Luna
3 of 3 flipped
Muse Glimmer 30B
3 of 4 flipped

Normalization in this category

Every judgment call made between the raw labels and the numbers above, listed so it is visible and reversible.

ShowHide
Category-scoped readings
Cisco AppDynamics read as AppDynamics
Datadog read as Datadog APM
Elastic Cloud read as Elastic APM
Grafana read as Grafana Cloud
Grafana (with Tempo/Prometheus) read as Grafana Cloud
Splunk APM read as Splunk Observability Cloud
Splunk Observability Cloud APM read as Splunk Observability Cloud
Unresolved, counted raw
Baselime
Datadog Community Edition
IBM APM
IBM AppDynamics
Kinsta APM
Logic Monitor
Odigos
Oracle DBCA
Self-hosted SigNoz
Site24x7 or ManageEngine OpManager
Discontinued, still offered
No shut-down product was recommended here.
← AIOpsBusiness phones →