AI Indexes
IT AI Index
Index › Research › Repeat measurement · October 2026 Edition
Repeat measurement · October 2026 Edition

Asked the same question twice, the models moved their own first choice 63% of the time. The leader still held in 35 of 36 readings (12 categories at three buyer sizes).

Twelve categories were asked again three days after the edition run, with nothing changed: the same prompts, the same model versions, the same settings, one buyer segment, six framings, fourteen models, 1,008 answers. Whatever differs between the two runs is noise, and the noise is what this page measures. That is the reason the index asks every question six ways of every model rather than once: on its own, one ask moved 63% of the time; eighty-four, read together, are what the index reports.

Model by model

For each model, the share of its questions where the top pick differed between the two runs. The pairs are one question asked in each run; a pair flips when the first choice is not the same product.
GLM 4.7 FlashX80%39 of 49
Qwen 3.7 Flash73%44 of 60
Mistral Small72%31 of 43
Claude Haiku 4.572%33 of 46
Muse Glimmer 30B69%36 of 52
Llama 4 Maverick69%20 of 29
MiniMax M2.567%33 of 49
Gemini 3.5 Flash67%40 of 60
Kimi K261%38 of 62
Perplexity Sonar56%24 of 43
Grok 4.1 Fast54%33 of 61
DeepSeek V4 Flash53%34 of 64
GPT-5.4 mini49%22 of 45
GPT-6 Luna41%18 of 44

GLM 4.7 FlashX changed its first choice most often, 80% of its 49 pairs; GPT-6 Luna least, 41% of 44. Pooled over every model and question, 63%.

A flip is the same model, the same question, a different first choice a few days apart. It says how much one answer can be trusted on its own, and nothing about why the model answered as it did.

Category by category

The leader's share of first choices in the edition run and in the repeat, and whether the leader was the same product both times. Sorted by the size of the move.
CategoryLeader in the edition runShare, run oneShare, repeatMoveLeader
FinOps · Small business Vantage41%57%+16 pointsheld
API management · Enterprise Apigee48%33%-15 pointsheld
CIAM · Enterprise Auth047%33%-14 pointsheld
APM · Mid-market New Relic46%57%+11 pointsheld
AI coding · Mid-market GitHub Copilot75%65%-9 pointsheld
API management · Mid-market Kong Gateway37%45%+9 pointsheld
ETL · Enterprise Fivetran49%40%-9 pointsheld
CDN · Small business Cloudflare85%77%-8 pointsheld
Backup and DR · Small business Acronis Cyber Protect23%16%-7 pointschanged: IDrive Business
Password managers · Mid-market 1Password47%54%+7 pointsheld
Data warehouses · Small business Google BigQuery63%69%+6 pointsheld
FinOps · Mid-market Vantage45%40%-5 pointsheld
CIAM · Small business Clerk30%25%-5 pointsheld
Password managers · Enterprise 1Password44%39%-5 pointsheld
Data warehouses · Mid-market Snowflake49%54%+5 pointsheld
CDN · Mid-market Cloudflare68%73%+5 pointsheld
Compliance automation · Mid-market Drata34%38%+5 pointsheld
ETL · Mid-market Fivetran40%35%-4 pointsheld
Backup and DR · Enterprise Veeam Data Platform32%36%+4 pointsheld
AI coding · Small business GitHub Copilot63%67%+4 pointsheld
SASE · Enterprise Palo Alto Networks Prisma SASE26%29%+3 pointsheld
APM · Small business New Relic53%50%-3 pointsheld
API management · Small business Kong Gateway28%31%+3 pointsheld
Password managers · Small business Bitwarden49%46%-3 pointsheld
Compliance automation · Small business Vanta35%38%+3 pointsheld
ETL · Small business Airbyte36%34%-2 pointsheld
Backup and DR · Mid-market Veeam Data Platform38%40%+2 pointsheld
FinOps · Enterprise IBM Apptio Cloudability49%47%-2 pointsheld
CIAM · Mid-market Frontegg21%19%-2 pointsheld
Data warehouses · Enterprise Snowflake62%64%+2 pointsheld
AI coding · Enterprise GitHub Copilot82%80%-1 pointsheld
SASE · Small business Cloudflare One69%67%-1 pointsheld
Compliance automation · Enterprise Vanta37%38%+1 pointheld
APM · Enterprise Dynatrace47%47%+1 pointheld
SASE · Mid-market Cato Networks35%35%+0 pointsheld
CDN · Enterprise Cloudflare46%46%+0 pointsheld

In 35 of the 36 readings (12 categories at three buyer sizes) the same product led both runs. Where the leader changed, the two products were within 5 points of each other in the edition run.

The share floor

The bar a change has to clear before the index calls it a change, in the unit of the change itself.
11 pointsthe floor: the 90th percentile of the moves above
4 pointsmedian move of a leader's share on a repeat
16 pointsthe largest move, FinOps

From the next edition on, a product's change in share counts as movement only when it is larger than 11 points, and a new leader is reported only when it clears the old one by more than that. Nine repeats in ten move a leader less. The floor is measured again with every edition and the method page carries the rule: how the floor is measured.

Cite this

IT AI Recommendation Index, October 2026 Edition: repeat measurement. it-ai-index.com/research/repeat-measurement/. Published under CC BY 4.0. Every figure on this page is computed from the published edition and changes with it; the edition and its date are the citation.

The output is the models' output. Nothing here says the models can be steered, and nothing here is a recommendation by the index.