IT AI Index
Index Vendors › Braintrust · September 2026 Edition
2 categories · Ranked

Braintrust

23Judge labels
3First choices
1Negative labels
11 of 12Models named it
2Categories
September 2026 Edition. Every number here is derived from the raw labels under vendor table vv2026-09.2, every buyer segment counted.
Best standing
2% in LLM gateways for mid-market buyers
Rank 8 of 63 in the mid-market standingaccepted challenger
0 of 12 models made it the first choice on the direct prompt; 0% of its 11 labels there were negative.
By buyer segmentRead the same way at every buyer size.
In llm gateways · each standing computed within its segment · bars are 0 to 100 · the accent bar is the product's own best reading

Standing by category

Every category where a model named Braintrust for a mid-market B2B company. Share is first choices across the direct, paraphrase, budget and scale prompts; rank is within every product named in that category.
CategoryFunctionShareRankNegative rateLabelsQuadrant
LLM gateways and observabilityData platform2%8 of 630%11accepted challenger
ML platformsData platform0%36 of 990%1under 10 labels · led by Azure Machine Learning at 17%

Movement

This is the first edition on this tier, so no move can be computed for Braintrust yet. From the next edition this section shows, per buyer segment, whether its share moved by more than the measured noise floor.

By model

How each model treated Braintrust across every prompt where it was named for a mid-market B2B company. Twelve models, six prompts per category.
ModelFirst choiceAlternativeMentionNegativeLabels
Claude Haiku 4.510001
GPT-5.4 mini00000
Gemini 3.5 Flash01001
Perplexity Sonar01102
Grok 4.1 Fast00000
Mistral Small01001
DeepSeek V4 Flash01102
Llama 4 Maverick00101
Qwen 3.7 Flash01001
Kimi K201102
GLM 4.7 FlashX00101
MiniMax M2.500000

By framing

Which of the six questions produced the naming. By model says how often; this says asked what. The first-choice count on the right carries the marks of the models that produced it.
FramingLabels by classFirst choices
Direct5 labelsNone
Paraphrase13 labels3
Comparative3 labelsNone
Budget-constrained0 labelsNone
Scale-constrained2 labelsNone
Negative0 labelsNone
First choiceAlternativeMentionNegative23 labels in all, every segment counted; 3 of the 3 first choices count toward share, since the comparative and negative framings do not. The bar is one segment per label class, to scale within the framing.

What the models said for it

Verbatim evidence the judge attached to positive labels.

“Best All-in-One Solution: Braintrust — offering monitoring, evaluation, and experimentation in one system” Claude Haiku 4.5 · LLM gateways · paraphrase prompt · first choice
“Offers robust features for governance and cost management, making it suitable for B2B companies with compliance and audit requirements.” Mistral Small · LLM gateways · direct prompt · alternative
“If you find yourself needing more advanced dataset management for evaluations, look at Braintrust.” Qwen 3.7 Flash · LLM gateways · paraphrase prompt · alternative
“Choose Braintrust if quality measurement and closed-loop routing are the main problem” Perplexity Sonar · LLM gateways · paraphrase prompt · alternative

And against it

Verbatim evidence attached to negative labels. A warning on a product with few labels is a warning; on a product with many, it is one voice among them.

No model argued against it.

Named alongside

The products named in the same answers as Braintrust, over the 23 answers that named it. Took the first choice instead counts the answers where the other product was the first choice and Braintrust was named but was not.
ProductSame answerTook the first choice insteadHead to head
Portkey14 of 236Not in the top three
Langfuse13 of 233Not in the top three
LiteLLM11 of 232Not in the top three
Helicone10 of 232Not in the top three
LangSmith10 of 230Not in the top three
OpenRouter7 of 230Not in the top three
Bifrost6 of 231Not in the top three
Cloudflare AI Gateway4 of 231Not in the top three
Confident AI4 of 230Not in the top three
Datadog4 of 230Not in the top three
A head-to-head page exists where both products are in a category's top three. The other rows are the same fact without a page behind them, so they link to the product instead.

What carried it into the answer

The sites and pages cited by the answers that named Braintrust. A fact about retrieval, not a lever on the model.

Citations exist only for the models that return a source list, four of the twelve in this edition, so these counts come from 21 of the 23 answers that named Braintrust and are not a share of its labels.

Domains cited

braintrust.devYour site17
respan.ai12
confident-ai.com10
dataiku.com9
toolradar.com9
getmaxim.ai8
fastrouter.ai7
gbhackers.com7
orq.ai7
dev.to6

Seventy-five of the ninety-two domain citations in answers naming Braintrust came from somebody else's page.

Pages cited

Pages are listed as the models cited them.

Is this your product?

Claim this page

Claiming is free and changes nothing in the data. A claimed page shows a verified contact who is told when each edition publishes and when Braintrust's standing changes by more than the noise floor; the right to propose corrections to the vendor table, meaning names the judge wrote that should or should not read as Braintrust, applied by version and listed in the change log; and a one-line description supplied by the vendor and marked as such.

It does not get any change to labels, shares or verdicts, any preview, or any say over which quotes appear. A verification link goes to your work email; an address at braintrust.dev is approved on the spot, any other address is reviewed by hand.

Your name and company appear on the claimed page, or the company alone if you ask below. A title and a LinkedIn address appear there too if you give them, and are left off if you do not. Your email address is never published.

Subscribe to the pack