Two of fourteen models named Prefect first on the direct prompt; six named Make. Prefect was named by fourteen of the fourteen models and Make by twelve and Prefect carries 55 labels and Make 27, so the shares are not directly comparable.
Named in three categories this edition.
Named in five categories this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all fourteen models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the workflow orchestration page.
Across every category in the October 2026 Edition, Prefect and Make were named in the same answer fifty-five times, of the 158 answers naming Prefect and the 118 naming Make. In those answers Make took the first choice thirteen times and Prefect twelve.
| Model | Direct | Paraphrase | Comparative | Budget-constrained | Scale-constrained | Negative |
|---|---|---|---|---|---|---|
| Claude Haiku 4.5 | ||||||
| GPT-5.4 mini | ||||||
| Gemini 3.5 Flash | ||||||
| Perplexity Sonar | ||||||
| Grok 4.1 Fast | ||||||
| Mistral Small | ||||||
| DeepSeek V4 Flash | ||||||
| Llama 4 Maverick | ||||||
| Qwen 3.7 Flash | ||||||
| Kimi K2 | ||||||
| GLM 4.7 FlashX | ||||||
| MiniMax M2.5 | ||||||
| GPT-6 Luna | ||||||
| Muse Glimmer 30B |
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of seven in this category shown.
“Better UX than Airflow, but sparse documentation/StackOverflow support makes troubleshooting hard.” Grok 4.1 Fast · negative prompt · soft negative
“excellent alternatives to Airflow. However ... they carry a distinct business risk” Gemini 3.5 Flash · negative prompt · soft negative
“Prefect/Dagster | Long-running business processes with complex state” Kimi K2 · negative prompt · soft negative
“the best default choice is usually Prefect if you want a modern, Python-friendly orchestration tool with a relatively fast time to value” GPT-5.4 mini · direct prompt · first choice
“Prefect | Python devs, dynamic flows | Easy UI, auto-retries, hybrid deploy. ... | Excellent starter.” Grok 4.1 Fast · scale prompt · first choice
“Start with Prefect. It typically offers the fastest return on investment for mid-sized teams” Qwen 3.7 Flash · paraphrase prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of six in this category shown.
“Avoid No-Code (Zapier/Make) for data infrastructure, ETL, and backend logic.” Gemini 3.5 Flash · negative prompt · hard negative
“If your budget is bleeding from expensive tools like Zapier or Make” Gemini 3.5 Flash · budget prompt · hard negative
“Good visual builder at a budget price ($9/mo) — but workflows get messy at scale.” DeepSeek V4 Flash · direct prompt · soft negative
“Make is usually the best overall choice if you want a low-cost, user-friendly automation platform with a generous free tier” Perplexity Sonar · budget prompt · first choice
“Make typically offers the best balance of features, pricing, and ease of use for mid-market B2B companies” GLM 4.7 FlashX · direct prompt · first choice
“Make.com is the most cost-effective and powerful engine for a mid-market B2B budget” Gemini 3.5 Flash · direct prompt · first choice
Comparisons are drawn for the top eight products in each category, each against each. The output is the models' output; nothing here is a recommendation by the index.