Zero of fourteen models named TaxAct Business first on the direct prompt; zero named Vertex. TaxAct Business was named by nine of the fourteen models and Vertex by seven and both carry 10 labels, so the shares below are directly comparable.
Named in one category this edition.
Named in five categories this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all fourteen models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the corporate tax provision page.
| Model | Direct | Paraphrase | Comparative | Budget-constrained | Scale-constrained | Negative |
|---|---|---|---|---|---|---|
| Claude Haiku 4.5 | ||||||
| GPT-5.4 mini | ||||||
| Gemini 3.5 Flash | ||||||
| Perplexity Sonar | ||||||
| Grok 4.1 Fast | ||||||
| Mistral Small | ||||||
| DeepSeek V4 Flash | ||||||
| Llama 4 Maverick | ||||||
| Qwen 3.7 Flash | ||||||
| Kimi K2 | ||||||
| GLM 4.7 FlashX | ||||||
| MiniMax M2.5 | ||||||
| GPT-6 Luna | ||||||
| Muse Glimmer 30B |
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of four in this category shown.
“TaxAct Business \u2013 Best Value for Small Companies ... at ~$170-280/year is your most cost-effective option” Kimi K2 · budget prompt · first choice
“Top Recommendation: TaxAct Business ... Often cited as the most cost-effective professional-grade option.” Qwen 3.7 Flash · budget prompt · first choice
“TaxAct Business is generally considered the best value for small businesses with limited budgets” Mistral Small · budget prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. One of one in this category shown.
“Vertex - a comprehensive tax compliance software that can handle complex tax logic and is used by large and mid-market companies.” Llama 4 Maverick · paraphrase prompt · first choice
Comparisons are drawn for the top eight products in each category, each against each. The output is the models' output; nothing here is a recommendation by the index.