Eleven of twelve models named Wellable first on the direct prompt; zero named Spring Health. Wellable was named by twelve of the twelve models and Spring Health by eight and Wellable carries 36 labels and Spring Health 14, so the shares are not directly comparable.
Named in three categories this edition.
Named in one category this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all twelve models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the employee wellbeing and mental health page.
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of four in this category shown.
“For most mid-market B2B companies, Wellable offers the best combination of engagement, flexibility, pricing, and ease of deployment” DeepSeek V4 Flash · direct prompt · first choice
“Wellable and GoPivot are the top choices due to their scalability, feature sets, and proven track records with mid-market clients” Mistral Small · direct prompt · first choice
“Choose Wellable if your main struggle is low engagement/buy-in and you want a fun, gamified interface.” Qwen 3.7 Flash · direct prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of four in this category shown.
“Start with a modern digital EAP platform (examples: Spring Health, Lyra, Headspace for Business, Unmind)” MiniMax M2.5 · paraphrase prompt · first choice
“combined with a dedicated therapy platform (like Lyra Health or Spring Health) for a mid-sized B2B company” Grok 4.1 Fast · paraphrase prompt · first choice
“Examples include providers like Modern Health, Lyra Health, Ginger, or Spring Health.” Qwen 3.7 Flash · paraphrase prompt · first choice
Comparisons are drawn for the top three products in each category. The output is the models' output; nothing here is a recommendation by the index.