Zero of twelve models named Lattice first on the direct prompt; zero named Leapsome. Lattice was named by eleven of the twelve models and Leapsome by eleven and Lattice carries 29 labels and Leapsome 20, so the shares are not directly comparable.
Named in nineteen categories this edition.
Named in thirteen categories this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all twelve models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the goals and okrs page.
Across every category in the September 2026 Edition, Lattice and Leapsome were named in the same answer 199 times, of the 672 answers naming Lattice and the 255 naming Leapsome. In those answers Leapsome took the first choice three times and Lattice fifty-six.
| Model | Direct | Paraphrase | Comparative | Budget-constrained | Scale-constrained | Negative |
|---|---|---|---|---|---|---|
| Claude Haiku 4.5 | ||||||
| GPT-5.4 mini | ||||||
| Gemini 3.5 Flash | ||||||
| Perplexity Sonar | ||||||
| Grok 4.1 Fast | ||||||
| Mistral Small | ||||||
| DeepSeek V4 Flash | ||||||
| Llama 4 Maverick | ||||||
| Qwen 3.7 Flash | ||||||
| Kimi K2 | ||||||
| GLM 4.7 FlashX | ||||||
| MiniMax M2.5 |
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
No label in this category carried a quote.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of three in this category shown.
“HR-suite "goals modules" (Lattice, Leapsome, ...) ... Risky if you want dedicated OKR tracking” MiniMax M2.5 · negative prompt · soft negative
“For a mid-sized B2B company, Leapsome is often the most recommended for its balance of features, ease of use, and scalability.” Mistral Small · paraphrase prompt · first choice
“Primary Recommendation: Leapsome” MiniMax M2.5 · paraphrase prompt · first choice
Comparisons are drawn for the top eight products in each category, each against each. The output is the models' output; nothing here is a recommendation by the index.