Four of fourteen models named Gem first on the direct prompt; zero named Phenom. Gem was named by twelve of the fourteen models and Phenom by fourteen and Gem carries 24 labels and Phenom 30, so the shares are not directly comparable.
Named in seven categories this edition.
Named in six categories this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all fourteen models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the talent intelligence page.
Across every category in the October 2026 Edition, Gem and Phenom were named in the same answer fifty-four times, of the 221 answers naming Gem and the 294 naming Phenom. In those answers Phenom took the first choice four times and Gem eleven.
| Model | Direct | Paraphrase | Comparative | Budget-constrained | Scale-constrained | Negative |
|---|---|---|---|---|---|---|
| Claude Haiku 4.5 | ||||||
| GPT-5.4 mini | ||||||
| Gemini 3.5 Flash | ||||||
| Perplexity Sonar | ||||||
| Grok 4.1 Fast | ||||||
| Mistral Small | ||||||
| DeepSeek V4 Flash | ||||||
| Llama 4 Maverick | ||||||
| Qwen 3.7 Flash | ||||||
| Kimi K2 | ||||||
| GLM 4.7 FlashX | ||||||
| MiniMax M2.5 | ||||||
| GPT-6 Luna | ||||||
| Muse Glimmer 30B |
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of three in this category shown.
“faces integration issues with some third-party systems and occasionally has UI and workflow clunkiness” Claude Haiku 4.5 · negative prompt · hard negative
“positioned more for outbound recruiting and pipeline management than broad talent intelligence” Perplexity Sonar · budget prompt · soft negative
“Gem is widely considered the best option for small businesses because they have a specific pricing tier designed to help you grow” GLM 4.7 FlashX · budget prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. One of one in this category shown.
“My default recommendation: Phenom, if you want one platform to support both external hiring and internal mobility” GPT-6 Luna · paraphrase prompt · first choice
Comparisons are drawn for the top eight products in each category, each against each. The output is the models' output; nothing here is a recommendation by the index.