# TestGorilla vs HackerRank: which do AI models recommend for candidate assessment, October 2026

HR AI Recommendation Index, October 2026 Edition, Candidate assessment. Four of fourteen models named TestGorilla first on the direct prompt; zero named HackerRank. Page: https://hr-ai-index.com/talent/candidate-assessment/testgorilla-vs-hackerrank/

| | First-choice share | Rank | Negative rate | Labels | Models naming it |
|---|---|---|---|---|---|
| TestGorilla | 52% | #1 of 12 | 11% | 53 | 14 of 14 |
| HackerRank | 2% | #3 of 12 | 9% | 23 | 10 of 14 |

## The direct prompt, model by model

- GPT-5.4 mini: testgorilla first (first choices: TestGorilla) (alternatives: Codility, HackerRank, Mercer | Mettl)
- Gemini 3.5 Flash: testgorilla first (first choices: TestGorilla) (alternatives: Criteria Corp, The Predictive Index, Vervoe)
- Qwen 3.7 Flash: testgorilla first (first choices: TestGorilla) (alternatives: CodeSignal, HireVue, Humantic AI)
- GLM 4.7 FlashX: testgorilla first (first choices: TestGorilla) (alternatives: Testlify, Vervoe)
- Perplexity Sonar: neither first, one named (first choices: Criteria Corp) (alternatives: HireVue, TestGorilla, The Predictive Index)
- Grok 4.1 Fast: neither first, one named (first choices: Criteria Corp) (alternatives: HireVue, TestGorilla, The Predictive Index)
- DeepSeek V4 Flash: neither first, one named (first choices: Criteria Corp) (alternatives: TestGorilla, eSkill)
- Kimi K2: neither first, one named (first choices: Criteria Corp) (alternatives: TestGorilla, The Predictive Index, Vervoe, Wonderlic)
- GPT-6 Luna: neither first, one named (first choices: Criteria Corp) (alternatives: TestGorilla)
- Claude Haiku 4.5: neither named
- Mistral Small: neither named (first choices: Criteria Corp, Evalart) (alternatives: Adaface)
- Llama 4 Maverick: neither named
- MiniMax M2.5: neither named (first choices: Evalart, Toggl Hire)
- Muse Glimmer 30B: neither named (first choices: Criteria Corp) (alternatives: Evalart)

## What the models said about TestGorilla

- "Why Avoid: If you have a limited budget or require responsive technical support, this platform may lead to frustration." (GLM 4.7 FlashX, negative prompt, hard negative)
- "While widely used and effective for skills testing, some candidates complain about the quality of "Soft Skills" or "Culture Add" multiple-choice questions." (Qwen 3.7 Flash, negative prompt, soft negative)
- "Some candidate feedback highlights poor question quality (broken English, unclear phrasing), unexpected subscription charges post-use" (Grok 4.1 Fast, negative prompt, soft negative)
- "The Best All-Rounder: TestGorilla ... arguably the best on the market for offering a mix of cognitive ability, culture add, and role-specific skills" (Qwen 3.7 Flash, paraphrase prompt, first choice)
- "For most mid-sized B2B companies, TestGorilla offers the best balance of test breadth, pricing flexibility, and ease of deployment." (DeepSeek V4 Flash, paraphrase prompt, first choice)

Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all fourteen models, for a mid-market B2B company; rank is within the category. Comparisons are drawn for the top eight products in each category. Published under CC BY 4.0; the output is the models' output, and nothing here is a recommendation by the index.
