# TestGorilla vs CodeSignal: which do AI models recommend for candidate assessment, October 2026

HR AI Recommendation Index, October 2026 Edition, Candidate assessment. Four of fourteen models named TestGorilla first on the direct prompt; zero named CodeSignal. Page: https://hr-ai-index.com/talent/candidate-assessment/testgorilla-vs-codesignal/

| | First-choice share | Rank | Negative rate | Labels | Models naming it |
|---|---|---|---|---|---|
| TestGorilla | 52% | #1 of 12 | 11% | 53 | 14 of 14 |
| CodeSignal | 2% | #5 of 12 | 8% | 13 | 8 of 14 |

## The direct prompt, model by model

- GPT-5.4 mini: testgorilla first (first choices: TestGorilla) (alternatives: Codility, HackerRank, Mercer | Mettl)
- Gemini 3.5 Flash: testgorilla first (first choices: TestGorilla) (alternatives: Criteria Corp, The Predictive Index, Vervoe)
- Qwen 3.7 Flash: testgorilla first (first choices: TestGorilla) (alternatives: CodeSignal, HireVue, Humantic AI)
- GLM 4.7 FlashX: testgorilla first (first choices: TestGorilla) (alternatives: Testlify, Vervoe)
- Perplexity Sonar: neither first, one named (first choices: Criteria Corp) (alternatives: HireVue, TestGorilla, The Predictive Index)
- Grok 4.1 Fast: neither first, one named (first choices: Criteria Corp) (alternatives: HireVue, TestGorilla, The Predictive Index)
- DeepSeek V4 Flash: neither first, one named (first choices: Criteria Corp) (alternatives: TestGorilla, eSkill)
- Kimi K2: neither first, one named (first choices: Criteria Corp) (alternatives: TestGorilla, The Predictive Index, Vervoe, Wonderlic)
- GPT-6 Luna: neither first, one named (first choices: Criteria Corp) (alternatives: TestGorilla)
- Claude Haiku 4.5: neither named
- Mistral Small: neither named (first choices: Criteria Corp, Evalart) (alternatives: Adaface)
- Llama 4 Maverick: neither named
- MiniMax M2.5: neither named (first choices: Evalart, Toggl Hire)
- Muse Glimmer 30B: neither named (first choices: Criteria Corp) (alternatives: Evalart)

## What the models said about TestGorilla

- "Why Avoid: If you have a limited budget or require responsive technical support, this platform may lead to frustration." (GLM 4.7 FlashX, negative prompt, hard negative)
- "While widely used and effective for skills testing, some candidates complain about the quality of "Soft Skills" or "Culture Add" multiple-choice questions." (Qwen 3.7 Flash, negative prompt, soft negative)
- "Some candidate feedback highlights poor question quality (broken English, unclear phrasing), unexpected subscription charges post-use" (Grok 4.1 Fast, negative prompt, soft negative)
- "The Best All-Rounder: TestGorilla ... arguably the best on the market for offering a mix of cognitive ability, culture add, and role-specific skills" (Qwen 3.7 Flash, paraphrase prompt, first choice)
- "For most mid-sized B2B companies, TestGorilla offers the best balance of test breadth, pricing flexibility, and ease of deployment." (DeepSeek V4 Flash, paraphrase prompt, first choice)

## What the models said about CodeSignal

- "The best skills testing tool for hiring for a mid-sized B2B company would be CodeSignal, TestGorilla, or Criteria Corp." (Llama 4 Maverick, paraphrase prompt, first choice)

Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all fourteen models, for a mid-market B2B company; rank is within the category. Comparisons are drawn for the top eight products in each category. Published under CC BY 4.0; the output is the models' output, and nothing here is a recommendation by the index.
