AI Indexes
HR AI Index
Index › Talent acquisition › Candidate assessment › TestGorilla vs CodeSignal
Candidate assessment · October 2026 Edition

TestGorilla vs CodeSignal

Four of fourteen models named TestGorilla first on the direct prompt; zero named CodeSignal. TestGorilla was named by fourteen of the fourteen models and CodeSignal by eight and TestGorilla carries 53 labels and CodeSignal 13, so the shares are not directly comparable.

TestGorilla

endorsed leader

Named in four categories this edition.

CodeSignal

accepted challenger

Named in four categories this edition.

First-choice share52%2%Of first choices across the direct, paraphrase, budget and scale prompts, 0 to 100.
Negative rate11%8%Negative labels as a share of the product's labels, 0 to 100.
Rank in category#1#5A position in a field of 12; printed, not drawn.
Labels5313A count; the two differ.
The two percentage rows are drawn on one 0 to 100 track, TestGorilla reading right to left. Rank and label count are printed, not drawn.Criteria Corp was named alongside these two in eight of the fourteen direct answers. TestGorilla vs Criteria Corp · TestGorilla vs HackerRank · TestGorilla vs Testlify

Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all fourteen models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the candidate assessment page.

By framing

How many of the fourteen models made each the first choice, per way of asking, and how many argued against it.
TestGorillaFirst choices, of fourteen modelsCodeSignal
Direct40
Paraphrase121
Comparative60
Budget-constrained901 against TestGorilla · 1 against CodeSignal
Scale-constrained00
Negative005 against TestGorilla
Bars are first choices, 0 to 14 each sideModels that argued againstA model can name both, so the two sides of a row do not sum to fourteen.

Across every category in the October 2026 Edition, TestGorilla and CodeSignal were named in the same answer eighty-four times, of the 265 answers naming TestGorilla and the 186 naming CodeSignal. In those answers CodeSignal took the first choice eleven times and TestGorilla twenty-two.

Every model, every framing

The eighty-four answers behind the chart above, one cell each: where TestGorilla and CodeSignal stood in it.
ModelDirectParaphraseComparativeBudget-constrainedScale-constrainedNegative
Claude Haiku 4.5
GPT-5.4 mini
Gemini 3.5 Flash
Perplexity Sonar
Grok 4.1 Fast
Mistral Small
DeepSeek V4 Flash
Llama 4 Maverick
Qwen 3.7 Flash
Kimi K2
GLM 4.7 FlashX
MiniMax M2.5
GPT-6 Luna
Muse Glimmer 30B
TestGorilla CodeSignal first choice named as an alternative argued againstblank: not namedEach cell is one answer, TestGorilla on the left and CodeSignal on the right.

The direct prompt

The plain question, one answer per model, grouped by where TestGorilla and CodeSignal stood in it.

TestGorilla first, CodeSignal an alternative

4 of 14 modelsCodeSignal was named in the answer but not as the choice, or not at all.
GPT-5.4 miniTestGorilla alternatives: Codility, HackerRank, Mercer | Mettl
Gemini 3.5 FlashTestGorilla alternatives: Criteria Corp, The Predictive Index, Vervoe
Qwen 3.7 FlashTestGorilla alternatives: CodeSignal, HireVue, Humantic AI
GLM 4.7 FlashXTestGorilla alternatives: Testlify, Vervoe

Neither was the first choice, one was named

5 of 14 modelsThe answer put something else first and named one of the two as an alternative.
Perplexity SonarCriteria Corp alternatives: HireVue, TestGorilla, The Predictive Index
Grok 4.1 FastCriteria Corp alternatives: HireVue, TestGorilla, The Predictive Index
DeepSeek V4 FlashCriteria Corp alternatives: TestGorilla, eSkill
Kimi K2Criteria Corp alternatives: TestGorilla, The Predictive Index, Vervoe, Wonderlic
GPT-6 LunaCriteria Corp alternatives: TestGorilla

Neither was named

5 of 14 modelsThe answer made no first choice from these two in this category.
Claude Haiku 4.5no first choice
Mistral SmallCriteria Corp, Evalart alternatives: Adaface
Llama 4 Maverickno first choice
MiniMax M2.5Evalart, Toggl Hire
Muse Glimmer 30BCriteria Corp alternatives: Evalart

Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.

By buyer segment

The same question asked on behalf of a different buyer. Each standing is computed within its segment and they are never added together. The figures above are the mid-market standing, which is the one the category orders by.
Small business
TestGorilla leads by seventy-three points.
TestGorilla76%#1 of 13
CodeSignal2%#– of 13
The full small business standing →
Mid-marketThe figures above
TestGorilla leads by fifty points.
TestGorilla52%#1 of 12
CodeSignal2%#5 of 12
The full mid-market standing →
Enterprise
TestGorilla leads by nine points.
TestGorilla9%#5 of 11
CodeSignal0%#– of 11
The full enterprise standing →

What the models said about TestGorilla

Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Five of five in this category shown.

“Why Avoid: If you have a limited budget or require responsive technical support, this platform may lead to frustration.” GLM 4.7 FlashX · negative prompt · hard negative
“While widely used and effective for skills testing, some candidates complain about the quality of "Soft Skills" or "Culture Add" multiple-choice questions.” Qwen 3.7 Flash · negative prompt · soft negative
“Some candidate feedback highlights poor question quality (broken English, unclear phrasing), unexpected subscription charges post-use” Grok 4.1 Fast · negative prompt · soft negative
“The Best All-Rounder: TestGorilla ... arguably the best on the market for offering a mix of cognitive ability, culture add, and role-specific skills” Qwen 3.7 Flash · paraphrase prompt · first choice
“For most mid-sized B2B companies, TestGorilla offers the best balance of test breadth, pricing flexibility, and ease of deployment.” DeepSeek V4 Flash · paraphrase prompt · first choice

What the models said about CodeSignal

Every negative label with a quote, up to three, then the highest-weighted positives, up to three. One of one in this category shown.

“The best skills testing tool for hiring for a mid-sized B2B company would be CodeSignal, TestGorilla, or Criteria Corp.” Llama 4 Maverick · paraphrase prompt · first choice
Also compared

Comparisons are drawn for the top eight products in each category, each against each. The output is the models' output; nothing here is a recommendation by the index.