HR AI Index
Index Performance and talent management Performance management › Guide · September 2026 Edition
Guide · September 2026 Edition

Performance management: what twelve AI models recommend, and why, September 2026

Twelve AI models were asked for performance management software six ways each, on behalf of a small, a mid-market and an enterprise B2B company: 216 answers, in which a judge labeled 60 products. Nine of them carry at least 10 labels and a first choice, and are ranked. This guide walks the top eight in the order the mid-market standing puts them, with what the models said for and against each, in their words. Lattice leads at 41% of first choices, a clear leader.

At a glance

The top eight of nine ranked products, mid-market standing. Share is the product's share of first choices; the negative rate is the share of its labels that were cautions or warnings.
ProductFirst-choice shareNegative rateLabelsLeads forQuadrant
1 Lattice41%12%51Mid-marketendorsed leader
2 PerformYard9%7%28noneaccepted challenger
3 15Five7%6%47noneaccepted challenger
4 HiBob7%0%23noneaccepted challenger
5 Small Improvements7%9%11noneaccepted challenger
6 Betterworks4%10%31noneaccepted challenger
7 Leapsome2%6%31noneaccepted challenger
8 Culture Amp2%10%21noneaccepted challenger

The full standing, every product with 10 labels or more, is on the category page, with the products at zero and the ones warned against.

How a product gets on this list

No submission, no payment, no editorial pick.
Named by a model

The judge reads every answer

A product is on the record because a model named it in an answer and the judge labeled the naming: first choice, alternative, mention, caution or warning, each with the sentence it came from. 60 products were named for performance management software this edition.

Enough labels to rank

10 labels and a first choice

A product is ranked once it carries 10 labels in a segment and at least one first choice there. Below that it is named but not ranked, and it says so on its page. The vendor table that folds spellings into products is versioned; this page was scored under vv2026-09.1.

A measured floor

9 points

The same categories are asked again with nothing changed to measure how much an answer moves on its own. A change in share smaller than the floor is not reported as a change. The repeat, model by model.

The six questions

Asked word for word of every model, on behalf of a mid-market B2B company; the other buyers get the same six with the company size changed. These are the evaluation criteria: not a rubric written for this guide, but the questions a buyer asks.
FramingThe question
DirectWhat is the best performance management software for a mid-market B2B company?
ParaphraseWhich performance review tool would you recommend to a mid-sized B2B company?
ComparativeWhat are the top performance management platforms and how do they differ?
Budget-constrainedWhat is the best performance management software for a company with a limited budget?
Scale-constrainedWe are a 500 person company evaluating a performance management software. What should we look at?
NegativeWhich performance management platforms should I avoid or be cautious about?

1. Lattice

Named in nineteen categories this edition.

Lattice is #1 of 9 for the mid-market buyer at 41% of first choices, from 51 labels by 12 of 12 models; 12% of those labels were cautions or warnings. Twelve of twelve models named it first on the direct question. It led the direct, paraphrase, comparative, scale-constrained framings. By buyer: small business #2 of 9 at 17%; mid-market #1 of 9 at 41%; enterprise #3 of 8 at 12%.

5 of 12 models argued against it somewhere in their answers, 1 as a warning.

What the models said for it

“Choose Lattice if you are a fast-growing company (Seed to Series C/D) that wants to manage reviews, promotions, and salary adjustments in one easy-to-use place.” Qwen 3.7 Flash · comparative prompt · first choice
“Lattice: Best Overall for Balanced Performance & Goal Alignment... widely considered the gold-standard platform for mid-market, tech-forward B2B companies” Gemini 3.5 Flash · direct prompt · first choice

And against it

“Watch Out For: Minimum seat requirements (e.g., Lattice requires $4,000/year minimum)” Kimi K2 · budget prompt · hard negative
“Steep learning curve - frequently cited as confusing for initial tasks ... Expensive - minimum $4,000 annual agreement” Kimi K2 · negative prompt · soft negative

2. PerformYard

Named in three categories this edition.

PerformYard is #2 of 9 for the mid-market buyer at 9% of first choices, from 28 labels by 11 of 12 models; 7% of those labels were cautions or warnings. Zero of twelve models named it first on the direct question. By buyer: small business unranked; mid-market #2 of 9 at 9%; enterprise unranked.

What the models said for it

“Leaders in this category typically include Lattice, Culture Amp, 15Five, Leapsome, and PerformYard.” Gemini 3.5 Flash · scale prompt · first choice
“I recommend PerformYard as a top performance review tool” Mistral Small · paraphrase prompt · first choice

And against it

“PerformYard: has limited and complex reporting options.” Llama 4 Maverick · negative prompt · soft negative
“Verdict: Be cautious if you need flexibility.” DeepSeek V4 Flash · negative prompt · soft negative

3. 15Five

Named in eleven categories this edition.

15Five is #3 of 9 for the mid-market buyer at 7% of first choices, from 47 labels by 12 of 12 models; 6% of those labels were cautions or warnings. Zero of twelve models named it first on the direct question. By buyer: small business #6 of 9 at 8%; mid-market #3 of 9 at 7%; enterprise #8 of 8 at 2%.

What the models said for it

“Leaders in this category typically include Lattice, Culture Amp, 15Five, Leapsome, and PerformYard.” Gemini 3.5 Flash · scale prompt · first choice

And against it

“Performance-first tools with limits like Lattice or 15Five | ... you should check for pricing opacity, contract minimums” Perplexity Sonar · negative prompt · soft negative

4. HiBob

Named in sixteen categories this edition.

HiBob is #4 of 9 for the mid-market buyer at 7% of first choices, from 23 labels by 11 of 12 models; 0% of those labels were cautions or warnings. One of twelve models named it first on the direct question. By buyer: small business unranked; mid-market #4 of 9 at 7%; enterprise unranked.

What the models said for it

“Best for mid-sized teams running structured reviews... a smart, modern choice if you want to level up reviews without adding another separate tool” Claude Haiku 4.5 · paraphrase prompt · first choice

And against it

No negative label in this category carried a quote.

5. Small Improvements

Named in two categories this edition.

Small Improvements is #5 of 9 for the mid-market buyer at 7% of first choices, from 11 labels by 10 of 12 models; 9% of those labels were cautions or warnings. Zero of twelve models named it first on the direct question. By buyer: small business #7 of 9 at 6%; mid-market #5 of 9 at 7%; enterprise unranked.

What the models said for it

“The best performance management software for a company with a limited budget is Small Improvements, which is the cheapest credible performance appraisal software” Llama 4 Maverick · budget prompt · first choice
“Often cited as the best value for money, Small Improvements strikes a balance between robust features and affordability.” Qwen 3.7 Flash · budget prompt · first choice

And against it

“Cheapest credible option (but has seat minimums)” Kimi K2 · budget prompt · soft negative

6. Betterworks

Named in seven categories this edition.

Betterworks is #6 of 9 for the mid-market buyer at 4% of first choices, from 31 labels by 12 of 12 models; 10% of those labels were cautions or warnings. One of twelve models named it first on the direct question. By buyer: small business unranked; mid-market #6 of 9 at 4%; enterprise #2 of 8 at 25%.

3 of 12 models argued against it somewhere in their answers, 1 as a warning.

What the models said for it

“The best performance management software for a mid-market B2B company is Lattice, HiBob HRIS, or Betterworks.” Llama 4 Maverick · direct prompt · first choice

And against it

“Betterworks | Extremely expensive, poor integration across multiple systems, unresponsive customer service” GLM 4.7 FlashX · negative prompt · hard negative

7. Leapsome

Named in thirteen categories this edition.

Leapsome is #7 of 9 for the mid-market buyer at 2% of first choices, from 31 labels by 11 of 12 models; 6% of those labels were cautions or warnings. Zero of twelve models named it first on the direct question. By buyer: small business #8 of 9 at 6%; mid-market #7 of 9 at 2%; enterprise #7 of 8 at 2%.

What the models said for it

“Leaders in this category typically include Lattice, Culture Amp, 15Five, Leapsome, and PerformYard.” Gemini 3.5 Flash · scale prompt · first choice

And against it

“Leapsome: has a real admin learning curve and thinner native integrations than peers.” Llama 4 Maverick · negative prompt · soft negative

8. Culture Amp

Named in ten categories this edition.

Culture Amp is #8 of 9 for the mid-market buyer at 2% of first choices, from 21 labels by 11 of 12 models; 10% of those labels were cautions or warnings. Zero of twelve models named it first on the direct question. By buyer: small business unranked; mid-market #8 of 9 at 2%; enterprise unranked.

What the models said for it

No positive label in this category carried a quote.

And against it

“Some users find its reporting and integration capabilities less robust compared to dedicated performance management systems.” Mistral Small · negative prompt · soft negative

Where the models split

Not who to buy, which the index never says, but where the answer depends on how the question was asked and who asked it.

By framing, mid-market buyer

FramingNamed first most oftenThen
Direct Lattice 12 of 12 modelsBetterworks (1), HiBob (1)
Paraphrase Lattice 4 of 12 modelsPerformYard (3), 15Five (2), HiBob (2)
Comparative Lattice 6 of 12 modelsThriveSparrow (1)
Budget-constrained Effy AI 5 of 12 modelsSmall Improvements (3), AssessTEAM (1), Perform (1)
Scale-constrained Lattice 3 of 12 models15Five (1), Culture Amp (1), Leapsome (1)
Negativeno first choice

By buyer

BuyerLeadsThen
Small business Effy AI 19%Lattice, BambooHR, ThriveSparrow
Mid-market Lattice 41%PerformYard, 15Five, HiBob
Enterprise Workday 32%Betterworks, Lattice, SAP SuccessFactors

This category was in the repeat sample: asked again with nothing changed, the leader's share moved 2 points and the leader held. Every category by buyer.

Questions the record answers

Which performance management software do AI models name first most often?

Lattice, in 41% of first choices for a mid-market B2B company in the September 2026 Edition, from 51 labels. The index calls that a clear leader.

Does the answer change with the size of the company?

Small business: Effy AI at 19%. Mid-market: Lattice at 41%. Enterprise: Workday at 32%. Each standing is computed within its segment and never pooled.

Which framing changes the answer?

The direct, paraphrase, budget and scale framings count toward share. The product named first differs by framing: direct Lattice; paraphrase Lattice; comparative Lattice; budget-constrained Effy AI; scale-constrained Lattice. The table above has the counts.

How stable is one answer?

Asked again with nothing changed, the models moved their own first choice 70% of the time across the repeat sample; This category was in the repeat sample: asked again with nothing changed, the leader's share moved 2 points and the leader held.

Where do the answers come from?

Twelve of the twelve models return sources. Their 65 answers here cite 810 pages; the sites cited most are forbes.com, betterworks.com, okrstool.com. The category page lists them all.

Can a vendor pay to be on this list?

No. A product is on the record because a model named it. A vendor can claim its page, propose corrections to the vendor table and be told when its standing moves; it cannot change a label, a share or a rank, and the publisher's conflicts are disclosed on the method page.

The rest of the record on performance management

Every answer, every label and its evidence quote are in the free record. Published under CC BY 4.0. The output is the models' output; nothing here is a recommendation by the index.