HR AI Index
Index Compensation and total rewards Flexible benefits › Benepass vs JOON
Flexible benefits and lifestyle spending accounts · September 2026 Edition

Benepass vs JOON

Six of twelve models named Benepass first on the direct prompt; zero named JOON. Benepass was named by twelve of the twelve models and JOON by nine and Benepass carries 42 labels and JOON 10, so the shares are not directly comparable.

Benepass

accepted challenger

Named in three categories this edition.

JOON

accepted challenger

Named in one category this edition.

First-choice share17%11%Of first choices across the direct, paraphrase, budget and scale prompts, 0 to 100.
Negative rate7%0%Negative labels as a share of the product's labels, 0 to 100.
Rank in category#3#4A position in a field of 6; printed, not drawn.
Labels4210A count; the two differ.
The two percentage rows are drawn on one 0 to 100 track, Benepass reading right to left. Rank and label count are printed, not drawn.Compt was named alongside these two in twelve of the twelve direct answers. Compt vs Benepass · Compt vs JOON · Forma vs Benepass

Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all twelve models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the flexible benefits and lifestyle spending accounts page.

By framing

How many of the twelve models made each the first choice, per way of asking, and how many argued against it.
BenepassFirst choices, of twelve modelsJOON
Direct60
Paraphrase10
Comparative20
Budget-constrained052 against Benepass
Scale-constrained10
Negative001 against Benepass
Bars are first choices, 0 to 12 each sideModels that argued againstA model can name both, so the two sides of a row do not sum to twelve.

Across every category in the September 2026 Edition, Benepass and JOON were named in the same answer twenty-one times, of the 134 answers naming Benepass and the 32 naming JOON. In those answers JOON took the first choice eight times and Benepass zero.

Every model, every framing

The seventy-two answers behind the chart above, one cell each: where Benepass and JOON stood in it.
ModelDirectParaphraseComparativeBudget-constrainedScale-constrainedNegative
Claude Haiku 4.5
GPT-5.4 mini
Gemini 3.5 Flash
Perplexity Sonar
Grok 4.1 Fast
Mistral Small
DeepSeek V4 Flash
Llama 4 Maverick
Qwen 3.7 Flash
Kimi K2
GLM 4.7 FlashX
MiniMax M2.5
Benepass JOON first choice named as an alternative argued againstblank: not namedEach cell is one answer, Benepass on the left and JOON on the right.

The direct prompt

The plain question, one answer per model, grouped by where Benepass and JOON stood in it.

Benepass first, JOON not the choice

6 of 12 modelsJOON was named in the answer but not as the choice, or not at all.
Mistral SmallBenepass alternatives: Compt, Forma
DeepSeek V4 FlashBenepass alternatives: Compt, Forma
Llama 4 MaverickBenepass, Compt, Guideflow
Qwen 3.7 FlashBenepass, Compt alternatives: Favorably
Kimi K2Benepass alternatives: Compt, Forma
MiniMax M2.5Benepass alternatives: Compt, Forma

Neither was the first choice, one was named

5 of 12 modelsThe answer put something else first and named one of the two as an alternative.
Claude Haiku 4.5Compt alternatives: Benepass, Forma
Gemini 3.5 FlashForma alternatives: Benepass, Compt, JOON
Perplexity SonarCompt alternatives: Benepass, Forma
Grok 4.1 FastForma alternatives: Benepass, Compt
GLM 4.7 FlashXCompt, Forma alternatives: Benepass, Espresa, ThrivePass

Neither was named

1 of 12 modelsThe answer made no first choice from these two in this category.
GPT-5.4 miniCompt, Forma alternatives: ThrivePass

Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.

By buyer segment

The same question asked on behalf of a different buyer. Each standing is computed within its segment and they are never added together. The figures above are the mid-market standing, which is the one the category orders by.
Small business
JOON leads by two points.
JOON17%#2 of 7
Benepass15%#3 of 7
The full small business standing →
Mid-marketThe figures above
The order flips: Benepass leads at mid-market.
Benepass17%#3 of 6
JOON11%#4 of 6
The full mid-market standing →
Enterprise
Benepass leads by eleven points.
Benepass11%#3 of 8
JOON0%#– of 8
The full enterprise standing →

What the models said about Benepass

Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Six of six in this category shown.

“Avoid higher-end ones like Benepass/Forma/Espresa” Grok 4.1 Fast · budget prompt · hard negative
“G2 reviewers have highlighted significant frustrations regarding the clarity of reimbursement policies” Qwen 3.7 Flash · negative prompt · soft negative
“Platform fees can add up quickly for small teams” DeepSeek V4 Flash · budget prompt · soft negative
“Benepass is often the top choice due to its balance of card-based convenience, compliance features, global scalability” MiniMax M2.5 · direct prompt · first choice
“Prioritize platforms with proven mid-market success (e.g., Compt for reimbursements, Benepass for cards).” Grok 4.1 Fast · scale prompt · first choice
“Benepass (Best Overall for Employee Experience & Distributed Teams) ... Benepass is a top-tier choice” Gemini 3.5 Flash · paraphrase prompt · first choice

What the models said about JOON

Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of four in this category shown.

“For a small team on a strict budget: JOON ($250/month flat) is the most straightforward and predictable option.” GLM 4.7 FlashX · budget prompt · first choice
“Joon and Compt are top choices, with Joon being the most affordable starting at $250/month” Mistral Small · budget prompt · first choice
“I'd suggest starting with Joon (simplest, lowest entry cost) or Compt” Kimi K2 · budget prompt · first choice
Also compared

Comparisons are drawn for the top eight products in each category, each against each. The output is the models' output; nothing here is a recommendation by the index.