Zero of twelve models named Traliant first on the direct prompt; eight named BizLibrary. Traliant was named by twelve of the twelve models and BizLibrary by ten and Traliant carries 29 labels and BizLibrary 13, so the shares are not directly comparable.
Named in one category this edition.
Named in two categories this edition.
Share is the count of first choices across the direct, paraphrase, budget and scale prompts over all twelve models, for a mid-market B2B company; rank is within the category; every quote names the model and the prompt it came from. Both figures come from the compliance training page.
Bold names in an answer are the products the judge labeled a first choice; a model naming several gives each of them that label. The full answer text for every row is in the record.
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of four in this category shown.
“Choose Traliant if you want: better learner engagement, workplace conduct/harassment-focused content, a training-first solution.” GPT-5.4 mini · comparative prompt · first choice
“I'd most often recommend Traliant ... Traliant is the safest default for a mid-sized B2B company” Perplexity Sonar · paraphrase prompt · first choice
“Start with a demo request for Traliant (for engagement) and NetExpedition (for legal rigor)” Qwen 3.7 Flash · paraphrase prompt · first choice
Every negative label with a quote, up to three, then the highest-weighted positives, up to three. Three of four in this category shown.
“The Best All-Rounder: BizLibrary ... It strikes the strongest balance between high-quality content and ease of use.” Qwen 3.7 Flash · direct prompt · first choice
“BizLibrary – Strong all-around for mid-sized (100–2,500 employees); combines compliance + upskilling, highly rated” DeepSeek V4 Flash · scale prompt · first choice
“BizLibrary is widely considered the best compliance training platform for mid-market B2B companies” MiniMax M2.5 · direct prompt · first choice
Comparisons are drawn for the top three products in each category. The output is the models' output; nothing here is a recommendation by the index.