//opengauntlet
← back to the leaderboard

bluestar-27b conversational AI benchmark

bluestar-27b

bluestar-27b benchmark: humanlikeness scores, pairwise rankings, transcripts, latency, and measured local speed.

Leaderboard rank
#36 of 41
Humanlikeness
43.1 / 100
Pairwise rating
1129
Measured speed
4.2 words/sec

Evidence: identical scenario pack and judge protocol; measured and judged results are shown separately below. Read the methodology.

Loading model…