//opengauntlet
← back to the leaderboard

Magistry 24B v1.1 (Q4_K_M) conversational AI benchmark

magistry-24b-v1-1-q4

Magistry 24B v1.1 (Q4_K_M) benchmark: humanlikeness scores, pairwise rankings, transcripts, latency, and measured local speed.

Leaderboard rank
#38 of 41
Humanlikeness
47.7 / 100
Pairwise rating
1065

Evidence: identical scenario pack and judge protocol; measured and judged results are shown separately below. Read the methodology.

Loading model…