HyperAIHyperAI

Command Palette

Search for a command to run...

Performance results of various models on this benchmark

Metrics

Chat
Safety
Overall
Chat-Hard
Reasoning
Length
WR (%)
LC WR (%)
2 rows total