HyperAIHyperAI

Command Palette

Search for a command to run...

このベンチマークにおける各種モデルの性能結果

メトリクス

Type
Health
Pluralistic
ChatGPT
Gemma-7B
LLaMA2-7B
LLaMA3-8B
LLaMA2-13B
LLaMA2-70B
Qwen2.5-7B
Qwen2.5-14B
ModPluralal Response
ModPlural Response
4 合計