HyperAI超神经

Mathematical Reasoning On Lila Ood

评估指标

Accuracy

评测结果

各个模型在此基准测试上的表现结果

比较表格
模型名称Accuracy
lila-a-unified-benchmark-for-mathematical0.448
lila-a-unified-benchmark-for-mathematical0.268
lila-a-unified-benchmark-for-mathematical0.586
lila-a-unified-benchmark-for-mathematical0.384
lila-a-unified-benchmark-for-mathematical0.177
lila-a-unified-benchmark-for-mathematical0.238