Logical Reasoning On Ruworldtree
Metrics
Accuracy
Results
Performance results of various models on this benchmark
Comparison Table
Model Name | Accuracy |
---|---|
tape-assessing-few-shot-russian-language | 38.0 |
tape-assessing-few-shot-russian-language | 83.7 |
tape-assessing-few-shot-russian-language | 34.0 |
tape-assessing-few-shot-russian-language | 40.7 |