| Rank | Model |
Accuracy |
Paper | Code | Year |
| 1 |
PaLM 2 (few-shot, k=3, CoT) |
91.2 |
PaLM 2 Technical Report
|
eternityyw/tram-benchmark |
2023 |
| 2 |
PaLM 2 (few-shot, k=3, Direct) |
61.2 |
PaLM 2 Technical Report
|
eternityyw/tram-benchmark |
2023 |
| 3 |
Chinchilla-70B (few-shot, k=5) |
59.7 |
Training Compute-Optimal Large Language Models
|
karpathy/llama2.c · nkluge-correa/teenytinyllama |
2022 |
| 4 |
Gopher-280B (few-shot, k=5) |
49.2 |
Scaling Language Models: Methods, Analysis & Insights from Training Gopher
|
allenai/dolma · rvlopes/gloria · bramiozo/PubScience |
2021 |
| 5 |
PaLM 540B (few-shot, k=3) |
38 |
BloombergGPT: A Large Language Model for Finance
|
yangletliu/finlora · open-finance-lab/finlora |
2023 |
| 6 |
BLOOM 176B (few-shot, k=3) |
36.8 |
BloombergGPT: A Large Language Model for Finance
|
yangletliu/finlora · open-finance-lab/finlora |
2023 |
| 7 |
Bloomberg GPT (few-shot, k=3) |
34.8 |
BloombergGPT: A Large Language Model for Finance
|
yangletliu/finlora · open-finance-lab/finlora |
2023 |
| 8 |
OPT 66B (few-shot, k=3) |
31.2 |
BloombergGPT: A Large Language Model for Finance
|
yangletliu/finlora · open-finance-lab/finlora |
2023 |
| 9 |
GPT-NeoX (few-shot, k=3) |
26 |
BloombergGPT: A Large Language Model for Finance
|
yangletliu/finlora · open-finance-lab/finlora |
2023 |