| Rank | Model |
Accuracy |
Paper | Code | Year |
| 1 |
PaLM 2 (few-shot, k=3, CoT) |
100 |
PaLM 2 Technical Report
|
eternityyw/tram-benchmark |
2023 |
| 2 |
PaLM 2 (few-shot, k=3, Direct) |
96.4 |
PaLM 2 Technical Report
|
eternityyw/tram-benchmark |
2023 |
| 3 |
PaLM 540B (few-shot, k=3) |
39.6 |
BloombergGPT: A Large Language Model for Finance
|
yangletliu/finlora · open-finance-lab/finlora |
2023 |
| 4 |
BLOOM 176B (few-shot, k=3) |
36.8 |
BloombergGPT: A Large Language Model for Finance
|
yangletliu/finlora · open-finance-lab/finlora |
2023 |
| 5 |
Chinchilla-70B (few-shot, k=5) |
32.0 |
Training Compute-Optimal Large Language Models
|
karpathy/llama2.c · nkluge-correa/teenytinyllama |
2022 |
| 6 |
Bloomberg GPT (few-shot, k=3) |
29.2 |
BloombergGPT: A Large Language Model for Finance
|
yangletliu/finlora · open-finance-lab/finlora |
2023 |
| 7 |
OPT 66B (few-shot, k=3) |
23.6 |
BloombergGPT: A Large Language Model for Finance
|
yangletliu/finlora · open-finance-lab/finlora |
2023 |
| 8 |
GPT-NeoX (few-shot, k=3) |
21.2 |
BloombergGPT: A Large Language Model for Finance
|
yangletliu/finlora · open-finance-lab/finlora |
2023 |
| 9 |
Gopher-280B (few-shot, k=5) |
19.0 |
Scaling Language Models: Methods, Analysis & Insights from Training Gopher
|
allenai/dolma · rvlopes/gloria · bramiozo/PubScience |
2021 |