Update README.md
Browse files
README.md
CHANGED
|
@@ -24,11 +24,10 @@ base_model:
|
|
| 24 |
|
| 25 |
ProX models are evaluated on 9 common math reasoning benchmarks.
|
| 26 |
|
| 27 |
-
| Model
|
| 28 |
-
|
| 29 |
-
| CodeLlama-7B
|
| 30 |
-
| CodeLlama-7B-ProXMath
|
| 31 |
-
|
| 32 |
|
| 33 |
### Citation
|
| 34 |
```
|
|
|
|
| 24 |
|
| 25 |
ProX models are evaluated on 9 common math reasoning benchmarks.
|
| 26 |
|
| 27 |
+
| Model | asdiv | gsm8k | mathqa | mawps | minerva_math | mmlu_stem | sat_math | svamp | tabmwp | average |
|
| 28 |
+
|-----------------------|:--------:|:--------:|:--------:|:--------:|:------------:|:---------:|:--------:|:--------:|:--------:|:--------:|
|
| 29 |
+
| CodeLlama-7B | 50.7 | 11.8 | 14.3 | 62.6 | 5.0 | 20.4 | 21.9 | 44.2 | 30.6 | 29.1 |
|
| 30 |
+
| CodeLlama-7B-ProXMath | **67.9** | **35.6** | **38.9** | **82.7** | **17.6** | **42.6** | **62.5** | **55.8** | **41.3** | **49.4** |
|
|
|
|
| 31 |
|
| 32 |
### Citation
|
| 33 |
```
|