PowerInfer
/

SmallThinker-21BA3B-Instruct

Text Generation

Model card Files Files and versions Community

yixinsong commited on about 17 hours ago

Commit

5f20f2f

·

verified ·

1 Parent(s): 003b904

Update README.md

Files changed (1) hide show

README.md +8 -1

README.md CHANGED Viewed

@@ -10,7 +10,14 @@ SmallThinker brings powerful, private, and low-latency AI directly to your perso
 without relying on the cloud.
 ## Performance
 For the MMLU evaluation, we use a 0-shot CoT setting.

 without relying on the cloud.
 ## Performance
+| Model                        | MMLU  | GPQA-diamond | MATH-500 | IFEVAL | LIVEBENCH | HUMANEVAL | Average |
+|------------------------------|-------|--------------|----------|--------|-----------|-----------|---------|
+| SmallThinker-21BA3B-Instruct | 84.43 | 55.05        | 82.4     | 85.77  | 60.3      | 89.63     | 76.26   |
+| Gemma3-12b-it                | 78.52 | 34.85        | 82.4     | 74.68  | 44.5      | 82.93     | 66.31   |
+| Qwen3-14B                    | 84.82 | 50           | 84.6     | 85.21  | 59.5      | 88.41     | 75.42   |
+| Qwen3-30BA3B                 | 85.1  | 44.4         | 84.4     | 84.29  | 58.8      | 90.24     | 74.54   |
+| Qwen3-8B                     | 81.79 | 38.89        | 81.6     | 83.92  | 49.5      | 85.9      | 70.26   |
+| Phi-4-14B                    | 84.58 | 55.45        | 80.2     | 63.22  | 42.4      | 87.2      | 68.84   |
 For the MMLU evaluation, we use a 0-shot CoT setting.