Update README.md
Browse files
README.md
CHANGED
|
@@ -52,7 +52,7 @@ These checkpoints are released to support continued pretraining, fine-tuning, an
|
|
| 52 |
<img src="https://intranetproxy.alipay.com/skylark/lark/0/2026/png/62256938/1787120625910-bb0c32bd-7e6b-4354-ac65-fa5b747d8bff.png" width="968.5" title="" crop="0,0,1,1" id="ud5147f99" class="ne-image">
|
| 53 |
|
| 54 |
## Base Model Evaluation
|
| 55 |
-
To systematically assess the capabilities of the base model, we use a comprehensive benchmark suite covering several key domains, including knowledge, coding, mathematics, reasoning, multilingual understanding, and long-context comprehension. The performance of the pretrained base checkpoint, i.e., `Ling-3.0-tiny-base`, is compared below:
|
| 56 |
|
| 57 |
<img src="https://intranetproxy.alipay.com/skylark/lark/0/2026/png/62256938/1787145422924-5c03d4b9-ee56-4d84-8927-58edb24b24c1.png" width="1578.5" title="" crop="0,0,1,1" id="u9090be40" class="ne-image">
|
| 58 |
|
|
|
|
| 52 |
<img src="https://intranetproxy.alipay.com/skylark/lark/0/2026/png/62256938/1787120625910-bb0c32bd-7e6b-4354-ac65-fa5b747d8bff.png" width="968.5" title="" crop="0,0,1,1" id="ud5147f99" class="ne-image">
|
| 53 |
|
| 54 |
## Base Model Evaluation
|
| 55 |
+
To systematically assess the capabilities of the base model, we use a self-built comprehensive benchmark suite covering several key domains, including knowledge, coding, mathematics, reasoning, multilingual understanding, and long-context comprehension. The performance of the pretrained base checkpoint, i.e., `Ling-3.0-tiny-base`, is compared below:
|
| 56 |
|
| 57 |
<img src="https://intranetproxy.alipay.com/skylark/lark/0/2026/png/62256938/1787145422924-5c03d4b9-ee56-4d84-8927-58edb24b24c1.png" width="1578.5" title="" crop="0,0,1,1" id="u9090be40" class="ne-image">
|
| 58 |
|