Update README.md
Browse files
README.md
CHANGED
|
@@ -52,7 +52,7 @@ These checkpoints are released to support continued pretraining, fine-tuning, an
|
|
| 52 |
<img src="https://intranetproxy.alipay.com/skylark/lark/0/2026/png/62256938/1787121194497-5a39e2a5-8f80-4ed8-81df-3304577bf317.png" width="952" title="" crop="0,0,1,1" id="u730ab9e7" class="ne-image">
|
| 53 |
|
| 54 |
## Base Model Evaluation
|
| 55 |
-
To systematically assess the capabilities of the base model, we use a comprehensive benchmark suite covering several key domains, including mathematics, coding, reasoning, multilingual understanding, and long-context comprehension. The performance of the pretrained base checkpoint, i.e., `Ling-3.0-flash-base`, is compared below:
|
| 56 |
|
| 57 |
<img src="https://intranetproxy.alipay.com/skylark/lark/0/2026/png/62256938/1787147691544-c61cbb0e-568c-4d5e-a2d9-12a709f003d0.png" width="1431.5" title="" crop="0,0,1,1" id="uf61ea7a7" class="ne-image">
|
| 58 |
|
|
|
|
| 52 |
<img src="https://intranetproxy.alipay.com/skylark/lark/0/2026/png/62256938/1787121194497-5a39e2a5-8f80-4ed8-81df-3304577bf317.png" width="952" title="" crop="0,0,1,1" id="u730ab9e7" class="ne-image">
|
| 53 |
|
| 54 |
## Base Model Evaluation
|
| 55 |
+
To systematically assess the capabilities of the base model, we use a self-built comprehensive benchmark suite covering several key domains, including mathematics, coding, reasoning, multilingual understanding, and long-context comprehension. The performance of the pretrained base checkpoint, i.e., `Ling-3.0-flash-base`, is compared below:
|
| 56 |
|
| 57 |
<img src="https://intranetproxy.alipay.com/skylark/lark/0/2026/png/62256938/1787147691544-c61cbb0e-568c-4d5e-a2d9-12a709f003d0.png" width="1431.5" title="" crop="0,0,1,1" id="uf61ea7a7" class="ne-image">
|
| 58 |
|