anurag051194 commited on
Commit
76f52bc
·
verified ·
1 Parent(s): ad1f9b6

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +5 -5
README.md CHANGED
@@ -106,9 +106,9 @@ length`, so it measures completed answers rather than raw decode rate.
106
  | **macro gate** | 0.4218 | **0.5896** | 0.3466 † | 0.2925 | 0.3473 | 0.3757 |
107
  | **strict-7** | 0.1971 | **0.3290** | 0.1493 | 0.1229 | 0.1579 | 0.1714 |
108
  | macro_primary | 0.4475 | **0.4958** | 0.4075 | 0.3450 | 0.4188 | 0.4213 |
109
- | tok/s | 15880 | | 15564 | 16160 | 25000 | 22000 |
110
- | mean gen length | **564** | | 999 | 696 | 2296 | 1830 |
111
- | **ans/s** | **28.1** | | 15.6 | 23.2 | 10.9 | 12.0 |
112
 
113
  \* **TwIL-LM3\*** is our latest version of TwIL-LM3. **The weights will be released soon** — the
114
  files in this repository are the current TwIL-LM3 release, not this one. Lanes marked — are not
@@ -162,9 +162,9 @@ more than cancels it.
162
  | math500 | 0.6900 | 0.7000 | 0.4233 | 0.7133 | 0.7800 | 0.6100 | **0.8433** |
163
  | **macro (10 CoT datasets)** | 0.7339 | 0.7193 | 0.6997 | 0.7523 | 0.7884 | 0.8493 | **0.8689** |
164
  | **macro (all 14)** | 0.6694 | 0.6612 | 0.6245 | 0.6814 | 0.7378 | 0.7591 | **0.8086** |
165
- | tok/s | 15880 | 15564 | 16160 | 25000 | 22000 | not measured | 3374 |
166
  | mean gen length | **482** | 626 | 510 | ≈796 | ≈1327 | ≈1931 | 801 |
167
- | **ans/s** | **32.9** | 24.9 | 31.7 | ≈31.4 | ≈16.6 | not measured | 4.2 |
168
 
169
  ‡ MXFP4 weights, tensor-parallel 2 — quantized and multi-GPU, so not directly comparable to the
170
  single-GPU BF16 rows. § 74% of its `rudas_ood` generations hit the length cap, so that cell is a
 
106
  | **macro gate** | 0.4218 | **0.5896** | 0.3466 † | 0.2925 | 0.3473 | 0.3757 |
107
  | **strict-7** | 0.1971 | **0.3290** | 0.1493 | 0.1229 | 0.1579 | 0.1714 |
108
  | macro_primary | 0.4475 | **0.4958** | 0.4075 | 0.3450 | 0.4188 | 0.4213 |
109
+ | tok/s | 15880 | 15840 | 15564 | 16160 | 25230 | 22480 |
110
+ | mean gen length | **564** | 572 | 999 | 696 | 2296 | 1830 |
111
+ | **ans/s** | **28.1** | 27.7 | 15.6 | 23.2 | 10.9 | 12.0 |
112
 
113
  \* **TwIL-LM3\*** is our latest version of TwIL-LM3. **The weights will be released soon** — the
114
  files in this repository are the current TwIL-LM3 release, not this one. Lanes marked — are not
 
162
  | math500 | 0.6900 | 0.7000 | 0.4233 | 0.7133 | 0.7800 | 0.6100 | **0.8433** |
163
  | **macro (10 CoT datasets)** | 0.7339 | 0.7193 | 0.6997 | 0.7523 | 0.7884 | 0.8493 | **0.8689** |
164
  | **macro (all 14)** | 0.6694 | 0.6612 | 0.6245 | 0.6814 | 0.7378 | 0.7591 | **0.8086** |
165
+ | tok/s | 15880 | 15564 | 16160 | 25230 | 22480 | 9420 | 3374 |
166
  | mean gen length | **482** | 626 | 510 | ≈796 | ≈1327 | ≈1931 | 801 |
167
+ | **ans/s** | **32.9** | 24.9 | 31.7 | ≈31.7 | ≈16.9 | 4.9 | 4.2 |
168
 
169
  ‡ MXFP4 weights, tensor-parallel 2 — quantized and multi-GPU, so not directly comparable to the
170
  single-GPU BF16 rows. § 74% of its `rudas_ood` generations hit the length cap, so that cell is a