anurag051194 commited on
Commit
1a29251
·
verified ·
1 Parent(s): dee09d0

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +1 -2
README.md CHANGED
@@ -28,8 +28,7 @@ GRPO reinforcement learning.
28
 
29
  It improves in-domain formal-logic performance by **+26% relative** over its base model
30
  (macro gate 0.336 → 0.422) **and improves held-out benchmark performance at the same time**
31
- (+0.022 core average). It is the only arm in this project that gains on both tracks, which is
32
- why it is the recommended release of the pair.
33
 
34
  <!-- ![TwIL-LM3 formal and general reasoning benchmarks against gpt-oss-120b, Qwen3-8B, LFM2-2.6B and Llama-3.2-3B](benchmarks.jpg) -->
35
 
 
28
 
29
  It improves in-domain formal-logic performance by **+26% relative** over its base model
30
  (macro gate 0.336 → 0.422) **and improves held-out benchmark performance at the same time**
31
+ (+0.022 core average).
 
32
 
33
  <!-- ![TwIL-LM3 formal and general reasoning benchmarks against gpt-oss-120b, Qwen3-8B, LFM2-2.6B and Llama-3.2-3B](benchmarks.jpg) -->
34