anurag051194 commited on
Commit
f399275
·
verified ·
1 Parent(s): 87464e0

Update TwIL-LM3: weights, tokenizer and model card

Browse files
Files changed (2) hide show
  1. README.md +2 -0
  2. benchmarks.jpg +0 -0
README.md CHANGED
@@ -31,6 +31,8 @@ It improves in-domain formal-logic performance by **+26% relative** over its bas
31
  (+0.022 core average). It is the only arm in this project that gains on both tracks, which is
32
  why it is the recommended release of the pair.
33
 
 
 
34
  ## Results
35
 
36
  ### Track A — in-domain formal logic
 
31
  (+0.022 core average). It is the only arm in this project that gains on both tracks, which is
32
  why it is the recommended release of the pair.
33
 
34
+ ![TwIL-LM3 formal and general reasoning benchmarks against gpt-oss-120b, Qwen3-8B, LFM2-2.6B and Llama-3.2-3B](benchmarks.jpg)
35
+
36
  ## Results
37
 
38
  ### Track A — in-domain formal logic
benchmarks.jpg ADDED