This is the result of an experimental technique called "MedIT Mesh", which is a method for improving the performance of large language models without the data and compute requirements of traditional fine-tuning, by looking only at the model's weights.

I evaluated it on GPQA-diamond +0.015 (0.485 vs 0.470) and IFBench +0.023 (0.620 vs 0.597).

Settings:

  • harness: prime eval
  • generation: temperature: 0.1 max seq: 16384
Downloads last month
2
Safetensors
Model size
3B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for mkurman/LFM2.5-2.6B-MedIT-Mesh

Finetuned
(14)
this model