merge

This is a merge of pre-trained language models created using mergekit.

Merge Details

An experimental merge of the legendary L3-8B-Stheno with Fizzarolli's Rosier. The aim is to improve Stheno's "ball-rolling" capabilities and reduce its awkwardness with more niche content. For a first go, I'm surprised at how well it's doing so far, but given that this is literally my first LLM project ever, probably temper your expectations.

Merge Method

This model was merged using the linear merge method.

Models Merged

The following models were included in the merge:

Configuration

The following YAML configuration was used to produce this model:

models:
  - model: Sao10K/L3-8B-Stheno-v3.2
    parameters:
      weight: 0.5
  - model: Fizzarolli/L3-8b-Rosier-v1
    parameters:
      weight: 0.5

merge_method: linear
parameters:
  normalize: true
dtype: float16

Downloads last month: 2

Safetensors

Model size

8B params

Tensor type

F16

Model tree for inflatebot/helide-alpha

Fizzarolli/L3-8b-Rosier-v1

Sao10K/L3-8B-Stheno-v3.2

Merge model

this model

Paper for inflatebot/helide-alpha

Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference time

Paper • 2203.05482 • Published Mar 10, 2022 • 8