helide-alpha / README.md
inflatebot's picture
Update README.md
0192afe verified
metadata
base_model:
  - Fizzarolli/L3-8b-Rosier-v1
  - Sao10K/L3-8B-Stheno-v3.2
library_name: transformers
tags:
  - mergekit
  - merge

merge

This is a merge of pre-trained language models created using mergekit.

Merge Details

An experimental merge of the legendary L3-8B-Stheno with Fizzarolli's Rosier. The aim is to improve Stheno's "ball-rolling" capabilities and reduce its awkwardness with more niche content. For a first go, I'm surprised at how well it's doing so far, but given that this is literally my first LLM project ever, probably temper your expectations.

Merge Method

This model was merged using the linear merge method.

Models Merged

The following models were included in the merge:

Configuration

The following YAML configuration was used to produce this model:

models:
  - model: Sao10K/L3-8B-Stheno-v3.2
    parameters:
      weight: 0.5
  - model: Fizzarolli/L3-8b-Rosier-v1
    parameters:
      weight: 0.5

merge_method: linear
parameters:
  normalize: true
dtype: float16