Edit model card

THIS IS A WIP. MANAGE YOUR EXPECTATIONS.

By NovelAI

merge

This is a merge of pre-trained language models created using mergekit.

Merge Details

An experimental merge of the legendary L3-8B-Stheno with Fizzarolli's Rosier. The aim is to improve Stheno's "ball-rolling" capabilities and reduce its awkwardness with more niche content. For a first go, I'm surprised at how well it's doing so far, but given that this is literally my first LLM project ever, probably temper your expectations.

Since R2: Changed to task-arithmetic.

Merge Method

This model was merged using the task arithmetic merge method using NousResearch/Meta-Llama-3-8B as a base.

Models Merged

The following models were included in the merge:

Configuration

The following YAML configuration was used to produce this model:

models:
  - model: Sao10K/L3-8B-Stheno-v3.2
    parameters:
      weight: 0.5
  - model: Fizzarolli/L3-8b-Rosier-v1
    parameters:
      weight: 0.5

merge_method: task_arithmetic
base_model: NousResearch/Meta-Llama-3-8B
parameters:
  normalize: true
dtype: float16
Downloads last month
8
Safetensors
Model size
8.03B params
Tensor type
FP16
·
Inference Examples
This model does not have enough activity to be deployed to Inference API (serverless) yet. Increase its social visibility and check back later, or deploy to Inference Endpoints (dedicated) instead.

Model tree for inflatebot/helide-alpha-r2