Edit model card

QuantFactory/llama3.1-cc-8B-GGUF

This is quantized version of nbeerbower/llama3.1-cc-8B created using llama.cpp

Original Model Card

This is an experimental finetune that formats the conversation data sequentially with the Llama 3 template.

Finetuned using an A100 on Google Colab for 3 epochs.

GGUF

Model size

8.03B params

Architecture

llama

2-bit

3-bit

4-bit

5-bit

6-bit

8-bit

Inference API

Unable to determine this model’s pipeline type. Check the docs .

Base model

Finetuned

Finetuned

Quantized

(16)

this model