Add model files and README

Files changed (6) hide show

README.md ADDED Viewed

+---
+datasets:
+- Squish42/bluemoon-fandom-1-1-rp-cleaned
+language:
+- en
+---
+## General
+Bluemoon roleplay finetune of LLaMA 33B (2 roleplayers only). This release also tests a longer 4k context token size achieved with AliBi.
+## Models
+*GGML 4-bit for llama.cpp*<br/>
+1. ggml-bluemoonrp-30b-4k-epoch6-q5_0.bin
+*GPTQ 4-bit CUDA:*<br/>
+1. bluemoonrp-30b-4k-epoch6-4bit-128g.safetensors
+## Remarks
+This model has been trained using the following prompt (Vicuna 1.1 format):
+```
+A transcript of a roleplay between two players, LEAD and ASSOCIATE. LEAD sets up a scenario and the characters, from which ASSOCIATE then assumes a character role and continues the story for that role in response to description given by LEAD. The story and characters are developed by exchange of detailed event descriptions and character dialogs, successively given by both LEAD and ASSOCIATE.
+LEAD: [role1 message]
+ASSOCIATE: [role2 message]</s>
+```

config.json ADDED Viewed

+{
+  "architectures": [
+    "LlamaForCausalLM"
+  ],
+  "bos_token_id": 1,
+  "eos_token_id": 2,
+  "hidden_act": "silu",
+  "hidden_size": 6656,
+  "initializer_range": 0.02,
+  "intermediate_size": 17920,
+  "max_position_embeddings": 2048,
+  "max_seq_len": 4096,
+  "model_type": "llama",
+  "num_attention_heads": 52,
+  "num_hidden_layers": 60,
+  "pad_token_id": 0,
+  "rms_norm_eps": 1e-06,
+  "tie_word_embeddings": false,
+  "torch_dtype": "bfloat16",
+  "transformers_version": "4.28.0.dev0",
+  "use_cache": true,
+  "vocab_size": 32000
+}

ggml-bluemoonrp-30b-4k-epoch6-q5_0.bin ADDED Viewed

+version https://git-lfs.github.com/spec/v1
+oid sha256:2b7ade7fb1abba1a478ba3fdc19526b79d87d11accf41b6d59caf966cd8f3718
+size 22366783872

special_tokens_map.json ADDED Viewed

+{
+  "bos_token": {
+    "content": "<s>",
+    "lstrip": false,
+    "normalized": true,
+    "rstrip": false,
+    "single_word": false
+  },
+  "eos_token": {
+    "content": "</s>",
+    "lstrip": false,
+    "normalized": true,
+    "rstrip": false,
+    "single_word": false
+  },
+  "pad_token": "<unk>",
+  "unk_token": {
+    "content": "<unk>",
+    "lstrip": false,
+    "normalized": true,
+    "rstrip": false,
+    "single_word": false
+  }
+}

tokenizer.model ADDED Viewed

+version https://git-lfs.github.com/spec/v1
+oid sha256:9e556afd44213b6bd1be2b850ebbbd98f5481437a8021afaf58ee7fb1818d347
+size 499723

tokenizer_config.json ADDED Viewed

+{
+  "add_bos_token": true,
+  "add_eos_token": false,
+  "bos_token": {
+    "__type": "AddedToken",
+    "content": "<s>",
+    "lstrip": false,
+    "normalized": true,
+    "rstrip": false,
+    "single_word": false
+  },
+  "clean_up_tokenization_spaces": false,
+  "eos_token": {
+    "__type": "AddedToken",
+    "content": "</s>",
+    "lstrip": false,
+    "normalized": true,
+    "rstrip": false,
+    "single_word": false
+  },
+  "model_max_length": 4096,
+  "pad_token": null,
+  "padding_side": "right",
+  "sp_model_kwargs": {},
+  "special_tokens_map_file": "special_tokens_map.json",
+  "tokenizer_class": "LlamaTokenizer",
+  "unk_token": {
+    "__type": "AddedToken",
+    "content": "<unk>",
+    "lstrip": false,
+    "normalized": true,
+    "rstrip": false,
+    "single_word": false
+  }
+}