2.4-bit Exllama v2 quant of airoboros-gpt4-1.4.1 70B model

Browse files

Files changed (9) hide show

README.md +47 -0
config.json +26 -0
output-00001-of-00003.safetensors +3 -0
output-00002-of-00003.safetensors +3 -0
output-00003-of-00003.safetensors +3 -0
special_tokens_map.json +23 -0
tokenizer.json +0 -0
tokenizer.model +3 -0
tokenizer_config.json +32 -0

README.md CHANGED Viewed

@@ -1,3 +1,50 @@
 ---
 license: other
 ---

 ---
 license: other
+datasets:
+- jondurbin/airoboros-gpt4-1.4.1
 ---
+### 2.4-bit Exllama v2 quant of airoboros-gpt4-1.4.1
+Simple quantization of original model.  This model should fit on a single 24 GB VRAM GPU where Exlalama v2 is supported.
+Should also support full 4096 context on a single GPU, without dsktop apps also running on the same GPU. Ideally,
+the GPU would be completely empty of any desktop or apps.
+### Overview
+Llama 2 70b fine tune using https://huggingface.co/datasets/jondurbin/airoboros-gpt4-1.4.1
+See the previous llama 65b model card for info:
+https://hf.co/jondurbin/airoboros-65b-gpt4-1.4
+### Contribute
+If you're interested in new functionality, particularly a new "instructor" type to generate a specific type of training data,
+take a look at the dataset generation tool repo: https://github.com/jondurbin/airoboros and either make a PR or open an issue with details.
+To help me with the OpenAI/compute costs:
+- https://bmc.link/jondurbin
+- ETH 0xce914eAFC2fe52FdceE59565Dd92c06f776fcb11
+- BTC bc1qdwuth4vlg8x37ggntlxu5cjfwgmdy5zaa7pswf
+### Licence and usage restrictions
+Base model has a custom Meta license:
+- See the [meta-license/LICENSE.txt](meta-license/LICENSE.txt) file attached for the original license provided by Meta.
+- See also [meta-license/USE_POLICY.md](meta-license/USE_POLICY.md) and [meta-license/Responsible-Use-Guide.pdf](meta-license/Responsible-Use-Guide.pdf), also provided by Meta.
+The fine-tuning data was generated by OpenAI API calls to gpt-4, via [airoboros](https://github.com/jondurbin/airoboros)
+The ToS for OpenAI API usage has a clause preventing the output from being used to train a model that __competes__ with OpenAI
+- what does *compete* actually mean here?
+- these small open source models will not produce output anywhere near the quality of gpt-4, or even gpt-3.5, so I can't imagine this could credibly be considered competing in the first place
+- if someone else uses the dataset to do the same, they wouldn't necessarily be violating the ToS because they didn't call the API, so I don't know how that works
+- the training data used in essentially all large language models includes a significant amount of copyrighted or otherwise non-permissive licensing in the first place
+- other work using the self-instruct method, e.g. the original here: https://github.com/yizhongw/self-instruct released the data and model as apache-2
+I am purposingly leaving this license ambiguous (other than the fact you must comply with the Meta original license for llama-2) because I am not a lawyer and refuse to attempt to interpret all of the terms accordingly.
+Your best bet is probably to avoid using this commercially due to the OpenAI API usage.
+Either way, by using this model, you agree to completely indemnify me.

config.json ADDED Viewed

	@@ -0,0 +1,26 @@

+{
+  "_name_or_path": "airoboros-l2-70b-gpt4-1.4.1-2.4bpw-h6-elx2",
+  "architectures": [
+    "LlamaForCausalLM"
+  ],
+  "bos_token_id": 1,
+  "eos_token_id": 2,
+  "hidden_act": "silu",
+  "hidden_size": 8192,
+  "initializer_range": 0.02,
+  "intermediate_size": 28672,
+  "max_position_embeddings": 4096,
+  "model_type": "llama",
+  "num_attention_heads": 64,
+  "num_hidden_layers": 80,
+  "num_key_value_heads": 8,
+  "pad_token_id": 0,
+  "pretraining_tp": 1,
+  "rms_norm_eps": 1e-05,
+  "rope_scaling": null,
+  "tie_word_embeddings": false,
+  "torch_dtype": "float16",
+  "transformers_version": "4.32.0.dev0",
+  "use_cache": true,
+  "vocab_size": 32000
+}

output-00001-of-00003.safetensors ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:a8dcfcff4a99284dd8ea4ed1349a18fdffac0b31489b8a40f9e9bc71c1c96af4
+size 8563868088

output-00002-of-00003.safetensors ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:b1131f2e3e4cb6f645f1a148271f1e836cb9ed841d6956085bbf4e95def0106b
+size 8572463208

output-00003-of-00003.safetensors ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:dba754cdbd6e1d4f5f36b05233508b53c5af679a9fcd513278b27442f6b04c69
+size 4156160000

special_tokens_map.json ADDED Viewed

	@@ -0,0 +1,23 @@

+{
+  "bos_token": {
+    "content": "<s>",
+    "lstrip": false,
+    "normalized": false,
+    "rstrip": false,
+    "single_word": false
+  },
+  "eos_token": {
+    "content": "</s>",
+    "lstrip": false,
+    "normalized": false,
+    "rstrip": false,
+    "single_word": false
+  },
+  "unk_token": {
+    "content": "<unk>",
+    "lstrip": false,
+    "normalized": false,
+    "rstrip": false,
+    "single_word": false
+  }
+}

tokenizer.json ADDED Viewed

The diff for this file is too large to render. See raw diff

tokenizer.model ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:9e556afd44213b6bd1be2b850ebbbd98f5481437a8021afaf58ee7fb1818d347
+size 499723

tokenizer_config.json ADDED Viewed

	@@ -0,0 +1,32 @@

+{
+  "bos_token": {
+    "__type": "AddedToken",
+    "content": "<s>",
+    "lstrip": false,
+    "normalized": false,
+    "rstrip": false,
+    "single_word": false
+  },
+  "clean_up_tokenization_spaces": false,
+  "eos_token": {
+    "__type": "AddedToken",
+    "content": "</s>",
+    "lstrip": false,
+    "normalized": false,
+    "rstrip": false,
+    "single_word": false
+  },
+  "legacy": false,
+  "model_max_length": 1000000000000000019884624838656,
+  "pad_token": null,
+  "sp_model_kwargs": {},
+  "tokenizer_class": "LlamaTokenizer",
+  "unk_token": {
+    "__type": "AddedToken",
+    "content": "<unk>",
+    "lstrip": false,
+    "normalized": false,
+    "rstrip": false,
+    "single_word": false
+  }
+}