SBDO1 commited on
Commit
006d174
·
verified ·
1 Parent(s): e95d0f6

Upload folder using huggingface_hub

Browse files
Files changed (7) hide show
  1. .gitattributes +1 -0
  2. README.md +28 -3
  3. merges.txt +0 -0
  4. sage-gpt-neo-1.3B.gguf +3 -0
  5. tokenizer.json +11 -0
  6. tokenizer_config.json +1 -0
  7. vocab.json +0 -0
.gitattributes CHANGED
@@ -33,3 +33,4 @@ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
33
  *.zip filter=lfs diff=lfs merge=lfs -text
34
  *.zst filter=lfs diff=lfs merge=lfs -text
35
  *tfevents* filter=lfs diff=lfs merge=lfs -text
 
 
33
  *.zip filter=lfs diff=lfs merge=lfs -text
34
  *.zst filter=lfs diff=lfs merge=lfs -text
35
  *tfevents* filter=lfs diff=lfs merge=lfs -text
36
+ sage-gpt-neo-1.3B.gguf filter=lfs diff=lfs merge=lfs -text
README.md CHANGED
@@ -1,3 +1,28 @@
1
- ---
2
- license: mit
3
- ---
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ # SAGE (Q8_0)
2
+
3
+ Sage - GPT-Neo 1.3B fine-tuned on educational data
4
+
5
+ - **Base Model**: EleutherAI/gpt-neo-1.3B
6
+ - **Quantization**: Q8_0
7
+ - **Format**: GGUF (llama.cpp / LM Studio compatible)
8
+ - **Architecture**: GPT-NeoX
9
+
10
+ ## Quick Start
11
+
12
+ ### llama.cpp
13
+ ```bash
14
+ ./main -m sage-q8_0.gguf -p "Your prompt here"
15
+ ```
16
+
17
+ ### LM Studio
18
+ 1. Download `sage-q8_0.gguf`
19
+ 2. Open LM Studio
20
+ 3. Load model and start chatting
21
+
22
+ ### Python
23
+ ```python
24
+ from llama_cpp import Llama
25
+ llm = Llama(model_path="sage-q8_0.gguf")
26
+ output = llm("Hello", max_tokens=128)
27
+ print(output['choices'][0]['text'])
28
+ ```
merges.txt ADDED
The diff for this file is too large to render. See raw diff
 
sage-gpt-neo-1.3B.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:f2dbf14b9121340af6b169afc3bfbdb98363ea25c0fb904482ad3dda19e37d58
3
+ size 4527693823
tokenizer.json ADDED
@@ -0,0 +1,11 @@
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "model": {
3
+ "type": "BPE",
4
+ "prefix": ""
5
+ },
6
+ "vocab": {},
7
+ "merges": [],
8
+ "special_tokens": {
9
+ "<|endoftext|>": 50256
10
+ }
11
+ }
tokenizer_config.json ADDED
@@ -0,0 +1 @@
 
 
1
+ {"unk_token": "<|endoftext|>", "bos_token": "<|endoftext|>", "eos_token": "<|endoftext|>", "add_prefix_space": false, "model_max_length": 2048, "special_tokens_map_file": null, "name_or_path": "gpt2"}
vocab.json ADDED
The diff for this file is too large to render. See raw diff