sd-inf commited on
Commit
8a51831
·
verified ·
1 Parent(s): 13d8f04

Create README.md

Browse files
Files changed (1) hide show
  1. README.md +81 -0
README.md ADDED
@@ -0,0 +1,81 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ language:
3
+ - en
4
+ tags:
5
+ - text-generation-inference
6
+ - text-generation
7
+ ---
8
+
9
+ <img src="https://cdn-uploads.huggingface.co/production/uploads/650707344a8839a8bd85ae2f/98wYFnJ1K_aQyfCoULlSz.png" style="height: 500px;">
10
+
11
+ # Synth 2.5 Pro Preview
12
+
13
+
14
+ > [!Note]
15
+ > This is a preview model, it may deliver performance below that of the final model.<br>
16
+ > This preview is a model finetuned with purely **SFT**, RLAIF has **not** been applied yet.<br>
17
+ > Alongside this, creative performance is currently best on **non-thinking mode** for this preview.<br>
18
+
19
+ Synth 2.5 Pro is the next model in the Synth line-up, based on Gemma 4 26BA4B, it has 26B total parameters and 4B active parameters.
20
+
21
+ ### Compared to Synth 2
22
+ Compared to Synth 2, 2.5 has the following changes/improvements:
23
+
24
+ - Trained on real-world creative data/usage from SOTA models (and past-sota, user-perferred models)
25
+
26
+ - Optimized for following user-wanted qualities (described [here](https://lucidity.sh/research/creative.html)) via RLAIF (READ NOTE FOR PREVIEW)
27
+
28
+ - Hybrid reasoning support, for optional further in-depth creative reasoning
29
+
30
+ ### Usage
31
+
32
+ #### Composite and LuciditySH Platform
33
+
34
+ You can test Synth 2.5 Pro at [Composite](https://composite.lucidity.sh/) (creative specific) or LuciditySH Platform (coming soon) for free, with up to 100 free requests a day for easy testing/usage.
35
+
36
+ #### Local
37
+
38
+ You can use Synth 2.5 Pro on Llama.cpp or VLLM
39
+
40
+ #### Generation Parameters
41
+
42
+ It is recommended to use Synth 2.5 Pro with the following generation settings:
43
+
44
+ - Temperature: 0.8-1
45
+ - Top-P: 0.95
46
+ - Top-K: 0
47
+
48
+ ### Training
49
+
50
+ For the SFT stage of Synth 2.5's training, it was trained on a closed dataset comprised of real-world creative interactions with the following models, with the amount of interactions:
51
+
52
+ | Family | Samples |
53
+ |---|---:|
54
+ | Gemini (2.5 Pro) | 1 |
55
+ | Gemini (3.X Pro/3.7) | 1 |
56
+ | DeepSeek (V3 0324) | 5 |
57
+ | DeepSeek (R1 0528) | 8 |
58
+ | DeepSeek (V4 Pro) | 804 |
59
+ | GLM (5.X) | 1981 |
60
+ | GLM (4.X) | 58 |
61
+ | Kimi (k2.X) | 284 |
62
+ | Kimi (k3) | 3 |
63
+ | Minimax M3 | 69 |
64
+ | StepFun | 40 |
65
+ | **total** | **3254** |
66
+
67
+
68
+ <img src="https://cdn-uploads.huggingface.co/production/uploads/650707344a8839a8bd85ae2f/OMfjLL7rzSJQKOZIVL6PR.png" style="height: 350px;">
69
+
70
+ For model replication, close data to our closed dataset can be found in our [PIPKIN datasets](https://huggingface.co/datasets/LucidityAI/PIPKIN-Creative-174k).
71
+
72
+ ### Limitations (for Preview)
73
+
74
+ As per the active parameter count of this model (4B) and the fact that it is only in preview, it comes with the following limitations, as flagged by Composite users:
75
+
76
+ - Synth 2.5 Pro Preview is unstable at higher temperatures (<1)
77
+ - Synth 2.5 Pro Preview may seem predictable, even at a higher temperature (0.8-1)
78
+
79
+ ### Considerations
80
+
81
+ Synth 2.5 Pro is capable of generating content that is harmful, illegal and/or generally NSFW, as with any LLM. It is recommended to put Synth 2.5 behind a moderation model or other safety layer for real world deployment.