Bc-AI commited on
Commit
ebbd6b4
·
verified ·
1 Parent(s): 21c9274

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +42 -8
README.md CHANGED
@@ -8,15 +8,49 @@ datasets:
8
  tags:
9
  - full-fine-tune
10
  - sft
 
11
  - qwen3.5
 
12
  ---
13
 
14
- # Qwen3.5-4B — Full SFT
15
 
16
- - **Base model:** `Qwen/Qwen3.5-4B`
17
- - **Dataset:** `Bc-AI/SFT-Ultra`
18
- - **Training type:** Full parameter supervised fine-tuning (no LoRA)
19
- - **Max sequence length:** 2048
20
- - **Max steps:** 2500
21
- - **Learning rate:** 1e-05
22
- - **Precision:** bf16
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
8
  tags:
9
  - full-fine-tune
10
  - sft
11
+ - dpo
12
  - qwen3.5
13
+ - smilyai
14
  ---
15
 
16
+ # SmilyAI Labs T1-Mini-Preview
17
 
18
+ **T1-Mini-Preview** is an early preview of the upcoming **T1-Mini** model from **SmilyAI Labs**.
19
+
20
+ T1-Mini-Preview is based on **Qwen/Qwen3.5-4B** and was fully fine-tuned on **Bc-AI/SFT-Ultra**, a curated instruction-tuning dataset prepared for SmilyAI's small-model research.
21
+
22
+ ## Training
23
+
24
+ The model was trained in two main stages:
25
+
26
+ 1. **Supervised Fine-Tuning (SFT)** — trained the base model on our curated instruction and reasoning data.
27
+ 2. **Direct Preference Optimization (DPO)** — further refined the model's responses using preference-based training.
28
+
29
+ This two-stage pipeline was designed to improve instruction following, response quality, and overall conversational behavior while keeping the model relatively small and efficient.
30
+
31
+ ## Model Status
32
+
33
+ **T1-Mini-Preview is a preview release**, not the final T1-Mini model. Training, evaluation, and further refinement are still ongoing.
34
+
35
+ We are releasing this version so the community can experiment with it and provide feedback while development continues.
36
+
37
+ ## Base Model
38
+
39
+ - **Base:** Qwen/Qwen3.5-4B
40
+ - **Training:** Full fine-tuning
41
+ - **Primary language:** English
42
+ - **Fine-tuning dataset:** Bc-AI/SFT-Ultra
43
+ - **Training stages:** SFT → DPO
44
+ - **License:** Apache 2.0
45
+
46
+ ## About SmilyAI Labs
47
+
48
+ **SmilyAI Labs** is a small open-source AI project focused on building capable, efficient, and accessible AI models.
49
+
50
+ We're experimenting with smaller models that can deliver strong performance without requiring enormous amounts of compute.
51
+
52
+ 🚀 **T1-Mini-Preview is one step toward that goal.**
53
+
54
+ ## Disclaimer
55
+
56
+ This is an experimental preview model. Its behavior and capabilities may differ from the final T1-Mini release, and it may occasionally produce incorrect, inconsistent, or undesirable outputs.