|
Download README.md from caccccc/111: direct link, hf CLI and curl.
- Browser
- Download file 1.09 kB
-
https://huggingface.co/caccccc/111/resolve/main/README.md
- Command line
-
hf download hf://caccccc/111/README.md
-
curl -L -o README.md https://huggingface.co/caccccc/111/resolve/main/README.md
1.09 kB
| license: apache-2.0 | |
| base_model: Qwen3.5-9B | |
| pipeline_tag: text-generation | |
| tags: | |
| - text-generation | |
| # 111 | |
| A chat model fine-tuned from **Qwen3.5-9B**. It takes a structured prompt and | |
| returns a strict JSON response. | |
| ## Serving | |
| OpenAI-compatible chat model (served with an inference engine such as SGLang / vLLM): | |
| ```python | |
| from openai import OpenAI | |
| client = OpenAI(base_url="http://127.0.0.1:8000/v1", api_key="EMPTY") | |
| resp = client.chat.completions.create( | |
| model="111", | |
| messages=[ | |
| {"role": "system", "content": system_prompt}, | |
| {"role": "user", "content": user_prompt}, | |
| ], | |
| temperature=0.0, | |
| max_tokens=8192, | |
| ) | |
| print(resp.choices[0].message.content) | |
| ``` | |
| ## Output format | |
| Strict JSON, e.g.: | |
| ```json | |
| {"thought_process": "...", "valid_steps": [1, 2, 5, 8]} | |
| ``` | |
| ## Details | |
| - Base: `Qwen3.5-9B` (`Qwen3_5ForConditionalGeneration`, 32 layers, hidden size 4096). | |
| - Precision: bfloat16. | |
| - Format: safetensors (4 shards) + HF `config.json` / tokenizer. | |
| ## License | |
| Inherits the base-model license. Set the correct license before publishing. | |