111 / README.md
caccccc's picture
upload README.md
84191ef verified
|
Raw History Blame Contribute Delete
1.09 kB
---
license: apache-2.0
base_model: Qwen3.5-9B
pipeline_tag: text-generation
tags:
- text-generation
---
# 111
A chat model fine-tuned from **Qwen3.5-9B**. It takes a structured prompt and
returns a strict JSON response.
## Serving
OpenAI-compatible chat model (served with an inference engine such as SGLang / vLLM):
```python
from openai import OpenAI
client = OpenAI(base_url="http://127.0.0.1:8000/v1", api_key="EMPTY")
resp = client.chat.completions.create(
model="111",
messages=[
{"role": "system", "content": system_prompt},
{"role": "user", "content": user_prompt},
],
temperature=0.0,
max_tokens=8192,
)
print(resp.choices[0].message.content)
```
## Output format
Strict JSON, e.g.:
```json
{"thought_process": "...", "valid_steps": [1, 2, 5, 8]}
```
## Details
- Base: `Qwen3.5-9B` (`Qwen3_5ForConditionalGeneration`, 32 layers, hidden size 4096).
- Precision: bfloat16.
- Format: safetensors (4 shards) + HF `config.json` / tokenizer.
## License
Inherits the base-model license. Set the correct license before publishing.