111 / README.md
caccccc's picture
upload README.md
84191ef verified
|
Raw History Blame Contribute Delete
1.09 kB
metadata
license: apache-2.0
base_model: Qwen3.5-9B
pipeline_tag: text-generation
tags:
  - text-generation

111

A chat model fine-tuned from Qwen3.5-9B. It takes a structured prompt and returns a strict JSON response.

Serving

OpenAI-compatible chat model (served with an inference engine such as SGLang / vLLM):

from openai import OpenAI

client = OpenAI(base_url="http://127.0.0.1:8000/v1", api_key="EMPTY")
resp = client.chat.completions.create(
    model="111",
    messages=[
        {"role": "system", "content": system_prompt},
        {"role": "user", "content": user_prompt},
    ],
    temperature=0.0,
    max_tokens=8192,
)
print(resp.choices[0].message.content)

Output format

Strict JSON, e.g.:

{"thought_process": "...", "valid_steps": [1, 2, 5, 8]}

Details

  • Base: Qwen3.5-9B (Qwen3_5ForConditionalGeneration, 32 layers, hidden size 4096).
  • Precision: bfloat16.
  • Format: safetensors (4 shards) + HF config.json / tokenizer.

License

Inherits the base-model license. Set the correct license before publishing.