Instructions to use PYTHAI/Qwen3.8-Flash-Next-fork with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use PYTHAI/Qwen3.8-Flash-Next-fork with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("image-text-to-text", model="PYTHAI/Qwen3.8-Flash-Next-fork") messages = [ { "role": "user", "content": [ {"type": "image", "url": "https://huggingface.co/datasets/huggingface/documentation-images/resolve/main/p-blog/candy.JPG"}, {"type": "text", "text": "What animal is on the candy?"} ] }, ] pipe(text=messages)# pip install -U transformers accelerate # Load model directly from transformers import AutoProcessor, AutoModelForMultimodalLM processor = AutoProcessor.from_pretrained("PYTHAI/Qwen3.8-Flash-Next-fork") model = AutoModelForMultimodalLM.from_pretrained("PYTHAI/Qwen3.8-Flash-Next-fork", device_map="auto") messages = [ { "role": "user", "content": [ {"type": "image", "url": "https://huggingface.co/datasets/huggingface/documentation-images/resolve/main/p-blog/candy.JPG"}, {"type": "text", "text": "What animal is on the candy?"} ] }, ] inputs = processor.apply_chat_template( messages, add_generation_prompt=True, tokenize=True, return_dict=True, return_tensors="pt", ).to(model.device) outputs = model.generate(**inputs, max_new_tokens=256) print(processor.decode(outputs[0][inputs["input_ids"].shape[-1]:])) - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- vLLM
How to use PYTHAI/Qwen3.8-Flash-Next-fork with vLLM:
Install from pip and serve model
# Install vLLM from pip: pip install vllm # Start the vLLM server: vllm serve "PYTHAI/Qwen3.8-Flash-Next-fork" # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:8000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "PYTHAI/Qwen3.8-Flash-Next-fork", "messages": [ { "role": "user", "content": [ { "type": "text", "text": "Describe this image in one sentence." }, { "type": "image_url", "image_url": { "url": "https://cdn.britannica.com/61/93061-050-99147DCE/Statue-of-Liberty-Island-New-York-Bay.jpg" } } ] } ] }'Use Docker
docker model run hf.co/PYTHAI/Qwen3.8-Flash-Next-fork
- SGLang
How to use PYTHAI/Qwen3.8-Flash-Next-fork with SGLang:
Install from pip and serve model
# Install SGLang from pip: pip install sglang # Start the SGLang server: python3 -m sglang.launch_server \ --model-path "PYTHAI/Qwen3.8-Flash-Next-fork" \ --host 0.0.0.0 \ --port 30000 # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:30000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "PYTHAI/Qwen3.8-Flash-Next-fork", "messages": [ { "role": "user", "content": [ { "type": "text", "text": "Describe this image in one sentence." }, { "type": "image_url", "image_url": { "url": "https://cdn.britannica.com/61/93061-050-99147DCE/Statue-of-Liberty-Island-New-York-Bay.jpg" } } ] } ] }'Use Docker images
docker run --gpus all \ --shm-size 32g \ -p 30000:30000 \ -v ~/.cache/huggingface:/root/.cache/huggingface \ --env "HF_TOKEN=<secret>" \ --ipc=host \ lmsysorg/sglang:latest \ python3 -m sglang.launch_server \ --model-path "PYTHAI/Qwen3.8-Flash-Next-fork" \ --host 0.0.0.0 \ --port 30000 # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:30000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "PYTHAI/Qwen3.8-Flash-Next-fork", "messages": [ { "role": "user", "content": [ { "type": "text", "text": "Describe this image in one sentence." }, { "type": "image_url", "image_url": { "url": "https://cdn.britannica.com/61/93061-050-99147DCE/Statue-of-Liberty-Island-New-York-Bay.jpg" } } ] } ] }' - Docker Model Runner
How to use PYTHAI/Qwen3.8-Flash-Next-fork with Docker Model Runner:
docker model run hf.co/PYTHAI/Qwen3.8-Flash-Next-fork
Download FORK.json from PYTHAI/Qwen3.8-Flash-Next-fork: direct link, hf CLI and curl.
- Browser
- Download file 2.55 kB
-
https://huggingface.co/PYTHAI/Qwen3.8-Flash-Next-fork/resolve/main/FORK.json
- Command line
-
hf download hf://PYTHAI/Qwen3.8-Flash-Next-fork/FORK.json
-
curl -L -o FORK.json https://huggingface.co/PYTHAI/Qwen3.8-Flash-Next-fork/resolve/main/FORK.json
2.55 kB
| { | |
| "kind": "licence-locked pointer fork", | |
| "source_repo": "Qwen/Qwen3.8-Flash-Next", | |
| "source_url": "https://huggingface.co/Qwen/Qwen3.8-Flash-Next", | |
| "source_commit": "de4b8e4d43b917e7706784d8bb445c9af86a3540", | |
| "forked_at_utc": "2026-09-13T23:55:35Z", | |
| "licence_tag": "other", | |
| "licence_file": { | |
| "path": "LICENSE", | |
| "bytes": 3235, | |
| "sha256": "a0dc422560841fd68e06d974907f8b4c709bca44a67daad2b528437bdf676c08" | |
| }, | |
| "weights_copied": false, | |
| "weights_not_copied": { | |
| "files": 131, | |
| "bytes": 360000192888 | |
| }, | |
| "parameters": 179999981459, | |
| "load_weights_with": "from_pretrained('Qwen/Qwen3.8-Flash-Next', revision='de4b8e4d43b917e7706784d8bb445c9af86a3540')", | |
| "files": [ | |
| { | |
| "path": "LICENSE", | |
| "bytes": 3235, | |
| "sha256": "a0dc422560841fd68e06d974907f8b4c709bca44a67daad2b528437bdf676c08" | |
| }, | |
| { | |
| "path": "README.md", | |
| "bytes": 65155, | |
| "sha256": "35ca37ccc366f1ba478dab33841a2c0c18ce53fd62f291ca05341f7728b225b2" | |
| }, | |
| { | |
| "path": "chat_template.jinja", | |
| "bytes": 8952, | |
| "sha256": "c3cf9e34abf4f9e36c2d72165aa9c132d3e2a725b6c2586aaa3a8af9d7a81041" | |
| }, | |
| { | |
| "path": "config.json", | |
| "bytes": 4745, | |
| "sha256": "889658f2508e8c61d409b02e70e0d78d8d4452ec65aaafbe129805d213d2e74b" | |
| }, | |
| { | |
| "path": "generation_config.json", | |
| "bytes": 202, | |
| "sha256": "e70c136c1b78ddc1fb0905bac8e733a4dc448d4f852a5dd75143fffc70be550e" | |
| }, | |
| { | |
| "path": "merges.txt", | |
| "bytes": 3353259, | |
| "sha256": "a9d356d7bdf1ef4949e3e748e95b8e10ad9d4e2e838eddc38a0a7b6b94d1db8d" | |
| }, | |
| { | |
| "path": "model.safetensors.index.json", | |
| "bytes": 170726, | |
| "sha256": "99e815241ef03325536b0aaa4441deea45174c17fae31e10f0bb456410c590de" | |
| }, | |
| { | |
| "path": "preprocessor_config.json", | |
| "bytes": 390, | |
| "sha256": "27225450ac9c6529872ee1924fcb0962ff5634834f817040f444118116f4e516" | |
| }, | |
| { | |
| "path": "tokenizer.json", | |
| "bytes": 12809320, | |
| "sha256": "0997f410c57a1f4e53b09e4be8f4a172d90edd9564368fb0847030937229b9f3" | |
| }, | |
| { | |
| "path": "tokenizer_config.json", | |
| "bytes": 17928, | |
| "sha256": "b11349aafa7cdc6a320767cf7ceb29ed82f7eda5d65e8e0819e76f0ce947bf27" | |
| }, | |
| { | |
| "path": "video_preprocessor_config.json", | |
| "bytes": 385, | |
| "sha256": "7768af27c1fafa9cc9011c1dc20067e03f8915e03b63504550e11d5066986d13" | |
| }, | |
| { | |
| "path": "vocab.json", | |
| "bytes": 6722759, | |
| "sha256": "ce99b4cb2983d118806ce0a8b777a35b093e2000a503ebde25853284c9dfa003" | |
| } | |
| ], | |
| "why": "preserve the licence terms of this exact commit and pin the revision; the weights stay at the source and are addressed by commit sha", | |
| "forked_by": "mindX (PYTHAI) via the node token" | |
| } |