Instructions to use ProCreations/Image-2.1-Calibrated-FP8 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Diffusers
How to use ProCreations/Image-2.1-Calibrated-FP8 with Diffusers:
pip install -U diffusers transformers accelerate
import torch from diffusers import DiffusionPipeline # switch to "mps" for apple devices pipe = DiffusionPipeline.from_pretrained("ProCreations/Image-2.1-Calibrated-FP8", dtype=torch.bfloat16, device_map="cuda") prompt = "Astronaut in a jungle, cold color palette, muted colors, detailed, 8k" image = pipe(prompt).images[0] - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- Draw Things
- DiffusionBee
Accelerate full 40-step FP8 generation with native precision, measured quality and real-time demo
1081be0 verified Download optimization/baseline-benchmark.json from ProCreations/Image-2.1-Calibrated-FP8: direct link, hf CLI and curl.
- Browser
- Download file 915 Bytes
-
https://huggingface.co/ProCreations/Image-2.1-Calibrated-FP8/resolve/main/optimization/baseline-benchmark.json
- Command line
-
hf download hf://ProCreations/Image-2.1-Calibrated-FP8/optimization/baseline-benchmark.json
-
curl -L -o baseline-benchmark.json https://huggingface.co/ProCreations/Image-2.1-Calibrated-FP8/resolve/main/optimization/baseline-benchmark.json
915 Bytes
| { | |
| "load_seconds": 2.426567278977018, | |
| "torch": "2.14.0+cu130", | |
| "gpu": "NVIDIA RTX PRO 6000 Blackwell Workstation Edition", | |
| "fp8_linears": 224, | |
| "steps": 40, | |
| "cfg": 1, | |
| "extra_quantization": false, | |
| "approximate_cache": false, | |
| "timing": { | |
| "1024": { | |
| "seconds": [ | |
| 6.9087469020159915, | |
| 6.936133738025092 | |
| ], | |
| "mean": 6.9224403200205415, | |
| "warmup_seconds": 8.378067673009355, | |
| "peak_gb": 32.590829568 | |
| }, | |
| "2048": { | |
| "seconds": [ | |
| 44.079785163048655, | |
| 44.1227304089698 | |
| ], | |
| "mean": 44.10125778600923, | |
| "warmup_seconds": 43.912163228029385, | |
| "peak_gb": 53.722432512 | |
| } | |
| }, | |
| "protocol": "CUDA synchronized; batch1; full40steps; includes encoder, denoising and VAE; excludes model load, resolution warmup and file writes. Weights and prefixKVcache unchanged. Compiled mode emulates intermediate precision casts." | |
| } |