Instructions to use DeepBeepMeep/MingImage with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Diffusion Single File
How to use DeepBeepMeep/MingImage with Diffusion Single File:
# No code snippets available yet for this library. # To use this model, check the repository files and the library's documentation. # Want to help? PRs adding snippets are welcome at: # https://github.com/huggingface/huggingface.js
- Notebooks
- Google Colab
- Kaggle
Document Ming Image Design-Layer checkpoints
Browse files
README.md
CHANGED
|
@@ -28,4 +28,24 @@ inference variant. The original model sources and conversion notes are in
|
|
| 28 |
WanGP's `models/ming_image/` directory.
|
| 29 |
|
| 30 |
WanGP supports text-to-image and one-reference image editing with these files.
|
| 31 |
-
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 28 |
WanGP's `models/ming_image/` directory.
|
| 29 |
|
| 30 |
WanGP supports text-to-image and one-reference image editing with these files.
|
| 31 |
+
|
| 32 |
+
## Ming Image 0.1 Design-Layer
|
| 33 |
+
|
| 34 |
+
Design-Layer separates a flattened reference image into ordered RGBA raster layers. WanGP places every layer in the gallery, frontmost first, and saves a ZIP. Supply one reference image and a front-to-back layer plan. Start with 12 steps, guidance 2, and the 1024 working bucket; 512 is faster.
|
| 35 |
+
|
| 36 |
+
Its BF16 and INT8 ConvRot transformer checkpoints are at the repository root. The Layer-specific Bailing encoder checkpoints and tokenizer are under `BailingMM2-Ming-Image-Layer/`, and runtime configs are under `ming_image_layer/`. The Design and Design-Layer checkpoints use the same VAE from `ming_image/`. Their Bailing base weights are identical, but the Layer connector, MLP, and tokenizer differ, so WanGP uses a separate complete Layer encoder.
|
| 37 |
+
|
| 38 |
+
Source: [inclusionAI/Ming-Image-0.1-Design-Layer](https://huggingface.co/inclusionAI/Ming-Image-0.1-Design-Layer), revision `9fabca8b62a67f1f53a957d46389c00451c11e52` (MIT). The released Layer transformer was converted from FP32 to BF16 to match upstream BF16 inference. The merged encoder and ConvRot files were verified against their source tensors and exercised in WanGP generation.
|
| 39 |
+
|
| 40 |
+
Example plan:
|
| 41 |
+
|
| 42 |
+
```text
|
| 43 |
+
Decompose this image into 4 layers with the following specifications:
|
| 44 |
+
Number of layers: 4
|
| 45 |
+
Layer 1: All clearly readable foreground text, preserving its exact wording and placement.
|
| 46 |
+
Layer 2: The card or panel directly behind the text.
|
| 47 |
+
Layer 3: The main foreground subject or illustration.
|
| 48 |
+
Layer 4: The complete background and remaining supporting shapes.
|
| 49 |
+
```
|
| 50 |
+
|
| 51 |
+
The layers are raster images, including any text. The last gallery image is the background.
|