Instructions to use DeepBeepMeep/MingImage with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Diffusion Single File
How to use DeepBeepMeep/MingImage with Diffusion Single File:
# No code snippets available yet for this library. # To use this model, check the repository files and the library's documentation. # Want to help? PRs adding snippets are welcome at: # https://github.com/huggingface/huggingface.js
- Notebooks
- Google Colab
- Kaggle
Update README.md
Browse files
README.md
CHANGED
|
@@ -1,51 +1,25 @@
|
|
| 1 |
---
|
| 2 |
-
license: mit
|
| 3 |
-
library_name: custom
|
| 4 |
-
pipeline_tag: text-to-image
|
| 5 |
tags:
|
| 6 |
-
|
| 7 |
-
|
| 8 |
-
|
| 9 |
-
- rgba
|
| 10 |
---
|
| 11 |
|
| 12 |
-
|
| 13 |
|
| 14 |
-
These are single-file checkpoints prepared for WanGP from
|
| 15 |
-
[inclusionAI/Ming-Image-0.1-Design](https://huggingface.co/inclusionAI/Ming-Image-0.1-Design)
|
| 16 |
-
at revision `208087ada1486931692c1896f38d4cd16ff3df82`.
|
| 17 |
-
The original checkpoint and inference code are MIT licensed. See `LICENSE`.
|
| 18 |
|
| 19 |
-
The
|
| 20 |
-
version at the root, a project-specific BailingMM2 text encoder and its INT8
|
| 21 |
-
ConvRot version under `BailingMM2-Ming-Image/`, and the VAE and runtime
|
| 22 |
-
configuration under `ming_image/`. The tokenizer lives beside the text encoder.
|
| 23 |
-
The INT8 files use WanGP's MMGP ConvRot loader rather than stock Diffusers.
|
| 24 |
|
| 25 |
-
|
| 26 |
-
The connector's upstream FP32 weights were converted to BF16 for the BF16
|
| 27 |
-
inference variant. The original model sources and conversion notes are in
|
| 28 |
-
WanGP's `models/ming_image/` directory.
|
| 29 |
|
| 30 |
-
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 31 |
|
| 32 |
-
|
| 33 |
-
|
| 34 |
-
Design-Layer separates a flattened reference image into ordered RGBA raster layers. WanGP places every layer in the gallery, frontmost first, and saves a ZIP. Supply one reference image and a front-to-back layer plan. Start with 12 steps, guidance 2, and the 1024 working bucket; 512 is faster.
|
| 35 |
-
|
| 36 |
-
Its BF16 and INT8 ConvRot transformer checkpoints are at the repository root. The Layer-specific Bailing encoder checkpoints and tokenizer are under `BailingMM2-Ming-Image-Layer/`, and runtime configs are under `ming_image_layer/`. The Design and Design-Layer checkpoints use the same VAE from `ming_image/`. Their Bailing base weights are identical, but the Layer connector, MLP, and tokenizer differ, so WanGP uses a separate complete Layer encoder.
|
| 37 |
-
|
| 38 |
-
Source: [inclusionAI/Ming-Image-0.1-Design-Layer](https://huggingface.co/inclusionAI/Ming-Image-0.1-Design-Layer), revision `9fabca8b62a67f1f53a957d46389c00451c11e52` (MIT). The released Layer transformer was converted from FP32 to BF16 to match upstream BF16 inference. The merged encoder and ConvRot files were verified against their source tensors and exercised in WanGP generation.
|
| 39 |
-
|
| 40 |
-
Example plan:
|
| 41 |
-
|
| 42 |
-
```text
|
| 43 |
-
Decompose this image into 4 layers with the following specifications:
|
| 44 |
-
Number of layers: 4
|
| 45 |
-
Layer 1: All clearly readable foreground text, preserving its exact wording and placement.
|
| 46 |
-
Layer 2: The card or panel directly behind the text.
|
| 47 |
-
Layer 3: The main foreground subject or illustration.
|
| 48 |
-
Layer 4: The complete background and remaining supporting shapes.
|
| 49 |
-
```
|
| 50 |
-
|
| 51 |
-
The layers are raster images, including any text. The last gallery image is the background.
|
|
|
|
| 1 |
---
|
|
|
|
|
|
|
|
|
|
| 2 |
tags:
|
| 3 |
+
- diffusion-single-file
|
| 4 |
+
base_model:
|
| 5 |
+
- inclusionAI/Ming-Image-0.1-Design
|
|
|
|
| 6 |
---
|
| 7 |
|
| 8 |
+
You will find here all the Ming-Image-0.1-Design models used with WanGP (https://github.com/deepbeepmeep/Wan2GP) :
|
| 9 |
|
|
|
|
|
|
|
|
|
|
|
|
|
| 10 |
|
| 11 |
+
WanGP by DeepBeepMeep : The best Open Source Video Generative Models Accessible to the GPU Poor
|
|
|
|
|
|
|
|
|
|
|
|
|
| 12 |
|
| 13 |
+
WanGP supports the Wan (and derived models), MiniMax H3, Hunyuan Video, Minimax H3, Krea-2, Flux 1 & 2, Qwen Image 1/2.1, Z-Image and LTX-2, LTX Video models with:
|
|
|
|
|
|
|
|
|
|
| 14 |
|
| 15 |
+
Low VRAM requirements (as low as 6 GB of VRAM is sufficient for certain models)
|
| 16 |
+
Support for old GPUs (RTX 10XX, 20xx, ...)
|
| 17 |
+
Very Fast on the latest GPUs
|
| 18 |
+
Easy to use Full Web based interface
|
| 19 |
+
Auto download of the required model adapted to your specific architecture
|
| 20 |
+
Tools integrated to facilitate Video Generation : Mask Editor, Prompt Enhancer, Temporal and Spatial Generation
|
| 21 |
+
Loras Support to customize each model
|
| 22 |
+
Queuing system : make your shopping list of videos to generate and come back later
|
| 23 |
+
Discord Server to get Help from Other Users and show your Best Videos: https://discord.gg/g7efUW9jGV
|
| 24 |
|
| 25 |
+
Follow DeepBeepMeep on Twitter/X to get the Latest News: https://x.com/deepbeepmeep
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|