Ming-Image-0.1-Design is a 6B text-to-image model for UI, infographics, posters, and other text-rich visual designs. It generates complete visual compositions and supports RGBA output with transparent backgrounds.
Use the companion Ming-Image repository for installation and inference:
git clone https://github.com/inclusionAI/Ming-Image
cd Ming-Image
pip install -r requirements.txt
python infer.py \
--model inclusionAI/Ming-Image-0.1-Design \
--task text-to-image \
--prompt assets/t2i_four_seasons_cabin_prompt.json \
--resolution 2048 \
--output-dir outputs/t2i
Prompt enhancement (PE) can use Ling-3.0-flash-VL or qwen3.8-27B; see
text-to-image prompt rewriting.
For transparent-background generation, prepend exactly one of the recommended RGBA phrases. See the transparent-background generation tip.
We recommend the following inference frameworks to serve the model:
The public inference code maps text-to-image resolution requests to the supported 1024 or 2048 bucket.
The checkerboard is used only to preview transparency; it is not part of the generated RGBA images.
This model is released under the MIT License.
5 commits
Ming-Image-0.1-Design is a 6B text-to-image model for UI, infographics, posters, and other text-rich visual designs. It generates complete visual compositions and supports RGBA output with transparent backgrounds.
Use the companion Ming-Image repository for installation and inference:
git clone https://github.com/inclusionAI/Ming-Image
cd Ming-Image
pip install -r requirements.txt
python infer.py \
--model inclusionAI/Ming-Image-0.1-Design \
--task text-to-image \
--prompt assets/t2i_four_seasons_cabin_prompt.json \
--resolution 2048 \
--output-dir outputs/t2i
Prompt enhancement (PE) can use Ling-3.0-flash-VL or qwen3.8-27B; see
text-to-image prompt rewriting.
For transparent-background generation, prepend exactly one of the recommended RGBA phrases. See the transparent-background generation tip.
We recommend the following inference frameworks to serve the model:
The public inference code maps text-to-image resolution requests to the supported 1024 or 2048 bucket.
The checkerboard is used only to preview transparency; it is not part of the generated RGBA images.
This model is released under the MIT License.
5 commits