Instructions to use FudanCVL/SceneDesigner with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Diffusers
How to use FudanCVL/SceneDesigner with Diffusers:
pip install -U diffusers transformers accelerate
import torch from diffusers import DiffusionPipeline # switch to "mps" for apple devices pipe = DiffusionPipeline.from_pretrained("FudanCVL/SceneDesigner", dtype=torch.bfloat16, device_map="cuda") prompt = "Astronaut in a jungle, cold color palette, muted colors, detailed, 8k" image = pipe(prompt).images[0] - Notebooks
- Google Colab
- Kaggle
File size: 2,807 Bytes
ee7c2d1 | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 40 41 42 43 44 45 46 47 48 49 50 51 52 53 54 55 56 57 58 59 60 61 62 63 64 65 66 67 68 69 70 71 72 73 74 75 76 77 78 79 80 81 82 83 84 85 86 87 88 89 90 91 92 | ---
license: apache-2.0
---
<p align="center">
<h2 align="center">SceneDesigner: Controllable Multi-Object Image Generation with 9-DoF Pose Manipulation</h2>
<p align="center">
<a href="https://github.com/qqzy/"><strong>Zhenyuan Qin<sup>*</sup></strong></a>
·
<a href="https://github.com/xinchengshuai/"><strong>Xincheng Shuai<sup>*</sup></strong></a>
·
<a href="https://henghuiding.com/"><strong>Henghui Ding </strong><sup>†</sup></a>
</p>
<p align="center">
Fudan University
</p>
<p align="center">
<a href="https://arxiv.org/abs/2511.16666"><img src="https://img.shields.io/static/v1?label=Paper&message=2511.16666&color=red&logo=arxiv"></a>
<a href="https://henghuiding.com/SceneDesigner/"><img src="https://img.shields.io/static/v1?label=Project%20Page&message=Github&color=blue&logo=github-pages"></a>
<a href="https://huggingface.co/datasets/FudanCVL/ObjectPose9D"><img src="https://img.shields.io/badge/🤗_HuggingFace-Dataset-ffbd45.svg" alt="HuggingFace">
</p>
<!-- ## 🎯 Introduction -->

## ⚙️ Quick Start
### 1. Installation
1. Install Python environment (recommended to use uv)
```bash
uv sync
```
Or alternatively:
```bash
pip install -r requirements.txt
```
2. Install Blender environment
```bash
cd render
python install.py
```
If the automatic installation script fails, you can install manually:
* First download [Blender](https://download.blender.org/release/Blender4.2/) and extract it to the `./render` directory
* Then locate the Blender Python path and install the Python dependencies for Blender, for example:
```bash
cd render
blender-4.2.8-linux-x64/4.2/python/bin/python3.11 -m pip install -r blender_requirements.txt
```
### 2. Download Checkpoints
1. Download the [SceneDesigner](https://huggingface.co/FudanCVL/SceneDesigner) weights to the `checkpoints` directory
2. Download the [Stable Diffusion 3.5](https://huggingface.co/stabilityai/stable-diffusion-3.5-medium) base model weights to the `checkpoints` directory
### 3. Run Demo
Launch the Gradio app:
```bash
python app.py \
--blender_path render/blender/blender \
--device cuda:0 \
--port 7861
```
- Adjust the 9D pose of the cube in the **Cube Controls** panel
- Enter text prompts in the **Generation Config** panel and click the **Generate Images** button to create images
## ✒️ Citation
If you find our work useful for your research and applications, please kindly cite using this BibTeX:
```latex
@inproceedings{SceneDesigner,
title={SceneDesigner: Controllable Multi-Object Image Generation with 9-DoF Pose Manipulation},
author={Qin, Zhenyuan and Shuai, Xincheng and Ding, Henghui},
booktitle={NeurIPS},
year={2025}
}
```
|