GD-ML/DreamX-World-5B
Image-to-Video • 5B • Updated • 1.07k • 45
Generate high-quality images from text prompts in seconds
Generate speech from text with voice design, cloning, or presets
Multimodal OCR model for complex document understanding.
A Step Towards Music Generation Foundation Model
Music Generation Foundation Model v1.5
Fast high quality video with audio generation