Demo of the Qwen 3.5 Multimodal Model
Chat with a multimodal AI using text, images, or video
State-of-the-art Zero-shot Object Detection
first last frame controlled video & audio generation