A hybrid reasoning model that runs locally in your browser.
Generate images from text prompts
Generate depth video from input video