Running on Zero Agents Featured 2.38k Bark 🐶 2.38k Generate realistic speech and sounds from typed text
Running on Zero Agents Featured 2.03k Chat With Janus-Pro-7B 🌍 2.03k A unified multimodal understanding and generation model.
meta-llama/Llama-3.2-11B-Vision-Instruct Image-Text-to-Text • 11B • Updated Dec 4, 2024 • 69.6k • 1.67k
Running on Zero MCP 82 LLM Agent from an Image 🤖 82 Get a LLM Assistant personality idea from an image
Running on Zero Agents Featured 5.1k MusicGen 🎵 5.1k Generate music from a text description and optional melody
Running Agents Featured 2.08k MagicPrompt Stable Diffusion 😻 2.08k Generate creative Stable Diffusion prompts
Running Agents Featured 118 Compressed Stable Diffusion 🌟 118 Compare image generation results from original and compressed AI models
Running on Zero Agents Featured 326 FABRIC: Personalizing Diffusion Models with Iterative Feedback 🎨 326 Generate images from prompts with feedback guidance