Bonsai 27B WebGPU Kernels
Run a 1-bit 27B LLM locally in your browser on WebGPU
Run a 1-bit 27B LLM locally in your browser on WebGPU
Resize images for visual token budgets while keeping aspect ratio
Unified audio-text intelligence
Identity-preserving instruction image editing on Krea 2
Extend images into larger canvases with Krea 2 outpaint
Stream structured Markdown from document images and PDFs.
Realtime VLM for image and video understanding
Distilled LTX-2.3 identity video from a reference photo
generate a video from an image with a text prompt
Run a 1-bit 27B LLM locally in your browser on WebGPU
Demo of the Collection of Qwen Image Edit LoRAs
Reproduce every ICML 2026 paper with your agent
Extract text from images and PDFs instantly
Voice chat over WebSocket against a HF speech-to-speech
Efficient native-resolution image generation and editing
Identity-preserving instruction image editing on Krea 2
Image edit, text to image, image upscale, remove watermark
One-click model liberation + chat playground
WAN2.2 based I2V
Extend images into larger canvases with Krea 2 outpaint
Use multiple FLUX.2-Klein LoRAs
generate a video from an image with a text prompt
Run complete 3.96M and 9.36M text-to-waveform models live.
Open agentic retriever for hard multi-step search
High-fidelity 3D Generation from images
Talk to Gemma 4 face to face, with a 3D lip-synced avatar
Animate an image into a video with custom prompts
Generate vivid images from text prompts in seconds
Fast Wan 2.2 image-to-video with first/last frames
ltx 2.3 improved image-to-video with 10eros & native audio
Stream structured Markdown from document images and PDFs.
Unified audio-text intelligence