Svd Keyframe Interpolation
🐨
66
Generate a smooth video between two keyframe images
Replace objects in images using prompts or reference images
Generate depth map from a single image
Segment objects in images using text prompts or scribbles
Predict depth map from a single image
Generate audio from text using VITS model
Generate anime character speech in English, Chinese, and Japanese
Restore and enhance faces in photos with optional upscaling
Transcribe audio to text with speaker diarization