Text-to-Audio
Diffusers
Safetensors
MossSoundEffectPipeline
sound-effects
foley
flow-matching
diffusion-transformer
dac-vae
english
chinese
moss
mirror
Instructions to use AEmotionStudio/moss-soundeffect-models with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Diffusers
How to use AEmotionStudio/moss-soundeffect-models with Diffusers:
pip install -U diffusers transformers accelerate
import torch from diffusers import DiffusionPipeline # switch to "mps" for apple devices pipe = DiffusionPipeline.from_pretrained("AEmotionStudio/moss-soundeffect-models", dtype=torch.bfloat16, device_map="cuda") prompt = "Astronaut in a jungle, cold color palette, muted colors, detailed, 8k" image = pipe(prompt).images[0] - Notebooks
- Google Colab
- Kaggle
Download vae/config.json from AEmotionStudio/moss-soundeffect-models: direct link, hf CLI and curl.
- Browser
- Download file 661 Bytes
-
https://huggingface.co/AEmotionStudio/moss-soundeffect-models/resolve/main/vae/config.json
- Command line
-
hf download hf://AEmotionStudio/moss-soundeffect-models/vae/config.json
-
curl -L -o config.json https://huggingface.co/AEmotionStudio/moss-soundeffect-models/resolve/main/vae/config.json
661 Bytes
| { | |
| "_class_name": "DAC", | |
| "type": "dac", | |
| "latent_dim": 128, | |
| "sample_rate": 48000, | |
| "note": "Converted from vae_128d_48k.pth (audiotools package). Load with DAC(**kwargs) + safetensors state dict \u2014 see MAESTRO's vendored diffsynth/pipelines/wan_audio.py.", | |
| "kwargs": { | |
| "encoder_dim": 128, | |
| "encoder_rates": [ | |
| 2, | |
| 3, | |
| 4, | |
| 5, | |
| 8 | |
| ], | |
| "latent_dim": 128, | |
| "decoder_dim": 2048, | |
| "decoder_rates": [ | |
| 8, | |
| 5, | |
| 4, | |
| 3, | |
| 2 | |
| ], | |
| "n_codebooks": 9, | |
| "codebook_size": 1024, | |
| "codebook_dim": 8, | |
| "quantizer_dropout": false, | |
| "sample_rate": 48000, | |
| "continuous": true | |
| } | |
| } | |