text_amon_API / config /config.yaml
Abdullahcoder54's picture
DreamX-Creator 1.0 on ZeroGPU: vendored videox_fun + dreamx_inference from AMAP-ML upstream; generate(image, prompt)->(mp4, last-frame PNG, seed), neutral keyframe when image empty, DREAMX_CKPT_DIR for persistent checkpoints, diffusers 0.37.1 stack
982899c
Raw History Blame Contribute Delete
1.32 kB
video_transformer_additional_kwargs:
transformer_low_noise_model_subpath: ./
transformer_combination_type: single
dict_mapping:
in_dim: in_channels
dim: hidden_size
audio_transformer_additional_kwargs:
transformer_low_noise_model_subpath: .
patch_size: [1]
in_dim: 128
out_dim: 128
vae_type: dac
video_vae_kwargs:
vae_type: AutoencoderKLWan3_8
vae_subpath: Wan2.2_VAE.pth
temporal_compression_ratio: 4
spatial_compression_ratio: 16
text_encoder_kwargs:
text_encoder_subpath: models_t5_umt5-xxl-enc-bf16.pth
tokenizer_subpath: google/umt5-xxl
text_length: 512
vocab: 256384
dim: 4096
dim_attn: 4096
dim_ffn: 10240
num_heads: 64
num_layers: 24
num_buckets: 32
shared_pos: false
dropout: 0.0
scheduler_kwargs:
scheduler_subpath: null
num_train_timesteps: 1000
shift: 5.0
use_dynamic_shifting: false
base_shift: 0.5
max_shift: 1.15
base_image_seq_len: 256
max_image_seq_len: 4096
creator_gating_kwargs:
use_temporal_rope: true
audio_fps: 50.0
vae_temporal_stride: 4
a2v_cross_attn_layers: [15, 16, 17, 18, 19, 20, 21, 22, 23, 24, 25, 26, 27, 28, 29]
v2a_cross_attn_layers: [15, 16, 17, 18, 19, 20, 21, 22, 23, 24, 25, 26, 27, 28, 29]
use_gating: true
zero_init_cross_attn: false
zero_init_gating: false
gate_init_value: 0.5