Image-to-Video
Wan2.2
Safetensors
English
world-model
physical-language
video-generation
motion-transfer
qwen3-vl
Instructions to use misumiuika/phi_zero with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Wan2.2
How to use misumiuika/phi_zero with Wan2.2:
# No code snippets available yet for this library. # To use this model, check the repository files and the library's documentation. # Want to help? PRs adding snippets are welcome at: # https://github.com/huggingface/huggingface.js
- Notebooks
- Google Colab
- Kaggle
File size: 830 Bytes
0bdb7fd | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 40 41 42 43 44 | {
"motion_tokenizer_config": {
"in_channels": 3,
"dim": 160,
"z_channels": 48,
"vae_arch": "wan2.2_vae38",
"dim_mult": [
1,
2,
4,
4
],
"num_res_blocks": 2,
"attn_scales": [],
"temperal_downsample": [
false,
true,
true
],
"dropout": 0.0,
"embedding_dim": 8,
"quantizer": "fsq",
"levels": [
8,
5,
5,
5,
5,
5
],
"num_embeddings": null,
"beta": 0.25,
"persistent_quantizer": false,
"act_embedding_num": 32,
"qformer_type": "QFormerAdjacentFSingleQ",
"qformer_depth": 2,
"qformer_num_heads": 4,
"connector_out_channels": 1024,
"name": "WanMotionHybridTokenizer"
},
"cross_attention_dim": 3072,
"projector_hidden_dim": 3072,
"max_motion_tokens": null
} |