Image-to-Video
Wan2.2
Safetensors
English
world-model
physical-language
video-generation
motion-transfer
qwen3-vl
Instructions to use misumiuika/phi_zero with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Wan2.2
How to use misumiuika/phi_zero with Wan2.2:
# No code snippets available yet for this library. # To use this model, check the repository files and the library's documentation. # Want to help? PRs adding snippets are welcome at: # https://github.com/huggingface/huggingface.js
- Notebooks
- Google Colab
- Kaggle
Download tokenizer/model_config.json from misumiuika/phi_zero: direct link, hf CLI and curl.
- Browser
- Download file 830 Bytes
-
https://huggingface.co/misumiuika/phi_zero/resolve/main/tokenizer/model_config.json
- Command line
-
hf download hf://misumiuika/phi_zero/tokenizer/model_config.json
-
curl -L -o model_config.json https://huggingface.co/misumiuika/phi_zero/resolve/main/tokenizer/model_config.json
830 Bytes
| { | |
| "motion_tokenizer_config": { | |
| "in_channels": 3, | |
| "dim": 160, | |
| "z_channels": 48, | |
| "vae_arch": "wan2.2_vae38", | |
| "dim_mult": [ | |
| 1, | |
| 2, | |
| 4, | |
| 4 | |
| ], | |
| "num_res_blocks": 2, | |
| "attn_scales": [], | |
| "temperal_downsample": [ | |
| false, | |
| true, | |
| true | |
| ], | |
| "dropout": 0.0, | |
| "embedding_dim": 8, | |
| "quantizer": "fsq", | |
| "levels": [ | |
| 8, | |
| 5, | |
| 5, | |
| 5, | |
| 5, | |
| 5 | |
| ], | |
| "num_embeddings": null, | |
| "beta": 0.25, | |
| "persistent_quantizer": false, | |
| "act_embedding_num": 32, | |
| "qformer_type": "QFormerAdjacentFSingleQ", | |
| "qformer_depth": 2, | |
| "qformer_num_heads": 4, | |
| "connector_out_channels": 1024, | |
| "name": "WanMotionHybridTokenizer" | |
| }, | |
| "cross_attention_dim": 3072, | |
| "projector_hidden_dim": 3072, | |
| "max_motion_tokens": null | |
| } |