Image-to-Video
LivePortrait
ONNX
MLX
face-detection
face-recognition
face-landmark
talking-head
audio-driven-animation
joyvasa
auraface
mediapipe
apple-silicon
Instructions to use talkyon/scribis_models with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- LivePortrait
How to use talkyon/scribis_models with LivePortrait:
# No code snippets available yet for this library. # To use this model, check the repository files and the library's documentation. # Want to help? PRs adding snippets are welcome at: # https://github.com/huggingface/huggingface.js
- MLX
How to use talkyon/scribis_models with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir scribis_models talkyon/scribis_models
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
File size: 8,171 Bytes
2941712 | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 40 41 42 43 44 45 46 47 48 49 50 51 52 53 54 55 56 57 58 59 60 61 62 63 64 65 66 67 68 69 70 71 72 73 74 75 76 77 78 79 80 81 82 83 84 85 86 87 88 89 90 91 92 93 94 95 96 97 98 99 100 101 102 103 104 105 106 107 108 109 110 111 112 113 114 115 116 117 118 119 120 121 122 123 124 125 126 127 128 129 130 131 132 133 134 135 136 137 138 139 140 141 142 143 144 145 146 147 148 149 150 151 152 153 154 155 156 157 158 159 160 161 162 163 164 165 166 167 168 169 170 171 172 173 174 175 176 177 178 179 180 181 182 183 184 185 186 187 188 189 190 191 192 193 194 195 196 197 198 199 200 201 202 203 204 205 206 207 208 209 210 211 212 213 214 215 216 217 218 219 220 221 222 223 224 225 226 227 228 229 230 231 232 233 234 235 236 237 238 239 240 241 242 243 244 245 246 247 248 249 250 251 252 253 254 255 256 257 258 259 260 261 262 263 264 265 266 267 268 269 270 271 272 273 274 275 276 277 278 279 280 281 282 283 284 285 286 287 288 289 290 291 292 293 294 295 296 297 298 299 300 301 302 303 304 305 306 307 308 309 310 311 312 313 314 315 316 317 318 319 320 321 | ---
license: other
license_name: multiple-upstream-model-licenses
license_link: LICENSE.md
tags:
- face-detection
- face-recognition
- face-landmark
- talking-head
- image-to-video
- audio-driven-animation
- joyvasa
- liveportrait
- auraface
- mediapipe
- onnx
- mlx
- apple-silicon
---
# Scribis Model Bundle
This repository contains third-party and converted model assets used by
Scribis for face analysis and talking-avatar features.
The repository contains multiple upstream model families and multiple runtime
formats. It therefore does not apply one blanket license to every model file.
See `LICENSE.md` before redistributing or using these assets.
## Repository layout
```text
auraface/
AuraFace identity-embedding model
mediapipe/
MediaPipe-based face detection and landmark models
avatar/
Talking-avatar models for ONNX Runtime and native Apple MLX
```
The repository is intentionally separated by model family so that provenance,
runtime format, and license attribution remain easier to track.
---
# AuraFace
## Identity embedding
Model:
```text
auraface/glintr100.onnx
```
Upstream source:
```text
Hugging Face: fal/AuraFace-v1
```
The local model originates from the AuraFace-v1 repository.
SHA-256:
```text
a7933ea5330113b01c9b60351d8f4c33003f145d8470ac5f0e52ee2effe25c60
```
AuraFace-v1 is marked Apache-2.0 by the upstream repository.
The AuraFace model card describes commercial applications among its intended
uses.
AuraFace examples may use InsightFace software as an inference wrapper.
This repository does not redistribute the InsightFace Python package or
InsightFace model packages as part of the AuraFace directory.
---
# MediaPipe
## Face detection and landmarks
Upstream source:
```text
Hugging Face: Heliosoph/mediapipe-face-onnx
```
The upstream repository identifies the model bundle as Apache-2.0 and
documents its provenance through Google MediaPipe, MediaPipePyTorch, and
Qualcomm AI Hub.
The float model bundle used by Scribis consists of the corresponding detector
and landmark ONNX models together with their external ONNX data files.
Typical layout:
```text
mediapipe/
βββ float/
βββ face_detector.onnx
βββ face_detector.data
βββ face_landmark_detector.onnx
βββ face_landmark_detector.data
```
The matching `.onnx` and `.data` files are parts of the same ONNX model and
must remain together when copied or redistributed.
Scribis uses these MediaPipe-based assets for face detection and landmark
processing instead of redistributing the InsightFace detection models used by
the original LivePortrait pipeline.
---
# Avatar
The `avatar/` directory contains the audio-driven motion and portrait-animation
models used by Scribis.
Two runtime paths are provided:
```text
avatar/onnx/ ONNX Runtime models
avatar/mlx/ Native Apple MLX models
```
## JoyVASA and Chinese HuBERT
Upstream sources:
```text
GitHub / Hugging Face: jdh-algo/JoyVASA
Hugging Face: TencentGameMate/chinese-hubert-base
```
Scribis uses Chinese HuBERT as the audio encoder and JoyVASA as the
audio-to-motion model.
### ONNX
```text
avatar/onnx/audio_encoder.onnx
avatar/onnx/motion_generator.onnx
avatar/onnx/joyvasa.json
avatar/onnx/joyvasa-template.json
```
`audio_encoder.onnx` is an ONNX conversion of the corresponding Chinese
HuBERT audio encoder.
`motion_generator.onnx` is an ONNX conversion of the JoyVASA motion generator.
The JSON files contain runtime metadata used by the Scribis ONNX pipeline.
### MLX
```text
avatar/mlx/JoyVASA/audio_encoder/hubert_chinese_mlx.npz
avatar/mlx/JoyVASA/motion_generator/motion_generator_hubert_chinese_mlx.npz
avatar/mlx/JoyVASA/motion_template/motion_template.pkl
```
The MLX audio and motion weights are converted runtime representations of
their corresponding upstream model assets.
Converted model weights retain the applicable upstream license obligations.
The upstream Chinese HuBERT and JoyVASA repositories are currently marked MIT.
---
# LivePortrait
Scribis uses LivePortrait human-animation models through ONNX and native MLX
runtime representations.
Relevant upstream projects:
```text
GitHub: KlingAIResearch/LivePortrait
Hugging Face: KlingTeam/LivePortrait
GitHub: warmshao/FasterLivePortrait
Hugging Face: warmshao/FasterLivePortrait
GitHub: ivanfioravanti/fasterliveportrait-mlx
Hugging Face: ivanfioravanti/FasterLivePortrait-MLX-weights
```
## ONNX LivePortrait assets
The ONNX models used by Scribis were obtained through FasterLivePortrait and
are derived from the LivePortrait model pipeline.
Scribis currently redistributes only:
```text
avatar/onnx/liveportrait_onnx/
βββ appearance_feature_extractor.onnx
βββ landmark.onnx
βββ motion_extractor.onnx
βββ stitching.onnx
βββ stitching_eye.onnx
βββ stitching_lip.onnx
βββ warping_spade.onnx
```
`warping_spade.onnx` represents the ONNX runtime path corresponding to the
LivePortrait warping and SPADE generation stages.
FasterLivePortrait licenses its source code under MIT but explicitly states
that machine-learning model files remain subject to their respective original
model licenses.
The official LivePortrait model repository is marked MIT.
## MLX LivePortrait assets
The native MLX weights used by Scribis are derived from the
FasterLivePortrait-MLX conversion project.
Current files:
```text
avatar/mlx/liveportrait_mlx/
βββ appearance_feature_extractor.npz
βββ landmark.npz
βββ motion_extractor.npz
βββ spade_generator.npz
βββ stitching.npz
βββ stitching_eye.npz
βββ stitching_lip.npz
βββ warping_module.npz
```
Unlike the combined ONNX `warping_spade.onnx` runtime model, the MLX path keeps
the warping module and SPADE generator as separate weight files:
```text
warping_module.npz
spade_generator.npz
```
The FasterLivePortrait-MLX source code is MIT.
Its converted model files remain derivative model weights and retain the
applicable license and attribution obligations of their upstream sources.
---
# Excluded model assets
This repository does not redistribute the InsightFace detection models from
the original LivePortrait pipeline.
In particular, the Scribis bundle does not include these FasterLivePortrait
ONNX assets:
```text
retinaface_det_static.onnx
face_2dpose_106_static.onnx
```
The official LivePortrait license states that InsightFace model weights are
for non-commercial research purposes and recommends replacing those detection
models for commercial use.
Scribis instead uses the MediaPipe-based models in `mediapipe/` for its face
detection and landmark pipeline.
The Scribis MLX bundle also does not include XPose.
---
# License and attribution
This is an aggregate model repository.
Different files originate from different upstream projects, so the repository
is intentionally marked:
```yaml
license: other
```
This designation does not mean the individual model assets have no license.
It means the repository as a whole cannot accurately be represented by one
single blanket license.
See `LICENSE.md` for the component-by-component license and attribution
summary.
Converted ONNX and MLX weights retain the applicable rights, restrictions,
copyright notices, and attribution obligations of their respective upstream
model sources.
When redistributing individual assets, consult the corresponding upstream
project and preserve the applicable license and attribution notices.
---
# Runtime notes
- The repository contains both ONNX and MLX model assets.
- `warping_spade.onnx` is the full-quality ONNX warping/SPADE runtime model
currently used by Scribis.
- `warping_spade_fp16.onnx` is not included.
- ONNX models that use external `.data` files must remain beside those files.
- `.DS_Store` and other local operating-system metadata files should not be
uploaded.
- Adding a new model family requires reviewing and documenting that model's
upstream source and license separately.
- Users are responsible for complying with applicable copyright, privacy,
biometric-data, publicity-rights, model-license, and other applicable laws.
|