Instructions to use mlboydaisuke/VoxCPM2-CoreAI with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- VoxCPM
How to use mlboydaisuke/VoxCPM2-CoreAI with VoxCPM:
import soundfile as sf from voxcpm import VoxCPM model = VoxCPM.from_pretrained("mlboydaisuke/VoxCPM2-CoreAI") wav = model.generate( text="VoxCPM is an innovative end-to-end TTS model from ModelBest, designed to generate highly expressive speech.", prompt_wav_path=None, # optional: path to a prompt speech for voice cloning prompt_text=None, # optional: reference text cfg_value=2.0, # LM guidance on LocDiT, higher for better adherence to the prompt, but maybe worse inference_timesteps=10, # LocDiT inference timesteps, higher for better result, lower for fast speed normalize=True, # enable external TN tool denoise=True, # enable external Denoise tool retry_badcase=True, # enable retrying mode for some bad cases (unstoppable) retry_badcase_max_times=3, # maximum retrying times retry_badcase_ratio_threshold=6.0, # maximum length restriction for bad case detection (simple but effective), it could be adjusted for slow pace speech ) sf.write("output.wav", wav, 16000) print("saved: output.wav") - Notebooks
- Google Colab
- Kaggle
macos voxcpm2_vocoder_fp16_t8
Browse files
.gitattributes
CHANGED
|
@@ -39,3 +39,4 @@ macos/voxcpm2_base_int8_prefill_t32/voxcpm2_base_int8_prefill_t32.aimodel/main.m
|
|
| 39 |
macos/voxcpm2_res_int8_prefill_t32/voxcpm2_res_int8_prefill_t32.aimodel/main.mlirb filter=lfs diff=lfs merge=lfs -text
|
| 40 |
macos/voxcpm2_feat_decoder_fp16/voxcpm2_feat_decoder_fp16.aimodel/main.mlirb filter=lfs diff=lfs merge=lfs -text
|
| 41 |
macos/voxcpm2_feat_encoder_fp16/voxcpm2_feat_encoder_fp16.aimodel/main.mlirb filter=lfs diff=lfs merge=lfs -text
|
|
|
|
|
|
| 39 |
macos/voxcpm2_res_int8_prefill_t32/voxcpm2_res_int8_prefill_t32.aimodel/main.mlirb filter=lfs diff=lfs merge=lfs -text
|
| 40 |
macos/voxcpm2_feat_decoder_fp16/voxcpm2_feat_decoder_fp16.aimodel/main.mlirb filter=lfs diff=lfs merge=lfs -text
|
| 41 |
macos/voxcpm2_feat_encoder_fp16/voxcpm2_feat_encoder_fp16.aimodel/main.mlirb filter=lfs diff=lfs merge=lfs -text
|
| 42 |
+
macos/voxcpm2_vocoder_fp16_t8/voxcpm2_vocoder_fp16_t8.aimodel/main.mlirb filter=lfs diff=lfs merge=lfs -text
|
macos/voxcpm2_vocoder_fp16_t8/voxcpm2_vocoder_fp16_t8.aimodel/main.hash
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
R:�$�饙����t�|W�N��
|
macos/voxcpm2_vocoder_fp16_t8/voxcpm2_vocoder_fp16_t8.aimodel/main.mlirb
ADDED
|
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:523afe031a24d7e9a5998805ccc1b40474e2ae7c57cb4e8c1805e017f3b097b3
|
| 3 |
+
size 91578309
|
macos/voxcpm2_vocoder_fp16_t8/voxcpm2_vocoder_fp16_t8.aimodel/metadata.json
ADDED
|
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
{
|
| 2 |
+
"assetVersion" : "2.0"
|
| 3 |
+
}
|