Instructions to use rav009/VoxCPM2-lora-xiaohai with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- VoxCPM
How to use rav009/VoxCPM2-lora-xiaohai with VoxCPM:
import soundfile as sf from voxcpm import VoxCPM model = VoxCPM.from_pretrained("rav009/VoxCPM2-lora-xiaohai") wav = model.generate( text="VoxCPM is an innovative end-to-end TTS model from ModelBest, designed to generate highly expressive speech.", prompt_wav_path=None, # optional: path to a prompt speech for voice cloning prompt_text=None, # optional: reference text cfg_value=2.0, # LM guidance on LocDiT, higher for better adherence to the prompt, but maybe worse inference_timesteps=10, # LocDiT inference timesteps, higher for better result, lower for fast speed normalize=True, # enable external TN tool denoise=True, # enable external Denoise tool retry_badcase=True, # enable retrying mode for some bad cases (unstoppable) retry_badcase_max_times=3, # maximum retrying times retry_badcase_ratio_threshold=6.0, # maximum length restriction for bad case detection (simple but effective), it could be adjusted for slow pace speech ) sf.write("output.wav", wav, 16000) print("saved: output.wav") - Notebooks
- Google Colab
- Kaggle
VoxCPM2-LoRA:小孩 音色适配器
English | 中文
📖 模型简介 (Model Description)
English:
This is a LoRA adapter fine-tuned on the VoxCPM2 base model. It is designed to generate speech with the "Child" (Xiǎo Hái) voice style, a popular voice on Douyin (TikTok China).
This voice style is widely used in meme videos, cute pet dubbing, and content featuring children's voices, often paired with playful, innocent, and slightly exaggerated expressions. The adapter captures the high-pitched, cute, and lively vocal characteristics of a child's voice, making it suitable for creating family-friendly content, dubbing for animations, comedy sketches, and social media videos.
中文:
这是一个基于 VoxCPM2 基座模型微调的 LoRA 适配器。它旨在生成具有 "小孩" 音色风格的语音,该音色是抖音(TikTok 中国)上广为流行的视频配音。
这种音色风格常被用于搞笑视频、萌宠配音以及儿童相关内容中,通常搭配俏皮、天真且略带夸张的表达方式。本适配器捕捉了小孩声音中高亢、可爱且活泼的嗓音特征,非常适合用于创作家庭友好型内容、动画配音、喜剧片段以及社交媒体短视频。
📊 训练信息 (Training Info)
- Training Steps: 1000 steps
⚠️ 重要声明 (Important Disclaimer)
English:
- For Educational & Research Purposes Only: This model is released for non-commercial research and learning only.
- No Illegal Use: You are strictly prohibited from using this model for any illegal, harmful, fraudulent, or malicious activities, including but not limited to generating deepfake content, impersonation, or spreading misinformation.
- User Responsibility: Users are solely responsible for complying with all applicable laws and regulations when using this model.
中文:
- 仅供学习与研究用途: 本模型仅供非商业性的研究和学习使用。
- 严禁非法用途: 严格禁止将本模型用于任何非法、有害、欺诈或恶意活动,包括但不限于生成深度伪造内容、冒充他人身份或传播虚假信息。
- 用户责任: 用户在使用本模型时,需自行承担遵守相关法律法规的责任。
Model tree for rav009/VoxCPM2-lora-xiaohai
Base model
openbmb/VoxCPM2