Instructions to use rav009/VoxCPM2-lora-lanxiaoyang with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- VoxCPM
How to use rav009/VoxCPM2-lora-lanxiaoyang with VoxCPM:
import soundfile as sf from voxcpm import VoxCPM model = VoxCPM.from_pretrained("rav009/VoxCPM2-lora-lanxiaoyang") wav = model.generate( text="VoxCPM is an innovative end-to-end TTS model from ModelBest, designed to generate highly expressive speech.", prompt_wav_path=None, # optional: path to a prompt speech for voice cloning prompt_text=None, # optional: reference text cfg_value=2.0, # LM guidance on LocDiT, higher for better adherence to the prompt, but maybe worse inference_timesteps=10, # LocDiT inference timesteps, higher for better result, lower for fast speed normalize=True, # enable external TN tool denoise=True, # enable external Denoise tool retry_badcase=True, # enable retrying mode for some bad cases (unstoppable) retry_badcase_max_times=3, # maximum retrying times retry_badcase_ratio_threshold=6.0, # maximum length restriction for bad case detection (simple but effective), it could be adjusted for slow pace speech ) sf.write("output.wav", wav, 16000) print("saved: output.wav") - Notebooks
- Google Colab
- Kaggle
VoxCPM2-LoRA:懒小羊 音色适配器
English | 中文
📖 模型简介 (Model Description)
English:
This is a LoRA adapter fine-tuned on the VoxCPM2 base model. It is designed to generate speech with the "Lan Xiao Yang" voice style, a popular internet phenomenon on Douyin (TikTok China)[reference:0]. This voice is inspired by the iconic character "Lazy Sheep" from the beloved Chinese animated series "Pleasant Goat and Big Big Wolf" (Xǐ Yáng Yáng yǔ Huī Tài Láng)[reference:1][reference:2].
This adapter captures the cute, soft, and slightly slow-paced vocal characteristics of this iconic voice[reference:3][reference:4], making it suitable for vlogs, sharing daily life anecdotes, and any creative project that requires a relaxed, cute, and healing presence[reference:5][reference:6].
中文:
这是一个基于 VoxCPM2 基座模型微调的 LoRA 适配器。它旨在生成具有 "懒小羊" 音色风格的语音,该音色是抖音(TikTok 中国)上广为流行的网络现象[reference:7]。其灵感源自中国经典动画 《喜羊羊与灰太狼》 中的标志性角色"懒羊羊"[reference:8][reference:9]。
本适配器捕捉了这一标志性声音中可爱、软萌且语速稍慢的嗓音特征[reference:10][reference:11],适用于Vlog、分享日常摸鱼生活等场景,以及任何需要轻松、治愈氛围的创意项目[reference:12][reference:13]。
📊 训练信息 (Training Info)
- Training Steps: 1000 steps
⚠️ 重要声明 (Important Disclaimer)
English:
- For Educational & Research Purposes Only: This model is released for non-commercial research and learning only.
- No Illegal Use: You are strictly prohibited from using this model for any illegal, harmful, fraudulent, or malicious activities, including but not limited to generating deepfake content, impersonation, or spreading misinformation.
- User Responsibility: Users are solely responsible for complying with all applicable laws and regulations when using this model.
中文:
- 仅供学习与研究用途: 本模型仅供非商业性的研究和学习使用。
- 严禁非法用途: 严格禁止将本模型用于任何非法、有害、欺诈或恶意活动,包括但不限于生成深度伪造内容、冒充他人身份或传播虚假信息。
- 用户责任: 用户在使用本模型时,需自行承担遵守相关法律法规的责任。
Model tree for rav009/VoxCPM2-lora-lanxiaoyang
Base model
openbmb/VoxCPM2