Casey M
AI & ML interests
Recent Activity
Organizations
LongCat-Video-Avatar 1.5
Audio-driven talking-head video generation (Meituan LongCat)
Stable Audio 3
Text-to-audio with SA3 Medium / Small Music / Small SFX.
Pixal3D-Server
High-fidelity pixel-aligned image-to-3D generation.
AudioπΉSeparator
Vocal and background audio separator
Khala β High-Fidelity Song Generation
Generate high-fidelity songs from text prompts
Scenema Audio
Zero-shot expressive voice cloning and speech generation
DramaBox
Expressive TTS with voice cloning β DramaBox demo
Basic Upscaler
Upscale images using various models
Wan2.2 14B Fast Preview
generate a video from an image with a text prompt
Wan2.2 14B Preview
generate a video from an image with a text prompt
MMAudio β generating synchronized audio from video/text
Generate synchronized audio for videos or from text prompts
OmniVoice
High-quality voice cloning TTS for 600+ languages
VoxCPM Demo
VoxCPM2 Nano-vLLM Demo
See-through: Layer Decomposition
Generate layered PSD of anime characters from a single image
Cohere Multilingual ASR
Transcribe audio clips to text in multiple languages
PrismAudio
Generate audio for your video using a text prompt