NexesQuants/Senku-70b-iMat.GGUF
69B β’ Updated β’ 616 β’ 16
Flexible Photo Recrafting While Preserving Your Identity
Generate customized images using text and multiple images
High-fidelity 3D Geometry Generation from single view image
Convert any voice to match another speaker instantly
Generate speech from text using a reference audio
Generate and edit images using text instructions
A text-to-speech model powered by SparkAudio and Mobvoi.
Conversational speech generation
VGGT (CVPR 2025)
Detect and estimate poses in images
Generate depth estimation from images
Generate depth map from any photo
Large Animatable Human Model
Generate high-quality images from text prompts
Generate a video synced to audio from an image and pose data
Generate speech from text with voice design, cloning, or presets