README / README.md
lmcfarlin's picture
Upload README.md with huggingface_hub
7a860ae verified
|
Raw
History Blame Contribute Delete
1.47 kB
metadata
title: README
emoji: 🎭
colorFrom: gray
colorTo: yellow
sdk: static
pinned: false

Artificial Humanity builds deeply expressive, local-first speech synthesis β€” the framework layer between text and genuinely dramatic spoken performance, running entirely on-device.

🎭 Prosodia β€” the on-device dramatic audiobook engine

Most text-to-speech reads. Prosodia performs. Every book is staged as a production: a Director (an on-device LLM) reads ahead and annotates each passage with emotional direction β€” valence, arousal, tension; an Actor (neural TTS) performs those notes through a Rust synthesis core; a Stage coordinator keeps the audio gap-free. No cloud, no telemetry β€” no page of your book leaves your hands.

πŸŽ™οΈ Sonora β€” training the voices

Our voice-actor training project. The first trained voice is live in the Sonora model registry β€” fidelity-verified ONNX and LiteRT/TFLite artifacts, including a mobile-ready split-graph export. You can hear it in the Prosodia Space. Emotional directability (VAT conditioning) and a multi-voice casting grid are in active development.


πŸ”— GitHub Β· 🎧 Hear the voice Β· πŸŽ™οΈ Sonora models