Spaces:
Running
Running
| title: README | |
| emoji: π | |
| colorFrom: gray | |
| colorTo: yellow | |
| sdk: static | |
| pinned: false | |
| **Artificial Humanity** builds deeply expressive, local-first speech synthesis β the framework | |
| layer between text and genuinely dramatic spoken performance, running entirely on-device. | |
| ## π Prosodia β the on-device dramatic audiobook engine | |
| Most text-to-speech reads. Prosodia *performs*. Every book is staged as a production: | |
| a **Director** (an on-device LLM) reads ahead and annotates each passage with emotional | |
| direction β valence, arousal, tension; an **Actor** (neural TTS) performs those notes through | |
| a Rust synthesis core; a **Stage** coordinator keeps the audio gap-free. No cloud, no | |
| telemetry β no page of your book leaves your hands. | |
| ## ποΈ Sonora β training the voices | |
| Our voice-actor training project. The first trained voice is live in the | |
| [Sonora model registry](https://huggingface.co/artificial-humanity/Sonora) β | |
| fidelity-verified ONNX and LiteRT/TFLite artifacts, including a mobile-ready split-graph | |
| export. You can hear it in the [Prosodia Space](https://huggingface.co/spaces/artificial-humanity/Prosodia). | |
| Emotional directability (VAT conditioning) and a multi-voice casting grid are in active | |
| development. | |
| --- | |
| π [GitHub](https://github.com/Artificial-Humanity) Β· π§ [Hear the voice](https://huggingface.co/spaces/artificial-humanity/Prosodia) Β· ποΈ [Sonora models](https://huggingface.co/artificial-humanity/Sonora) | |