File size: 1,471 Bytes
61c0d92
 
7a860ae
 
 
61c0d92
 
 
 
7a860ae
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
---
title: README
emoji: 🎭
colorFrom: gray
colorTo: yellow
sdk: static
pinned: false
---

**Artificial Humanity** builds deeply expressive, local-first speech synthesis β€” the framework
layer between text and genuinely dramatic spoken performance, running entirely on-device.

## 🎭 Prosodia β€” the on-device dramatic audiobook engine

Most text-to-speech reads. Prosodia *performs*. Every book is staged as a production:
a **Director** (an on-device LLM) reads ahead and annotates each passage with emotional
direction β€” valence, arousal, tension; an **Actor** (neural TTS) performs those notes through
a Rust synthesis core; a **Stage** coordinator keeps the audio gap-free. No cloud, no
telemetry β€” no page of your book leaves your hands.

## πŸŽ™οΈ Sonora β€” training the voices

Our voice-actor training project. The first trained voice is live in the
[Sonora model registry](https://huggingface.co/artificial-humanity/Sonora) β€”
fidelity-verified ONNX and LiteRT/TFLite artifacts, including a mobile-ready split-graph
export. You can hear it in the [Prosodia Space](https://huggingface.co/spaces/artificial-humanity/Prosodia).
Emotional directability (VAT conditioning) and a multi-voice casting grid are in active
development.

---

πŸ”— [GitHub](https://github.com/Artificial-Humanity) Β· 🎧 [Hear the voice](https://huggingface.co/spaces/artificial-humanity/Prosodia) Β· πŸŽ™οΈ [Sonora models](https://huggingface.co/artificial-humanity/Sonora)