Current release standard: spontaneous speech, human-validated transcripts, word-level forced alignment. Built for evaluation, not training.
AI & ML interests
Speech data for the next generation of speech models. 2M+ consented-contributors | 180+ countries | 250+ languages - High-quality, ethically-sourced voice data for ASR, TTS and conversational AI. š 250,000+ hours OTS available; Human-Validated Transcriptions available š§ info@silencio.network
Recent Activity
View all activity
Current release standard: spontaneous speech, human-validated transcripts, word-level forced alignment. Built for evaluation, not training.
Philippine-language speech under one protocol. 7,500 hours in active collection across Hiligaynon, Tagalog and Cebuano.
East and West African speech samples. Off-the-shelf inventory runs to thousands of hours per language ā 12,030 for Swahili alone.
Accented English, global French and clinical-domain speech. Transcript coverage stated per card; human transcription available on request.