Current release standard: spontaneous speech, human-validated transcripts, word-level forced alignment. Built for evaluation, not training.
AI & ML interests
Speech data for underrepresented languages, accents and niche domains, at scale. 2.5M+ consented contributors | 180+ countries | ~350 languages collectable | ~500,000 hours off the shelf 📊 Proprietary, first-party recordings with consent and provenance records, not available in any other dataset. Human-validated transcription for every language, tailored to your needs. Anything not off the shelf, sourced through our community. 📧 https://www.silencio.network/contact
Recent Activity
Organization Card
Silencio Network
This Space holds the organization card. See the datasets and collections at huggingface.co/SilencioNetwork.
models 0
None public yet
datasets 15
SilencioNetwork/indic-languages-speech
Viewer • Updated • 322 • 106
SilencioNetwork/slavic-accents-english-speech
Viewer • Updated • 100 • 89
SilencioNetwork/english-accents-speech
Viewer • Updated • 513 • 245
SilencioNetwork/spanish-accents-speech
Viewer • Updated • 19 • 106 • 1
SilencioNetwork/amharic-speech-transcribed
Viewer • Updated • 45 • 80
SilencioNetwork/french-accents-speech
Viewer • Updated • 32 • 118
SilencioNetwork/hausa-speech-transcribed
Viewer • Updated • 49 • 104
SilencioNetwork/yoruba-speech-transcribed
Viewer • Updated • 49 • 96
SilencioNetwork/swahili-speech
Viewer • Updated • 103 • 176
SilencioNetwork/tagalog-filipino-speech
Viewer • Updated • 90 • 249 • 1