Current release standard: spontaneous speech, human-validated transcripts, word-level forced alignment. Built for evaluation, not training.
AI & ML interests
Speech data for underrepresented languages, accents and niche domains, at scale. 2.5M+ consented contributors | 180+ countries | ~350 languages collectable | ~500,000 hours off the shelf 📊 Proprietary, first-party recordings with consent and provenance records, not available in any other dataset. Human-validated transcription for every language, tailored to your needs. Anything not off the shelf, sourced through our community. 📧 https://www.silencio.network/contact
Recent Activity
Organization Card
Silencio Network
This Space holds the organization card. See the datasets and collections at huggingface.co/SilencioNetwork.
models 0
None public yet
datasets 15
SilencioNetwork/indic-languages-speech
Viewer • Updated • 322 • 66
SilencioNetwork/slavic-accents-english-speech
Viewer • Updated • 100 • 77
SilencioNetwork/english-accents-speech
Viewer • Updated • 513 • 114
SilencioNetwork/spanish-accents-speech
Viewer • Updated • 19 • 51
SilencioNetwork/amharic-speech-transcribed
Viewer • Updated • 45 • 66
SilencioNetwork/french-accents-speech
Viewer • Updated • 32 • 75
SilencioNetwork/hausa-speech-transcribed
Viewer • Updated • 49 • 70
SilencioNetwork/yoruba-speech-transcribed
Viewer • Updated • 49 • 79
SilencioNetwork/swahili-speech
Viewer • Updated • 103 • 206
SilencioNetwork/tagalog-filipino-speech
Viewer • Updated • 90 • 317 • 1