For teams and researchers

License voice and dialect data collected by real speakers.

Dialect Library's dataset is growing every day. Tell us what you're building and we'll follow up about access, coverage, and licensing terms.

Speech recognition

Train or fine-tune ASR models on real accents and dialects instead of a single standardized reference.

Voice assistants

Improve wake-word and command recognition for users your product currently underserves.

Linguistic research

Access word-level translations paired with native pronunciation for underrepresented dialects.

Manual review

There is no self-serve subscription yet — every request is reviewed manually while the dataset and licensing terms are still being defined.