Buckeye Speech Corpus
The Buckeye Corpus of conversational speech contains high-quality recordings from 40 speakers in Columbus OH conversing freely with an interviewer. The speech has been orthographically transcribed and phonetically labeled. The audio and text files, together with time-aligned phonetic labels, are stored in a format for use with speech analysis software (Xwaves and Wavesurfer). Software for searching the transcription files is currently being written.
Getting access
Registration required Restricted (registration)
- How to apply
- Not yet verified — we don’t publish a route we haven’t checked.
Summarised from a registry — the exact application route is not yet verified.
About this database
| Website | https://buckeyecorpus.osu.edu/ |
|---|---|
| Subjects | Humanities and Social Sciences · Humanities · Linguistics |
| Last updated | 2026-09-13 |