Description of corpora
Romance Language Survey
Description
- The word list for this corpus was originally designed by Laura Colantoni and Jeffrey Steele's Romance Phonetic Database (containing audio recordings).
- The current data are from 11 speakers representing two varieties of French and three varieties of Spanish.
- The simultaneous EPG and audio recordings were collected in the Linguistics Phonetics Lab in 2008-09 (Spanish) and 2014-15 (French).
Speaker codes
- French: FRCf01, FRCf02, FRQf01, FRQf02 (where C = Continental/France, Q = Quebec; f = female)
- Spanish: SPAf01, SPAf02, SPAf03, SPAf04, SPAm05, SPCf01, SPPf01 (where A = Argentine, C = Cuban, P = Peninsular/Spain, f = female, m = male)
Materials
English (L1 and L2)
Description
- The 'EPG Set 1' word list was partially based on Peter Ladefoged's 'American English' chapter of the Handbook of IPA illustrating all consonant and vowel phonemes, with additional words included to illustrate allophonic variation in consonants. The set was used in this corpus for L1 speakers only.
- The 'EPG Set 2' word list was originally designed by Laura Colantoni and Jeffrey Steele to study the acquisition of English consonants, vowels, and consonants clusters by L2 learners. In this corpus it was used for both L1 and L2 speakers.
- The current data are from 3 L1 Canadian English speakers and 11 L2 speakers whose L1 is French, Japanese, Korean, or Spanish.
- The simultaneous EPG and audio recordings were collected in the Linguistics Phonetics Lab in 2009-11 (EPG Set 1) and 2015-18 (EPG Set 2).
Speaker codes
- L1 English: ENf01, ENm01, ENm02
- L2 English, L1 French: FRCf01, FRCf02, FRQf01, FRQf02 (C = Continental/France, Q = Quebec)
- L2 English, L1 Japanese: JPf03, JPf04, JPf05
- L2 English, L1 Korean: KRm02
- L2 English, L1 Spanish: SPAf04, SPCf01, SPPf01 (A = Argentine, C = Cuban, P = Peninsular/Spain)
Materials
Japanese
Description
- The word list was designed by Alexei Kochetov (with the assistance of Maho Kobayashi) to illustrate the production of word-initial consonants in several vowel contexts.
- The data are from 5 speakers from various locations on the island of Honshu.
- The simultaneous EPG and audio recordings were collected in the Linguistics Phonetics Lab in 2008-09.
Speaker codes
- JPf01, JPf02, JPf03, JPf04, JPf05
Materials
- Words (alphabetical), produced in the carrier phrase Sore ga __ to itta.
Serbian
Description
- The word list was designed by Milica Radišić to illustrate various consonants and vowels of the language; it was partially based on Landau et al. (1995)'s JIPA Illustration of IPA 'Croatian' (adapted for Serbian), with additional sentences added to illustrate the contrasts in laterals, nasals, and affricates across word positions and vowel contexts.
- The data are from 4 speakers, two of whom are from Niš and two from Novi Sad.
- The simultaneous EPG and audio recordings were collected in the Linguistics Phonetics Lab in 2008-10.
Speaker codes
- SRf01, SRm01, SRf02, SRf03
Materials
The Cross-Language Articulatory Database (CLAD) @ CHASS / University of Toronto Copyright © 2026 University of Toronto