CHILDES English kidLUCID Corpus
|
Valerie Hazan
SHaPS - emeritus
UCL
v.hazan@ucl.ac.uk
|
| Participants: | 96 |
| Type of Study: | speaker controlled variability |
| Location: | UK |
| Media type: | audio |
| DOI: | doi:10.21415/kye9-r061 |
Browsable transcripts
Download transcripts
Link to media folder
Citation information
- Baker, R., & Hazan, V (2011). DiapixUK: a task for the
elicitation of spontaneous speech dialogs. Behavior Research Methods,
43(3), 761-770.
- Hazan, V. L., & Baker, R. (2011). Acoustic-phonetic characteristics of
speech produced with communicative intent to counter adverse listening
conditions. Journal of the Acoustical Society of America, 130(4),
2139-2152.
In accordance with TalkBank rules, any use of data from this corpus
must be accompanied by at least one of the above references.
Project Description
A complete description of the project is in
this pdf
Picture stimuli as in this .zip file.
Praat textGrids are in this .zip file.
Acronym: LUCID = London UCL Clear Speech in Interaction Database.
Project title: Speaker-controlled Variability in Children's Speech in
Interaction (A research project funded by the ESRC)
- >96 talkers (all native southern British English speakers, 46 male, 50 female). Recorded in pairs.
- A total of 288 conversations distributed across three conditions as follows:
- NOB (No barrier): 96 conversations while they both heard each other normally.
- BAB (Babble): 96 conversations where one conversational partner
heard the other's speech in a background for multi-talker babble at
an approximate SNR of 0 dB. The talker hearing the babble was a
confederate. Exceptions: 4 conversations with CBB ('child multitalker
babble' as opposed to 'adult multitalker babble')
- VOC (Vocoded): 96 conversations where one conversational partner
heard the other's speech after it had been processed in real time
through a noise-excited three channel vocoder
- Each condition contains 1xTextgrid and 1x sound file that both
contain the two speakers. SpA (i.e., Speaker A) is the leader (always
on the right audio channel and bottom two Tiers in TextGrid files) whose
speech has been impaired (i.e., the one who has to clarify their
speech). SpA is coded in the filenames as follows: for example,
F01F02BAB1S2.wav, that is, F01 is the one on the left audio channel
(SpB) and F02 the one on the right audio channel (SpA).
- >(NB! The
transcribed data [textgrids] has been automatically aligned with the
audio. Right channel data has been manually checked for word level
alignment and vowel alignment but note that this is for three vowels
only (iy/ao/ae) whereas the left channel alignment has not been
checked.)
- File naming has this info: F/M female/male; 01/02 speaker order;
S = Street, F = Farm scene , B = Beach scene; BAB/NOB/VOC