Abstract | ||
---|---|---|
This paper describes data collection efforts conducted as part of the RedDots project which is dedicated to the study of speaker recognition under conditions where test utterances are of short duration and of variable phonetic content. At the current stage, we focus on English speakers, both native and non-native, recruited worldwide. This is made possible through the use of a recording front-end consisting of an application running on mobile devices communicating with a centralized web server at the back-end. Speech recordings are collected by having speakers read text prompts displayed on the screen of the mobile devices. We aim to collect a large number of sessions from each speaker over a long time span, typically one session per week over a one year period. The corpus is expected to include rich inter-speaker and intra-speaker variations, both intrinsic and extrinsic (that is, due to recording channel and acoustic environment). |
Year | Venue | Keywords |
---|---|---|
2015 | 16TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION (INTERSPEECH 2015), VOLS 1-5 | speaker recognition, crowd sourcing, corpus collection |
Field | DocType | Citations |
Data collection,Computer science,Communication channel,Speech recognition,Speaker recognition,Mobile device,Speaker diarisation,Web server | Conference | 19 |
PageRank | References | Authors |
1.12 | 13 | 15 |
Name | Order | Citations | PageRank |
---|---|---|---|
Kong-Aik Lee | 1 | 709 | 60.64 |
Anthony Larcher | 2 | 362 | 24.91 |
Guangsen Wang | 3 | 28 | 4.98 |
Patrick Kenny | 4 | 2700 | 214.80 |
Niko Brümmer | 5 | 595 | 44.01 |
David A. van Leeuwen | 6 | 631 | 59.01 |
Hagai Aronowitz | 7 | 240 | 22.95 |
Marcel Kockmann | 8 | 32 | 2.87 |
Carlos Vaquero | 9 | 19 | 1.46 |
Bin Ma | 10 | 60 | 5.69 |
Haizhou Li | 11 | 3678 | 334.61 |
Themos Stafylakis | 12 | 431 | 30.12 |
jahangir alam | 13 | 320 | 38.69 |
Albert Swart | 14 | 22 | 3.21 |
Javier Pérez | 15 | 24 | 1.69 |