Abstract
We present a Cantonese emotional speech dataset that is suitable for use in research investigating the auditory and visual expression of emotion in tonal languages. This unique dataset consists of auditory and visual recordings of ten native speakers of Cantonese uttering 50 sentences each in the six basic emotions plus neutral (angry, happy, sad, surprise, fear, and disgust). The visual recordings have a full HD resolution of 1920 × 1080 pixels and were recorded at 50 fps. The important features of the dataset are outlined along with the factors considered when compiling the dataset. A validation study of the recorded emotion expressions was conducted in which 15 native Cantonese perceivers completed a forced-choice emotion identification task. The variability of the speakers and the sentences was examined by testing the degree of concordance between the intended and the perceived emotion. We compared these results with those of other emotion perception and evaluation studies that have tested spoken emotions in languages other than Cantonese. The dataset is freely available for research purposes. A correction to the original article is available at: https://doi.org/10.3758/s13428-023-02270-7
| Original language | English |
|---|---|
| Pages (from-to) | 5264-5278 |
| Number of pages | 15 |
| Journal | Behavior Research Methods |
| Volume | 56 |
| Issue number | 5 |
| DOIs | |
| Publication status | Published - Aug 2024 |
Bibliographical note
Publisher Copyright:© The Author(s) 2023.
Open Access - Access Right Statement
Creative Commons licence and your intended use is not permitted by statutory regulation or exceeds the permitted use, you will need to obtain permission directly from the copyright holder. To view a copy of this licence, visit http://creativecommons.org/licenses/by/4.0/.Keywords
- Auditory and visual expressions
- Cantonese dataset
- Dataset evaluation
- Emotional speech
Fingerprint
Dive into the research topics of 'A Cantonese audio-visual emotional speech (CAVES) dataset'. Together they form a unique fingerprint.Datasets
-
Cantonese Audio-Visual Emotional Speech (CAVES) dataset
Davis, C., Kim, J. & Chong, C. S., Western Sydney University, 2024
DOI: 10.26183/3se5-s316, https://research-data.westernsydney.edu.au/published/56177c30011711ef8f04a5d757634468
Dataset
Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver