Dataset Open Access

The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS)

Livingstone, Steven R.; Russo, Frank A.


Dublin Core Export

<?xml version='1.0' encoding='utf-8'?>
<oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
  <dc:creator>Livingstone, Steven R.</dc:creator>
  <dc:creator>Russo, Frank A.</dc:creator>
  <dc:date>2018-04-05</dc:date>
  <dc:description>Contact Information

If you experience any issues downloading the RAVDESS, or if would like further information about the database, please contact us at ravdess@gmail.com. 

Construction and Validation

Construction and validation of the RAVDESS is described in our paper: Livingstone SR, Russo FA (2018) The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS): A dynamic, multimodal set of facial and vocal expressions in North American English. PLoS ONE 13(5): e0196391. https://doi.org/10.1371/journal.pone.0196391. 

Our Open Access paper is made freely available and can be downloaded without restriction from PLoS ONE.

The RAVDESS contains 7356 files. Each file was rated 10 times on emotional validity, intensity, and genuineness. Ratings were provided by 247 individuals who were characteristic of untrained adult research participants from North America. A further set of 72 participants provided test-retest data. High levels of emotional validity, interrater reliability, and test-retest intrarater reliability were reported. Validation data is open-access, and can be downloaded along with our paper from PLOS ONE.

Description

This dataset contains the complete set of 7356 RAVDESS files (total size: 24.8 GB). Each of the 24 actors consists of three modality formats: Audio-only (16bit, 48kHz .wav), Audio-Video (720p H.264, AAC 48kHz, .mp4), and Video-only (no sound).  Note, there are no song files for Actor_18.

Audio-only files

Audio-only files of all actors (01-24) are available as two separate zip files (~200 MB each):


	Speech file (Audio_Speech_Actors_01-24.zip, 215 MB) contains 1440 files: 60 trials per actor x 24 actors = 1440. 
	Song file (Audio_Song_Actors_01-24.zip, 198 MB) contains 1012 files: 44 trials per actor x 23 actors = 1012.


Audio-Visual and Video-only files

Video files are provided as separate zip downloads for each actor (01-24, ~500 MB each), and are split into separate speech and song downloads:


	Speech files (Video_Speech_Actor_01.zip to Video_Speech_Actor_24.zip) collectively contains 2880 files: 60 trials per actor x 2 modalities (AV, VO) x 24 actors = 2880.
	Song files (Video_Song_Actor_01.zip to Video_Song_Actor_24.zip) collectively contains 2024 files: 44 trials per actor x 2 modalities (AV, VO) x 23 actors = 2024.


File Summary

In total, the RAVDESS collection includes 7356 files (2880+2024+1440+1012 files).

License information

The RAVDESS is released under a Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International License, CC BY-NA-SC 4.0 

How to cite the RAVDESS

Academic citation 
If you use the RAVDESS in an academic publication, please use the following citation: 

Livingstone SR, Russo FA (2018) The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS): A dynamic, multimodal set of facial and vocal expressions in North American English. PLoS ONE 13(5): e0196391. https://doi.org/10.1371/journal.pone.0196391.

All other attributions 
If you use the RAVDESS in a form other than an academic publication, such as in a blog post, school project, or non-commercial product, please use the following attribution: "The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS)" by Livingstone &amp; Russo is licensed under CC BY-NA-SC 4.0.

File naming convention

Each of the 7356 RAVDESS files has a unique filename. The filename consists of a 7-part numerical identifier (e.g., 02-01-06-01-02-01-12.mp4). These identifiers define the stimulus characteristics: 

Filename identifiers 


	Modality (01 = full-AV, 02 = video-only, 03 = audio-only).
	Vocal channel (01 = speech, 02 = song).
	Emotion (01 = neutral, 02 = calm, 03 = happy, 04 = sad, 05 = angry, 06 = fearful, 07 = disgust, 08 = surprised).
	Emotional intensity (01 = normal, 02 = strong). NOTE: There is no strong intensity for the 'neutral' emotion.
	Statement (01 = "Kids are talking by the door", 02 = "Dogs are sitting by the door").
	Repetition (01 = 1st repetition, 02 = 2nd repetition).
	Actor (01 to 24. Odd numbered actors are male, even numbered actors are female).



Filename example: 02-01-06-01-02-01-12.mp4 


	Video-only (02)
	Speech (01)
	Fearful (06)
	Normal intensity (01)
	Statement "dogs" (02)
	1st Repetition (01)
	12th Actor (12)
	Female, as the actor ID number is even.
</dc:description>
  <dc:identifier>https://zenodo.org/record/1188976</dc:identifier>
  <dc:identifier>10.5281/zenodo.1188976</dc:identifier>
  <dc:identifier>oai:zenodo.org:1188976</dc:identifier>
  <dc:relation>doi:10.1371/journal.pone.0196391</dc:relation>
  <dc:relation>doi:10.5281/zenodo.1188975</dc:relation>
  <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
  <dc:rights>https://creativecommons.org/licenses/by-nc-sa/4.0/</dc:rights>
  <dc:source>PLoS ONE 13(5) e0196391</dc:source>
  <dc:subject>emotion</dc:subject>
  <dc:subject>emotion expression</dc:subject>
  <dc:subject>emotion perception</dc:subject>
  <dc:subject>emotion database</dc:subject>
  <dc:subject>facial expressions</dc:subject>
  <dc:subject>vocal expressions</dc:subject>
  <dc:subject>stimulus validation</dc:subject>
  <dc:subject>face</dc:subject>
  <dc:subject>voice</dc:subject>
  <dc:subject>multimodal communication</dc:subject>
  <dc:subject>RAVDESS</dc:subject>
  <dc:subject>emotion classification</dc:subject>
  <dc:title>The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS)</dc:title>
  <dc:type>info:eu-repo/semantics/other</dc:type>
  <dc:type>dataset</dc:type>
</oai_dc:dc>

Share

Cite as