Published July 10, 2024
| Version v1
Dataset
Open
SLF Evaluation Dataset
Authors/Creators
Description
This dataset was constructed from the test set split of the VoxCeleb 2 dataset (VoxCeleb). The VoxCeleb 2 test set contains 118 speakers each in several different videos. To develop this dataset, only one video per speaker was selected. A face image was also extracted from the video, as well as, a low resolution face image (8x8). Age, gender and ethnicity of the person in the face image were determined using the “DeepFace” library, a face recognition and facial attribute analysis library.
This dataset can be used to evaluate speech2face, speech conditioned face generation and speech conditioned face super-resolution systems.
Files
persons.csv
Additional details
Additional titles
- Alternative title
- "Speaking the Language of Faces" Evaluation Dataset
References
- Chung, J. S., Nagrani, A., & Zisserman, A. (2018). Voxceleb2: Deep speaker recognition. Interspeech 2018. https://doi.org/10.21437/interspeech.2018-1929
- Serengil, S. I., & Ozpinar, A. (2020). Lightface: A hybrid deep face recognition framework. 2020 Innovations in Intelligent Systems and Applications Conference (ASYU). https://doi.org/10.1109/asyu50717.2020.9259802
- Serengil, S. I., & Ozpinar, A. (2021). Hyperextended Lightface: A facial attribute analysis framework. 2021 International Conference on Engineering and Emerging Technologies (ICEET). https://doi.org/10.1109/iceet53442.2021.9659697