Published July 10, 2024 | Version v1

SLF Evaluation Dataset

Authors/Creators

Description

This dataset was constructed from the test set split of the VoxCeleb 2 dataset (VoxCeleb). The VoxCeleb 2 test set contains 118 speakers each in several different videos. To develop this dataset, only one video per speaker was selected. A face image was also extracted from the video, as well as, a low resolution face image (8x8). Age, gender and ethnicity of the person in the face image were determined using the “DeepFace” library, a face recognition and facial attribute analysis library.

This dataset can be used to evaluate speech2face, speech conditioned face generation and speech conditioned face super-resolution systems.

Files

persons.csv

Files (33.2 MB)

Name Size Download all
md5:3d14087fb3d59343066553ed49bf817c
4.3 MB Preview Download
md5:ae6412538590c0791e6995af4480fb33
2.8 MB Download
md5:a7ece9e4d2e68590fc80ff57243e40cb
26.1 kB Download
md5:f0a92a5f12a1087af6b2dfdc517f4f2d
26.0 MB Download

Additional details

Additional titles

Alternative title
"Speaking the Language of Faces" Evaluation Dataset

References