Published April 15, 2020 | Version v1

LRRo: A Lip Reading Data Set for the Under-resourced Romanian Language

  • 1. University Politehnica of Bucharest

Description

Two distinct collections are presented in this repository:

(i) wild LRRo data is designed for an Internet in-the-wild, ad-hoc scenario, coming with more than 35 different speakers, 1.1k words, a vocabulary of 21 words, and more than 20 hours;

(ii) lab LRRo data, addresses a lab controlled scenario for more accurate data, coming with 19 different speakers, 6.4k words, a vocabulary of 48 words, and more than 5 hours.

Notes

If you make use of this collection, please acknowledge the work of the authors by citing the following publication: Andrei Cosmin Jitaru, Şeila Abdulamit, and Bogdan Ionescu. 2020. LRRo: A Lip Reading Data Set for the Under-resourced Romanian Language. In 11th ACM Multimedia Systems Conference (MMSys'20), June 8–11, 2020, Istanbul, Turkey. ACM, New York, NY, USA, 6 pages. https://doi.org/10.1145/3339825.3394932

Files

Files (291.5 MB)

Name Size Download all
md5:17f10ef83c01e596dacf39b7cab658be
291.5 MB Download

Additional details

References

  • Andrei Cosmin Jitaru, Şeila Abdulamit, and Bogdan Ionescu, "LRRo:A Lip Reading Data Set for the Under-resourced Romanian Language" in MMSys '20: ACM Open Dataset and Software Track, June 8-11, 2020, Istanbul, Turkey