Performance comparison of multimodal and textual intermediate tasks in zero-shot cross-lingual transfer on XTREME-R
Description
In zero-shot cross-lingual transfer, a supervised NLP task trained on a corpus in one language is directly applicable to another language without any additional training. A source of cross-lingual transfer can be as straightforward as lexical overlap between languages (e.g., use of the same scripts, shared subwords) that naturally forces text embeddings to occupy a similar representation space. Recently introduced cross-lingual language model (XLM) pretraining brings out neural parameter sharing in Transformer-style networks as the most important factor for the transfer. In this paper, we aim
Research goal: How does the performance of multimodal intermediate-task training (e.g., image-text alignment) compare to purely textual intermediate tasks in zero-shot cross-lingual transfer accuracy on the XTREME-R benchmark?
Autonomous synthesis report generated by Assignee Research. Tribunal consensus score: 7.7/10.
Notes
Files
paper.pdf
Files
(87.1 kB)
| Name | Size | Download all |
|---|---|---|
|
md5:9423370fa66b3d0395b57056bd589e0a
|
87.1 kB | Preview Download |