Learning translations via images with a massively multilingual image dataset

Hewitt, John; Ippolito, Daphne; Callahan, Brendan; Kriz, Reno; Wijaya, Derry; Callison-Burch, Chris

Learning translations via images with a massively multilingual image dataset

Files

P18-1239.pdf(3.4 MB)

Published version

Date

2018-07-15

DOI

10.18653/v1/P18-1239

Authors

Hewitt, John

Ippolito, Daphne

Callahan, Brendan

Kriz, Reno

Wijaya, Derry

Callison-Burch, Chris

Version

Published version

URI

https://hdl.handle.net/2144/38467

Citation

John Hewitt, Daphne Ippolito, Brendan Callahan, Reno Kriz, Derry Wijaya, Chris Callison-Burch. 2018. "Learning Translations via Images with a Massively Multilingual Image Dataset." Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics. Annual Meeting of the Association for Computational Linguistics. Melbourne, Australia, 2018-07-15 - 2018-07-20. https://doi.org/10.18653/v1/P18-1239

Abstract

We conduct the most comprehensive study to date into translating words via images. To facilitate research on the task, we introduce a large-scale multilingual corpus of images, each labeled with the word it represents. Past datasets have been limited to only a few high-resource languages and unrealistically easy translation settings. In contrast, we have collected by far the largest available dataset for this task, with images for approximately 10,000 words in each of 100 languages. We run experiments on a dozen high resource languages and 20 low resources languages, demonstrating the effect of word concreteness and part-of-speech on translation quality. %We find that while image features work best for concrete nouns, they are sometimes effective on other parts of speech. To improve image-based translation, we introduce a novel method of predicting word concreteness from images, which improves on a previous state-of-the-art unsupervised technique. This allows us to predict when image-based translation may be effective, enabling consistent improvements to a state-of-the-art text-based word translation system. Our code and the Massively Multilingual Image Dataset (MMID) are available at http://multilingual-images.org/.

License

cb

Collections

BU Open Access Articles
CAS: Computer Science: Scholarly Papers

Full item page