A Large-Scale Multilingual Study of Visual Constraints on Linguistic Selection of Descriptions

Uri Berger, Lea Frermann, Gabriel Stanovsky, Omri Abend

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

2 Scopus citations

Abstract

We present a large, multilingual study into how vision constrains linguistic choice, covering four languages and five linguistic properties, such as verb transitivity or use of numerals. We propose a novel method that leverages existing corpora of images with captions written by native speakers, and apply it to nine corpora, comprising 600k images and 3M captions. We study the relation between visual input and linguistic choices by training classifiers to predict the probability of expressing a property from raw images, and find evidence supporting the claim that linguistic properties are constrained by visual context across languages. We complement this investigation with a corpus study, taking the test case of numerals. Specifically, we use existing annotations (number or type of objects) to investigate the effect of different visual conditions on the use of numeral expressions in captions, and show that similar patterns emerge across languages. Our methods and findings both confirm and extend existing research in the cognitive literature. We additionally discuss possible applications for language generation. We make our codebase publicly available.

Original languageEnglish
Title of host publicationEACL 2023 - 17th Conference of the European Chapter of the Association for Computational Linguistics, Findings of EACL 2023
PublisherAssociation for Computational Linguistics (ACL)
Pages2240-2254
Number of pages15
ISBN (Electronic)9781959429470
StatePublished - 2023
Externally publishedYes
Event17th Conference of the European Chapter of the Association for Computational Linguistics, EACL 2023 - Findings of EACL 2023 - Dubrovnik, Croatia
Duration: 2 May 20236 May 2023

Publication series

NameEACL 2023 - 17th Conference of the European Chapter of the Association for Computational Linguistics, Findings of EACL 2023

Conference

Conference17th Conference of the European Chapter of the Association for Computational Linguistics, EACL 2023 - Findings of EACL 2023
Country/TerritoryCroatia
CityDubrovnik
Period2/05/236/05/23

Bibliographical note

Publisher Copyright:
© 2023 Association for Computational Linguistics.

Funding

We would like to thank the anonymous reviewers for their helpful comments and feedback. We would also like to thank Rotem Dror, Sharon Goldwater and Grzegorz Chrupała for consulting, and the native speakers that consulted and validated our annotation tool: Assaf Porat, Kozue Watan-abe, Arie Cattan, and Yilin Geng. This work was supported in part by the Israel Science Foundation (grant no. 2424/21), the Israeli Ministry of Science and Technology (grant no. 2336), and by the HUJI-UoM joint PhD program.

FundersFunder number
Israel Science Foundation2424/21
Ministry of science and technology, Israel2336

    Fingerprint

    Dive into the research topics of 'A Large-Scale Multilingual Study of Visual Constraints on Linguistic Selection of Descriptions'. Together they form a unique fingerprint.

    Cite this