A Large-Scale Multilingual Study of Visual Constraints on Linguistic Selection of Descriptions

Berger, Uri; Frermann, Lea; Stanovsky, Gabriel; Abend, Omri

Computer Science > Computation and Language

arXiv:2302.04811 (cs)

[Submitted on 9 Feb 2023]

Title:A Large-Scale Multilingual Study of Visual Constraints on Linguistic Selection of Descriptions

Authors:Uri Berger, Lea Frermann, Gabriel Stanovsky, Omri Abend

View PDF

Abstract:We present a large, multilingual study into how vision constrains linguistic choice, covering four languages and five linguistic properties, such as verb transitivity or use of numerals. We propose a novel method that leverages existing corpora of images with captions written by native speakers, and apply it to nine corpora, comprising 600k images and 3M captions. We study the relation between visual input and linguistic choices by training classifiers to predict the probability of expressing a property from raw images, and find evidence supporting the claim that linguistic properties are constrained by visual context across languages. We complement this investigation with a corpus study, taking the test case of numerals. Specifically, we use existing annotations (number or type of objects) to investigate the effect of different visual conditions on the use of numeral expressions in captions, and show that similar patterns emerge across languages. Our methods and findings both confirm and extend existing research in the cognitive literature. We additionally discuss possible applications for language generation.

Comments:	Accepted to EACL 2023 Findings
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2302.04811 [cs.CL]
	(or arXiv:2302.04811v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2302.04811

Submission history

From: Uri Berger [view email]
[v1] Thu, 9 Feb 2023 17:57:58 UTC (3,487 KB)

Computer Science > Computation and Language

Title:A Large-Scale Multilingual Study of Visual Constraints on Linguistic Selection of Descriptions

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:A Large-Scale Multilingual Study of Visual Constraints on Linguistic Selection of Descriptions

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators