Multilingual Word Embeddings using Multigraphs

Soricut, Radu; Ding, Nan

Computer Science > Computation and Language

arXiv:1612.04732 (cs)

[Submitted on 14 Dec 2016]

Title:Multilingual Word Embeddings using Multigraphs

Authors:Radu Soricut, Nan Ding

View PDF

Abstract:We present a family of neural-network--inspired models for computing continuous word representations, specifically designed to exploit both monolingual and multilingual text. This framework allows us to perform unsupervised training of embeddings that exhibit higher accuracy on syntactic and semantic compositionality, as well as multilingual semantic similarity, compared to previous models trained in an unsupervised fashion. We also show that such multilingual embeddings, optimized for semantic similarity, can improve the performance of statistical machine translation with respect to how it handles words not present in the parallel data.

Comments:	12 pages
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:1612.04732 [cs.CL]
	(or arXiv:1612.04732v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.1612.04732

Submission history

From: Radu Soricut [view email]
[v1] Wed, 14 Dec 2016 17:13:01 UTC (79 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CL

< prev | next >

new | recent | 2016-12

Change to browse by:

References & Citations

DBLP - CS Bibliography

listing | bibtex

Radu Soricut
Nan Ding

export BibTeX citation

Computer Science > Computation and Language

Title:Multilingual Word Embeddings using Multigraphs

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Multilingual Word Embeddings using Multigraphs

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators