Discovering Stylistic Variations in Distributional Vector Space Models via Lexical Paraphrases

Xing Niu, Marine Carpuat


Abstract
Detecting and analyzing stylistic variation in language is relevant to diverse Natural Language Processing applications. In this work, we investigate whether salient dimensions of style variations are embedded in standard distributional vector spaces of word meaning. We hypothesizes that distances between embeddings of lexical paraphrases can help isolate style from meaning variations and help identify latent style dimensions. We conduct a qualitative analysis of latent style dimensions, and show the effectiveness of identified style subspaces on a lexical formality prediction task.
Anthology ID:
W17-4903
Volume:
Proceedings of the Workshop on Stylistic Variation
Month:
September
Year:
2017
Address:
Copenhagen, Denmark
Editors:
Julian Brooke, Thamar Solorio, Moshe Koppel
Venue:
Style-Var
SIG:
Publisher:
Association for Computational Linguistics
Note:
Pages:
20–27
Language:
URL:
https://aclanthology.org/W17-4903/
DOI:
10.18653/v1/W17-4903
Bibkey:
Cite (ACL):
Xing Niu and Marine Carpuat. 2017. Discovering Stylistic Variations in Distributional Vector Space Models via Lexical Paraphrases. In Proceedings of the Workshop on Stylistic Variation, pages 20–27, Copenhagen, Denmark. Association for Computational Linguistics.
Cite (Informal):
Discovering Stylistic Variations in Distributional Vector Space Models via Lexical Paraphrases (Niu & Carpuat, Style-Var 2017)
Copy Citation:
PDF:
https://aclanthology.org/W17-4903.pdf