Abstract
Detecting and analyzing stylistic variation in language is relevant to diverse Natural Language Processing applications. In this work, we investigate whether salient dimensions of style variations are embedded in standard distributional vector spaces of word meaning. We hypothesizes that distances between embeddings of lexical paraphrases can help isolate style from meaning variations and help identify latent style dimensions. We conduct a qualitative analysis of latent style dimensions, and show the effectiveness of identified style subspaces on a lexical formality prediction task.- Anthology ID:
- W17-4903
- Volume:
- Proceedings of the Workshop on Stylistic Variation
- Month:
- September
- Year:
- 2017
- Address:
- Copenhagen, Denmark
- Editors:
- Julian Brooke, Thamar Solorio, Moshe Koppel
- Venue:
- Style-Var
- SIG:
- Publisher:
- Association for Computational Linguistics
- Note:
- Pages:
- 20–27
- Language:
- URL:
- https://aclanthology.org/W17-4903
- DOI:
- 10.18653/v1/W17-4903
- Cite (ACL):
- Xing Niu and Marine Carpuat. 2017. Discovering Stylistic Variations in Distributional Vector Space Models via Lexical Paraphrases. In Proceedings of the Workshop on Stylistic Variation, pages 20–27, Copenhagen, Denmark. Association for Computational Linguistics.
- Cite (Informal):
- Discovering Stylistic Variations in Distributional Vector Space Models via Lexical Paraphrases (Niu & Carpuat, Style-Var 2017)
- PDF:
- https://preview.aclanthology.org/nschneid-patch-2/W17-4903.pdf