An empirical study on the contribution of formal and semantic features to the grammatical gender of nouns

Publikation: Bidrag til tidsskriftTidsskriftartikelForskningfagfællebedømt

This study conducts an experimental evaluation of two hypotheses about the contributions of formal and semantic features to the grammatical gender assignment of nouns. One of the hypotheses (Corbett and Fraser 2000) claims that semantic features dominate formal ones. The other hypothesis, formulated within the optimal gender assignment theory (Rice 2006), states that form and semantics contribute equally. Both hypotheses claim that the combination of formal and semantic features yields the most accurate gender identification. In this paper, we operationalize and test these hypotheses by trying to predict grammatical gender using only character-based embeddings (that capture only formal features), only context-based embeddings (that capture only semantic features) and the combination of both. We performed the experiment using data from three languages with different gender systems (French, German and Russian). Formal features are a significantly better predictor of gender than semantic ones, and the difference in prediction accuracy is very large. Overall, formal features are also significantly better than the combination of form and semantics, but the difference is very small and the results for this comparison are not entirely consistent across languages.

OriginalsprogEngelsk
TidsskriftLinguistics Vanguard
Vol/bind7
Udgave nummer1
Sider (fra-til)2020-0048
DOI
StatusUdgivet - 1 jan. 2021
Eksternt udgivetJa

Bibliografisk note

Publisher Copyright:
© 2021 Walter de Gruyter GmbH. All rights reserved.

ID: 366046076