Search: onr:"swepub:oai:DiVA.org:su-192115" >
A Multi-Word Expres...
A Multi-Word Expression Dataset for Swedish
-
- Kurfali, Murathan, 1990- (author)
- Stockholms universitet,Institutionen för lingvistik
-
- Östling, Robert (author)
- Stockholms universitet,Institutionen för lingvistik
-
- Sjons, Johan (author)
- Stockholms universitet,Institutionen för lingvistik
-
show more...
-
- Wirén, Mats (author)
- Stockholms universitet,Institutionen för lingvistik
-
show less...
-
(creator_code:org_t)
- Marseille : European Language Resources Association (ELRA), 2020
- 2020
- English.
-
In: Proceedings of the 12th Conference on Language Resources and Evaluation (LREC 2020). - Marseille : European Language Resources Association (ELRA). ; , s. 4402-4409
- Related links:
-
http://www.lrec-conf...
-
show more...
-
https://urn.kb.se/re...
-
show less...
Abstract
Subject headings
Close
- We present a new set of 96 Swedish multi-word expressions annotated with degree of (non-)compositionality. In contrast to most previous compositionality datasets we also consider syntactically complex constructions and publish a formal specification of each expression. This allows evaluation of computational models beyond word bigrams, which have so far been the norm. Finally, we use the annotations to evaluate a system for automatic compositionality estimation based on distributional semantics. Our analysis of the disagreements between human annotators and the distributional model reveal interesting questions related to the perception of compositionality, and should be informative to future work in the area.
Subject headings
- NATURVETENSKAP -- Data- och informationsvetenskap -- Språkteknologi (hsv//swe)
- NATURAL SCIENCES -- Computer and Information Sciences -- Language Technology (hsv//eng)
Keyword
- multi-word expressions
- compositionality
- distributional semantic
- datorlingvistik
- Computational Linguistics
Publication and Content Type
- ref (subject category)
- kon (subject category)
To the university's database