Neural morphosyntactic tagging for Rusyn

Show full item record



Permalink

http://hdl.handle.net/10138/309779

Citation

Scherrer , Y & Rabus , A 2019 , ' Neural morphosyntactic tagging for Rusyn ' , Natural Language Engineering , vol. 25 , no. 5 , 1351324919000287 , pp. 633-650 . https://doi.org/10.1017/S1351324919000287

Title: Neural morphosyntactic tagging for Rusyn
Author: Scherrer, Yves; Rabus, Achim
Contributor: University of Helsinki, Department of Digital Humanities
Date: 2019-09
Language: eng
Number of pages: 18
Belongs to series: Natural Language Engineering
ISSN: 1351-3249
URI: http://hdl.handle.net/10138/309779
Abstract: The paper presents experiments on part-of-speech and full morphological tagging of the Slavic minority language Rusyn. The proposed approach relies on transfer learning and uses only annotated resources from related Slavic languages, namely Russian, Ukrainian, Slovak, Polish, and Czech. It does not require any annotated Rusyn training data, nor parallel data or bilingual dictionaries involving Rusyn. Compared to earlier work, we improve tagging performance by using a neural network tagger and larger training data from the neighboring Slavic languages.We experiment with various data preprocessing and sampling strategies and evaluate the impact of multitask learning strategies and of pretrained word embeddings. Overall, while genre discrepancies between training and test data have a negative impact, we improve full morphological tagging by 9% absolute micro-averaged F1 as compared to previous research.
Subject: 6121 Languages
Rights:


Files in this item

Total number of downloads: Loading...

Files Size Format View
second_submission.pdf 641.0Kb PDF View/Open

This item appears in the following Collection(s)

Show full item record