Ayuda
Ir al contenido

Dialnet


An automatic part-of-speech tagger for Middle Low German

    1. [1] Ghent University

      Ghent University

      Arrondissement Gent, Bélgica

  • Localización: International journal of corpus linguistics, ISSN-e 1569-9811, ISSN 1384-6655, Vol. 22, Nº 1, 2017, págs. 107-140
  • Idioma: inglés
  • Texto completo no disponible (Saber más ...)
  • Resumen
    • Syntactically annotated corpora are highly important for enabling large-scale diachronic and diatopic language research. Such corpora have recently been developed for a variety of historical languages, or are still under development. One of those under development is the fully tagged and parsed Corpus of Historical Low German (CHLG), which is aimed at facilitating research into the highly under-researched diachronic syntax of Low German. The present paper reports on a crucial step in creating the corpus, viz. the creation of a part-of-speech tagger for Middle Low German (MLG). Having been transmitted in several non-standardised written varieties, MLG poses a challenge to standard POS taggers, which usually rely on normalized spelling. We outline the major issues faced in the creation of the tagger and present our solutions to them.


Fundación Dialnet

Dialnet Plus

  • Más información sobre Dialnet Plus

Opciones de compartir

Opciones de entorno