Skip to search boxSkip to navigationSkip to main content

Discriminating Between Similar Nordic Languages

Research Output:
Conference Article in Proceeding or Book/Report chapter
Article in proceedings
Peer-review

Open access

Publication Information

Output type

Research Output:
Conference Article in Proceeding or Book/Report chapter
Article in proceedings
Peer-review

Original language

English

Pages from-to (Number of pages)

Pages 67–75

Publication milestones

  • Published - 20/04/2021

Publication status

Published - 20/04/2021

Publisher

Association for Computational Linguistics, United States

Host publication title

Proceedings of the Eighth Workshop on NLP for Similar Languages, Varieties and Dialects

Abstract

Automatic language identification is a challenging problem. Discriminating between closely related languages is especially difficult. This paper presents a machine learning approach for automatic language identification for the Nordic languages, which often suffer miscategorisation by existing state-of-the-art tools. Concretely we will focus on discrimination between six Nordic languages: Danish, Swedish, Norwegian (Nynorsk), Norwegian (Bokmål), Faroese and Icelandic.

Related Event

Title

Workshop on NLP for Similar Languages, Varieties and Dialects

Event type

Workshop

Date

20/04/2021 - 20/04/2021

Location

VIRTUAL