strombergnlp/nordic_langid download history

strombergnlp/nordic_langid is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 281 times (61 in the last 7 days), and 8,522 times in total. It ranks #51,683 among datasets by monthly downloads.

Automatic language identification is a challenging problem. Discriminating between closely related languages is especially difficult. This paper presents a machine learning approach for automatic language identification for the Nordic languages, which often suffer miscategorisation by existing stat

Models trained on nordic_langid

2 models list it as training data.

Open strombergnlp/nordic_langid on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.