strombergnlp/twitter_pos_vcb download history

strombergnlp/twitter_pos_vcb is a token classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 18 times (5 in the last 7 days), and 2,078 times in total. It ranks #402,156 among datasets by monthly downloads.

Part-of-speech information is basic NLP task. However, Twitter text is difficult to part-of-speech tag: it is noisy, with linguistic errors and idiosyncratic style. This data is the vote-constrained bootstrapped data generate to support state-of-the-art results. The data is about 1.5 million Englis

Open strombergnlp/twitter_pos_vcb on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.