josefbednar/11m-czech-sentences download history

josefbednar/11m-czech-sentences is a token classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 16 times (4 in the last 7 days), and 270 times in total. It ranks #433,426 among datasets by monthly downloads.

Dataset Card for Dataset Name This dataset contains 11 million raw Czech sentences. Dataset Description 11m-czech-sentences was created by filtering out sentences which contain at least one comma from the SYN2006PUB corpus. Direct Use This dataset is suitable

Open josefbednar/11m-czech-sentences on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.