mteb/sib200.v2 download history
mteb/sib200.v2 is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 546 times (57 in the last 7 days), and 3,314 times in total. It ranks #30,974 among datasets by monthly downloads.
SIB-200 v2 SIB-200 is the largest publicly available topic classification dataset based on Flores-200 covering 205 languages and dialects annotated. The original SIB-200 v1 dataset (mteb/sib200) includes only seven labels. This v2 dataset includes all 14 labels (excluding "uncategorized")