tdr-research-team/tdr-dataset download history
tdr-research-team/tdr-dataset is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 7,140 times (7,140 in the last 7 days), and 7,140 times in total. It ranks #4,187 among datasets by monthly downloads.
TDR: Tagalog Diacritic Restoration Dataset TDR is a manually annotated corpus of contemporary Tagalog news text covering 40 frequent homographs, with diacritical marks (tuldik) restored to disambiguate meaning and pronunciation. Dataset Description Tagalog uses diacriti
Open tdr-research-team/tdr-dataset on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.