Hemanth-thunder/tamil-lemmatizer-dataset-12k download history

Hemanth-thunder/tamil-lemmatizer-dataset-12k is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 13 times (5 in the last 7 days), and 147 times in total. It ranks #492,381 among datasets by monthly downloads.

A high-quality dataset curated for training and evaluating Tamil lemmatization models. This dataset contains inflected Tamil word forms paired with their base lemma, enabling machine learning models to learn morphological normalization for Tamil. Dataset Overview Tamil is a morphologicall

Open Hemanth-thunder/tamil-lemmatizer-dataset-12k on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.