Biomedical-TeMU/SPACCC_Tokenizer download history
Biomedical-TeMU/SPACCC_Tokenizer is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 4,173 times (81 in the last 7 days), and 12,425 times in total. It ranks #6,560 among datasets by monthly downloads.
The Tokenizer for Clinical Cases Written in Spanish Introduction This repository contains the tokenization model trained using the SPACCC_TOKEN corpus (https://github.com/PlanTL-SANIDAD/SPACCC_TOKEN). The model was trained using the 90% of the corpus (900 clinical cases) and te
Open Biomedical-TeMU/SPACCC_Tokenizer on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.