nagohachi/NDL_pdm-ocr-part2_cropped download history
nagohachi/NDL_pdm-ocr-part2_cropped is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 34 times (8 in the last 7 days), and 141 times in total. It ranks #248,758 among datasets by monthly downloads.
NDL Public Domain OCR Dataset — Line-Cropped (WebDataset) This dataset is a derivative of the Public Domain OCR Training Dataset (FY2021) published by the National Diet Library of Japan (国立国会図書館). Each sample is a line-level crop of a historical document page, paired with its OCR text tra
Open nagohachi/NDL_pdm-ocr-part2_cropped on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.