SBB/page_extraction_dataset download history
SBB/page_extraction_dataset is an image segmentation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 11 times (4 in the last 7 days), and 229 times in total. It ranks #541,719 among datasets by monthly downloads.
Title Page Extraction Dataset Description In digitised cultural heritage items such as books, newspapers and archival records, a problem that can negatively affect OCR are black margins around a page caused by document scanning. In order to enable document layout analysis (DLA)
Open SBB/page_extraction_dataset on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.