SBB/page_extraction_dataset download history

SBB/page_extraction_dataset is an image segmentation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 11 times (4 in the last 7 days), and 229 times in total. It ranks #541,719 among datasets by monthly downloads.

Title Page Extraction Dataset Description In digitised cultural heritage items such as books, newspapers and archival records, a problem that can negatively affect OCR are black margins around a page caused by document scanning. In order to enable document layout analysis (DLA)

Open SBB/page_extraction_dataset on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.