storytracer/LoC-PD-Books download history

storytracer/LoC-PD-Books is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 1,644 times (477 in the last 7 days), and 19,040 times in total. It ranks #13,216 among datasets by monthly downloads.

Library of Congress Public Domain Books (English) This dataset contains more than 140,000 English books (~ 8 billion words) digitised by the Library of Congress (LoC) that are in the public domain in the United States. The dataset was compiled by Sebastian Majstorovic. Curation me

Models trained on LoC-PD-Books

15 models list it as training data.

Open storytracer/LoC-PD-Books on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.