Chenhangcui/ChronoDiscovery-corpus download history

Chenhangcui/ChronoDiscovery-corpus is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 433 times (193 in the last 7 days), and 1,489 times in total. It ranks #37,063 among datasets by monthly downloads.

ChronoDiscovery Historical Training Corpus (1930–1970) Real, period-authentic full-text corpus (newspapers + books) staged by decade, for continual / time-sliced language-model training in the ChronoDiscovery project. The purpose is causal validation of LLM "surprise": if a model trai

Open Chenhangcui/ChronoDiscovery-corpus on Hugging Face