Chenhangcui/ChronoDiscovery-corpus download history
Chenhangcui/ChronoDiscovery-corpus is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 433 times (193 in the last 7 days), and 1,489 times in total. It ranks #37,063 among datasets by monthly downloads.
ChronoDiscovery Historical Training Corpus (1930–1970) Real, period-authentic full-text corpus (newspapers + books) staged by decade, for continual / time-sliced language-model training in the ChronoDiscovery project. The purpose is causal validation of LLM "surprise": if a model trai