alwaysgood/financial-english-source-corpus download history
alwaysgood/financial-english-source-corpus is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 445 times (97 in the last 7 days), and 1,353 times in total. It ranks #36,279 among datasets by monthly downloads.
Financial English Source Corpus This dataset is a filtered, fuzzy-deduplicated English source-text corpus for financial-domain language-model training and translation-data generation. This version preserves the final pre-split source rows. Derived 1280-token split versions are availab
Models trained on financial-english-source-corpus
2 models list it as training data.
- alwaysgood/Qwen3.5_4B_ADS 95 downloads in 30 days
- alwaysgood/Gemma4_E2B_ADS 88 downloads in 30 days
Open alwaysgood/financial-english-source-corpus on Hugging Face