shofikul-1234/titulm-bangla-corpus download history

shofikul-1234/titulm-bangla-corpus is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 996 times (162 in the last 7 days), and 2,477 times in total. It ranks #19,383 among datasets by monthly downloads.

TituLM Bangla Corpus This dataset is associated with the paper TituLLMs: A Family of Bangla LLMs with Comprehensive Benchmarking TituLM Bangla Corpus is one of the largest Bangla clean corpus prepared for pretraining, continual pretraining or fine-tuning Large Language Model(LLM) for

Open shofikul-1234/titulm-bangla-corpus on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.