hishab/titulm-bangla-corpus download history

hishab/titulm-bangla-corpus is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 1,923 times (245 in the last 7 days), and 24,892 times in total. It ranks #11,636 among datasets by monthly downloads.

TituLM Bangla Corpus This dataset is associated with the paper TituLLMs: A Family of Bangla LLMs with Comprehensive Benchmarking TituLM Bangla Corpus is one of the largest Bangla clean corpus prepared for pretraining, continual pretraining or fine-tuning Large Language Model(LLM) for impr

Models trained on titulm-bangla-corpus

2 models list it as training data.

Open hishab/titulm-bangla-corpus on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.