Madras1/corpus-ptbr-v1 download history

Madras1/corpus-ptbr-v1 is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 351 times (78 in the last 7 days), and 3,278 times in total. It ranks #43,701 among datasets by monthly downloads.

🇧🇷 Corpus PT-BR v1 Um corpus em Português Brasileiro voltado para pré-treinamento e fine-tuning de LLMs. Combina dados reais curados com uma camada sintética construída para ampliar a diversidade estilística, lexical e discursiva em português. 🔄 Visão Geral do Pipeline

Models trained on corpus-ptbr-v1

3 models list it as training data.

Open Madras1/corpus-ptbr-v1 on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.