kenhktsui/code-natural-language-classification-dataset download history
kenhktsui/code-natural-language-classification-dataset is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 648 times (62 in the last 7 days), and 11,682 times in total. It ranks #26,938 among datasets by monthly downloads.
Sampling from codeparrot/github-code under more permissive license ['mit', 'apache-2.0', 'bsd-3-clause', 'bsd-2-clause', 'cc0-1.0'] + sampling from minipile. It is intended to be used for training code natural language classifier.
Models trained on code-natural-language-classification-dataset
1 models list it as training data.
- kenhktsui/code-natural-language-fasttext-classifier 77 downloads in 30 days
Open kenhktsui/code-natural-language-classification-dataset on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.