kenhktsui/code-natural-language-classification-dataset download history

kenhktsui/code-natural-language-classification-dataset is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 648 times (62 in the last 7 days), and 11,682 times in total. It ranks #26,938 among datasets by monthly downloads.

Sampling from codeparrot/github-code under more permissive license ['mit', 'apache-2.0', 'bsd-3-clause', 'bsd-2-clause', 'cc0-1.0'] + sampling from minipile. It is intended to be used for training code natural language classifier.

Models trained on code-natural-language-classification-dataset

1 models list it as training data.

Open kenhktsui/code-natural-language-classification-dataset on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.