jpwahle/autoencoder-paraphrase-dataset download history

jpwahle/autoencoder-paraphrase-dataset is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 59 times (13 in the last 7 days), and 3,360 times in total. It ranks #164,922 among datasets by monthly downloads.

Dataset Card for Machine Paraphrase Dataset (MPC) Dataset Summary The Autoencoder Paraphrase Corpus (APC) consists of ~200k examples of original, and paraphrases using three neural language models. It uses three models (BERT, RoBERTa, Longformer) on three source texts (Wikipedi

Models trained on autoencoder-paraphrase-dataset

1 models list it as training data.

Open jpwahle/autoencoder-paraphrase-dataset on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.