TIDE-dllm/distill_wedlm_sft download history

TIDE-dllm/distill_wedlm_sft is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 128 times (18 in the last 7 days), and 1,752 times in total. It ranks #95,881 among datasets by monthly downloads.

distill_wedlm_sft — Pre-tokenized SFT mixture for WeDLM-teacher distillation Pre-tokenized SFT corpus used to train every checkpoint in the Shared-Tokenizer (Pipeline B) of the TIDE framework — i.e. the distill-WeDLM-* student checkpoints distilled from tencent/WeDLM-8B-Instruct. T

Open TIDE-dllm/distill_wedlm_sft on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.