yuhuanstudio/twdict_pretrain download history

yuhuanstudio/twdict_pretrain is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 21 times (5 in the last 7 days), and 964 times in total. It ranks #361,218 among datasets by monthly downloads.

Dataset Card for "yuhuanstudio/twdict_pretrain" 資料集摘要 本資料集將「成語典」與「重編字典」的內容合併為單一資料集,提供繁體中文的詞彙、成語與其解釋、用例等資訊。此資料集適用於大型語言模型(LLM)預訓練以及自然語言處理(NLP)各類應用,如關鍵字抽取、問答系統、語義分析等。 原始資料來源: 《重編國語辭典修訂本》 《成語典》 內容說明 • 數據來源: – 「成語典」:收錄常見與罕見成語,含釋義、典故、例句等相關資訊 – 「重編字典」:收錄繁體中文詞彙與解釋,含詞性、詞義、

Open yuhuanstudio/twdict_pretrain on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.