keshan/lk_hansard_trilingual download history
keshan/lk_hansard_trilingual is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 34 times (3 in the last 7 days), and 236 times in total. It ranks #248,758 among datasets by monthly downloads.
Sri Lanka Parliament Hansard Trilingual Dataset This dataset contains text extracted from the Official Reports (Hansards) of the Parliament of Sri Lanka. The data has been processed from PDF sources, OCRed, and split into English, Sinhala, and Tamil subsets based on script detection.
Open keshan/lk_hansard_trilingual on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.