rifathridoy/bengali-ocr-synthetic download history
rifathridoy/bengali-ocr-synthetic is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 1,078 times (85 in the last 7 days), and 12,768 times in total. It ranks #18,274 among datasets by monthly downloads.
Bengali OCR Synthetic Dataset A high-quality synthetic Bengali OCR dataset for fine-tuning vision-language models like DeepSeek-OCR 2. Generated using 100+ professional Bengali Unicode fonts and 13K+ unique Bengali words with advanced text rendering via FreeType and HarfBuzz. Data
Open rifathridoy/bengali-ocr-synthetic on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.