renhehuang/formosa-vlm-knowledge-v3 download history
renhehuang/formosa-vlm-knowledge-v3 is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 7 times (2 in the last 7 days), and 95 times in total. It ranks #677,853 among datasets by monthly downloads.
台灣視覺知識庫 v3 (Formosa VLM Knowledge Base) 概述 台灣視覺知識庫 (Visual RAG Knowledge Base) 是 Formosa-VLM 的核心模組, 透過 CLIP 視覺檢索 + 知識注入,讓 VLM 能正確辨識台灣在地事物(建築、美食、寺廟、自然景觀等)。 此資料集為 **備援用 (backup)**,包含完整的實體 metadata、參考圖片、CLIP embeddings。 資料集統計 指標 數值 實體數 253 分類數 12 參考圖片 164