loganbolton/sketchvlm-maze-navigation download history
loganbolton/sketchvlm-maze-navigation is an image text to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 62 times (5 in the last 7 days), and 331 times in total. It ranks #159,286 among datasets by monthly downloads.
SketchVLM: Maze Navigation This dataset is associated with the paper: SketchVLM: Vision Language Models Can Annotate Images to Explain Thoughts and Guide Users. SketchVLM is a training-free, model-agnostic framework that enables Vision-Language Models (VLMs) to produce non-destructive, ed
Open loganbolton/sketchvlm-maze-navigation on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.