#Florence-2
Showing 3 of 3 repositories tagged #florence-2, ranked by stars
gokayfem
ComfyUI_VLM_nodes
ComfyUI nodes for vision-language models: Qwen3-VL, Moondream 3, Florence-2, SmolVLM2, InternVL, Gemma 3, MiniCPM-V. Plus open-vocabulary detection, SAM2/SAM3 segmentation, video temporal reasoning, GGUF via llama.cpp, and hosted LLM/VLM APIs.
Score
100
★ 582
⑂ 61
+2/day
Python
Ravi-Teja-konda
Surveillance_Video_Summarizer
VLM driven tool that processes surveillance videos, extracts frames, and generates insightful annotations using a fine-tuned Florence-2 Vision-Language Model. Includes a Gradio-based interface for querying and analyzing video footage.
Score
0
★ 135
⑂ 17
—
Python
sayedmohamedscu
Vision-language-models-VLM
vision language models finetuning notebooks & use cases (Medgemma - paligemma - florence .....)
Score
0
★ 63
⑂ 15
—
Jupyter Notebook