Nodes to use Florence2 VLM for image vision tasks: object detection, captioning, segmentation and ocr
https://github.com/ttulttul/ComfyUI-Iterative-Mixer
git clone https://github.com/ttulttul/ComfyUI-Iterative-Mixer
comfy node install ComfyUI-Iterative-Mixer