Florence 2
π
845
Generate captions, detections, and segmentations from images
Generate captions, detections, and segmentations from images
Segment objects in images or videos using text prompts
Generate object masks and masked video from your MP4
Segment and track objects in videos
In-browser tool calling, powered by Transformers.js