microsoft
Microsoft
https://github.com/microsoft
https://github.com/microsoft
Models

microsoft / florence-2-large/dense-region-caption
Detect the salient regions in a photo and get back an annotated image where each region is boxed and labeled with a short natural-language caption of what it contains.

microsoft / florence-2-large/ocr-with-region
Locate and read the text in a photo and get back an annotated image with each detected text region boxed and labeled with what it says.
microsoft / florence-2-large/caption
Generate a concise one-sentence caption describing any photo — no prompt needed.

microsoft / florence-2-large/object-detection
Detect and label every object in a photo and get back an annotated image with bounding boxes drawn on it.
