Open-source fibo-scene-analyzer outputs structured captions, bboxes and poses in one model

linoy_tsaban · x · 2026-10-06

The team behind fibo-scene-analyzer released an open-source image understanding model fine-tuned from Qwen3.6-35B-A3B. A single model produces highly detailed structured captions, RGB colors, detection bounding boxes, human pose, and more — aimed at fine-grained structured image annotation.

Original post →

More from Multimodal

Multimodal channel →