Fetching the paper…

Instruction-guided Multi-Granularity Segmentation and Captioning with Large Multimodal Model · Around