Fetching the paper…

Multimodal Self-Instruct: Synthetic Abstract Image and Visual Reasoning Instruction Using Language Model · Around