Fetching the paper…

Advancing Large Multi-modal Models with Explicit Chain-of-Reasoning and Visual Question Generation · Around