Fetching the paper…

GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing · Around