Fetching the paper…

RelationVLM: Making Large Vision-Language Models Understand Visual Relations · Around