Fetching the paper…

Bi-VLDoc: Bidirectional Vision-Language Modeling for Visually-Rich Document Understanding · Around