Fetching the paper…
Reading the bibliography…
Recent advancements in large vision-language models (LVLMs) have led to significant progress in generating natural language descriptions for visual content and thus enhancing various applications.
A new measure of rank correlation
Maurice G Kendall. 1938 · 1938
Earlier work this paper cites.
Binary codes capable of correcting deletions, insertions, and reversals
Vladimir I Levenshtein et al. 1966 · 1966
Earlier work this paper cites.
Measuring nominal scale agreement among many raters
Joseph L Fleiss. 1971 · 1971
Earlier work this paper cites.
Figureqa: An annotated figure dataset for visual reasoning
Samira Ebrahimi Kahou, Adam Atkinson, Vincent Michalski, Ákos Kádár, Adam Trischler, and Yoshua Bengio. 2017 · 2017
Earlier work this paper cites.
Situational awareness from social media photographs using automated image captioning
João Monteiro, Asanobu Kitamoto, and Bruno Martins. 2017 · 2017
Earlier work this paper cites.
Dvqa: Understanding data visualizations via question answering
Kushal Kafle, Scott Cohen, Brian Price, and Christopher Kanan. 2018 · 2018
Earlier work this paper cites.
Factual error correction for abstractive summarization models
Meng Cao, Yue Dong, Jiapeng Wu, and Jackie Chi Kit Cheung. 2020 · 2020
Earlier work this paper cites.
Plotqa: Reasoning over scientific plots
Nitesh Methani, Pritha Ganguly, Mitesh M. Khapra, and Pratyush Kumar. 2020 · 2020
Earlier work this paper cites.
Automatic fact-guided sentence modification
Darsh Shah, Tal Schuster, and Regina Barzilay. 2020 · 2020
Earlier work this paper cites.
Improving faithfulness in abstractive summarization with contrast candidate generation and selection
Sihao Chen, Fan Zhang, Kazoo Sone, and Dan Roth. 2021 · 2021
Earlier work this paper cites.
InfoSurgeon: Cross-media fine-grained information consistency checking for fake news detection
Yi Fung, Christopher Thomas, Revanth Gangi Reddy, Sandeep Polisetty, Heng Ji, Shih-Fu Chang, Kathleen McKeown, Mohit Bansal, and Avi Sil. 2021 · 2021
Earlier work this paper cites.
Visual news: Benchmark and challenges in news image captioning
Fuxiao Liu, Yinghan Wang, Tianlu Wang, and Vicente Ordonez. 2021 · 2021
Earlier work this paper cites.
Understanding factuality in abstractive summarization with FRANK: A benchmark for factuality metrics
Artidoro Pagnoni, Vidhisha Balachandran, and Yulia Tsvetkov. 2021 · 2021
Earlier work this paper cites.
Evidence-based factual error correction
James Thorne and Andreas Vlachos. 2021 · 2021
Cited alongside, same era.
Enhancing factual consistency of abstractive summarization
Chenguang Zhu, William Hinthorn, Ruochen Xu, Qingkai Zeng, Michael Zeng, Xuedong Huang, and Meng Jiang. 2021 · 2021
Cited alongside, same era.
Learning to revise references for faithful summarization
Griffin Adams, Han-Chin Shing, Qing Sun, Christopher Winestock, Kathleen McKeown, and Noémie Elhadad. 2022 · 2022
Cited alongside, same era.
Improving factual consistency in summarization with compression-based post-editing
Alex Fabbri, Prafulla Kumar Choubey, Jesse Vig, Chien-Sheng Wu, and Caiming Xiong. 2022a · 2022
Cited alongside, same era.
QAFactEval: Improved QA-based factual consistency evaluation for summarization
Alexander Fabbri, Chien-Sheng Wu, Wenhao Liu, and Caiming Xiong. 2022b · 2022
Cited alongside, same era.
Doc2ppt: automatic presentation slides generation from scientific documents
Chartllama: A multimodal llm for chart understanding and generation
Yucheng Han, Chi Zhang, Xin Chen, Xu Yang, Zhibin Wang, Gang Yu, Bin Fu, and Hanwang Zhang. 2023 · 2023
Closest in time.
DePlot: One-shot visual language reasoning by plot-to-table translation
Fangyu Liu, Julian Eisenschlos, Francesco Piccinno, Syrine Krichene, Chenxi Pang, Kenton Lee, Mandar Joshi, Wenhu Chen, Nigel Collier, and Yasemin Altun. 2023a · 2023
Closest in time.
Unichart: A universal vision-language pretrained model for chart comprehension and reasoning
Ahmed Masry, Parsa Kavehzadeh, Xuan Long Do, Enamul Hoque, and Shafiq Joty. 2023 · 2023
Closest in time.
Gender biases in automatic evaluation metrics for image captioning
Haoyi Qiu, Zi-Yi Dou, Tianlu Wang, Asli Celikyilmaz, and Nanyun Peng. 2023a · 2023
Closest in time.
Can large language models really improve by self-critiquing their own plans?
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Tsu-Jui Fu, William Yang Wang, Daniel McDuff, and Yale Song. 2022 · 2022
Cited alongside, same era.
The battlefront of combating misinformation and coping with media bias
Yi R. Fung, Kung-Hsiang Huang, Preslav Nakov, and Heng Ji. 2022 · 2022
Cited alongside, same era.
Chart-to-text: A large-scale benchmark for chart summarization
Shankar Kantharaj, Rixie Tiffany Leong, Xiang Lin, Ahmed Masry, Megh Thakkar, Enamul Hoque, and Shafiq Joty. 2022 · 2022
Cited alongside, same era.
SummaC: Re-visiting NLI-based models for inconsistency detection in summarization
Philippe Laban, Tobias Schnabel, Paul N. Bennett, and Marti A. Hearst. 2022 · 2022
Cited alongside, same era.
ChartQA: A benchmark for question answering about charts with visual and logical reasoning
Ahmed Masry, Xuan Long Do, Jia Qing Tan, Shafiq Joty, and Enamul Hoque. 2022 · 2022
Cited alongside, same era.
FactPEGASUS: Factuality-aware pre-training and fine-tuning for abstractive summarization
David Wan and Mohit Bansal. 2022 · 2022
Cited alongside, same era.
Cross-document misinformation detection based on event graph reasoning
Xueqing Wu, Kung-Hsiang Huang, Yi Fung, and Heng Ji. 2022 · 2022
Cited alongside, same era.
Karthik Valmeekam, Matthew Marquez, and Subbarao Kambhampati. 2023 · 2023
Closest in time.
Paxion: Patching video-language foundation models with action knowledge
Zhenhailong Wang, Ansel Blume, Sha Li, Genglin Liu, Jaemin Cho, Zineng Tang, Mohit Bansal, and Heng Ji. 2023 · 2023
Closest in time.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing, et al. 2023 · 2023
Closest in time.
Enhanced chart understanding via visual language pre-training on plot table pairs
Mingyang Zhou, Yi Fung, Long Chen, Christopher Thomas, Heng Ji, and Shih-Fu Chang. 2023 · 2023
Closest in time.
Kung-Hsiang Huang, Hou Pong Chan, Yi Ren Fung, Haoyi Qiu, Mingyang Zhou, Shafiq R. Joty, Shih-Fu Chang, and Heng Ji. 2024 · 2024
Closest in time.
Fanqing Meng, Wenqi Shao, Quanfeng Lu, Peng Gao, Kaipeng Zhang, Yu Qiao, and Ping Luo. 2024 · 2024
Closest in time.
Valor-eval: Holistic coverage and faithfulness evaluation of large vision-language models
Haoyi Qiu, Wenbo Hu, Zi-Yi Dou, and Nanyun Peng. 2024 · 2024
Closest in time.
New job, new gender? measuring the social bias in image generation models
Wenxuan Wang, Haonan Bai, Jen tse Huang, Yuxuan Wan, Youliang Yuan, Haoyi Qiu, Nanyun Peng, and Michael R. Lyu. 2024 · 2024
Closest in time.