Fetching the paper…
Reading the bibliography…
Large Vision-Language Models (LVLMs) and Multimodal Large Language Models (MLLMs) have demonstrated outstanding performance in various general multimodal applications and have shown increasing promise in specialized domains.
3d object representations for fine-grained categorization
Jonathan Krause, Michael Stark, Jia Deng, and Li Fei-Fei · 2013
Earlier work this paper cites.
Preparing a collection of radiology examinations for distribution and retrieval
Dina Demner-Fushman, Marc D Kohli, Marc B Rosenman, Sonya E Shooshan, Laritza Rodriguez, Sameer Antani, George R Thoma, and Clement J McDonald · 2016
Earlier work this paper cites.
Insurance literacy in australia: Not knowing the value of personal insurance
Tania Driver, Mark Brimble, Brett Freudenberg, and Katherine Hunt · 2018
Earlier work this paper cites.
The impact of digitalization on the insurance value chain and the insurability of risks
Martin Eling and Martin Lehmann · 2018
Earlier work this paper cites.
An anti-fraud system for car insurance claim based on visual evidence
Pei Li, Bingyu Shen, and Weishan Dong · 2018
Earlier work this paper cites.
Towards end-to-end license plate detection and recognition: A large dataset and baseline
Zhenbo Xu, Wei Yang, Ajin Meng, Nanxue Lu, Huan Huang, Changchun Ying, and Liusheng Huang · 2018
Earlier work this paper cites.
Vqa-med: Overview of the medical visual question answering task at imageclef 2019
Asma Ben Abacha, Sadid A Hasan, Vivek V Datla, Joey Liu, Dina Demner-Fushman, and Henning Müller · 2019
Earlier work this paper cites.
Creating xbd: A dataset for assessing building damage from satellite imagery
Ritwik Gupta, Bryce Goodman, Nirav Patel, Ricky Hosfelt, Sandra Sajeev, Eric Heim, Jigar Doshi, Keane Lucas, Howie Choset, and Matthew Gaston · 2019
Earlier work this paper cites.
Decision making in personal insurance: Impact of insurance literacy
Sampath Sanjeewa Weedige, Hongbing Ouyang, Yao Gao, and Yaqing Liu · 2019
Earlier work this paper cites.
Deep neural networks and transfer learning for food crop identification in uav images
Robert Chew, Jay Rineer, Robert Beach, Maggie O’Neil, Noel Ujeneza, Daniel Lapidus, Thomas Miano, Meghan Hegarty-Craver, Jason Polly, and Dorota S Temple · 2020
Earlier work this paper cites.
Agriculture-vision: A large aerial image database for agricultural pattern analysis
Mang Tik Chiu, Xingqian Xu, Yunchao Wei, Zilong Huang, Alexander G Schwing, Robert Brunner, Hrant Khachatrian, Hovnatan Karapetyan, Ivan Dozier, Greg Rose, et al · 2020
Earlier work this paper cites.
Insurance fraud identification using computer vision and iot: a study of field fires
Srishti Sahni, Anmol Mittal, Farzil Kidwai, Ajay Tiwari, and Kanak Khandelwal · 2020
Earlier work this paper cites.
Automatic car damage assessment system: Reading and understanding videos as professional insurance inspectors
Wei Zhang, Yuan Cheng, Xin Guo, Qingpei Guo, Jian Wang, Qing Wang, Chen Jiang, Meng Wang, Furong Xu, and Wei Chu · 2020
Earlier work this paper cites.
Robust deep learning-based driver distraction detection and classification
Amal Ezzouhri, Zakaria Charouh, Mounir Ghogho, and Zouhair Guennoun · 2021
Earlier work this paper cites.
Dada: Driver attention prediction in driving accident scenarios
Jianwu Fang, Dingxin Yan, Jiahuan Qiao, Jianru Xue, and Hongkai Yu · 2021
Earlier work this paper cites.
Trodo: A public vehicle odometers dataset for computer vision
Kaouther Mouheb, Ali Yürekli, and Burcu Yılmazel · 2021
Earlier work this paper cites.
Computer vision techniques in construction: a critical review
Shuyuan Xu, Jun Wang, Wenchi Shou, Tuan Ngo, Abdul-Manan Sadick, and Xiangyu Wang · 2021
Earlier work this paper cites.
The impact of artificial intelligence along the insurance value chain and on the insurability of risks
Martin Eling, Davide Nuessle, and Julian Staubli · 2022
Earlier work this paper cites.
Automated vehicle insurance claims processing using computer vision, natural language processing
Nisaja Fernando, Abimani Kumarage, Vithyashagar Thiyaganathan, Radesh Hillary, and Lakmini Abeywardhana · 2022
Earlier work this paper cites.
Large language models can self-improve
Jiaxin Huang, Shixiang Shane Gu, Le Hou, Yuexin Wu, Xuezhi Wang, Hongkun Yu, and Jiawei Han · 2022
Earlier work this paper cites.
Emergent abilities of large language models
Jason Wei, Yi Tay, Rishi Bommasani, Colin Raffel, Barret Zoph, Sebastian Borgeaud, Dani Yogatama, Maarten Bosma, Denny Zhou, Donald Metzler, et al · 2022
Earlier work this paper cites.
Automatic chain of thought prompting in large language models
Zhuosheng Zhang, Aston Zhang, Mu Li, and Alex Smola · 2022
Earlier work this paper cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Earlier work this paper cites.
Qwen-vl: A frontier large vision-language model with versatile abilities
Jinze Bai, Shuai Bai, Shusheng Yang, Shijie Wang, Sinan Tan, Peng Wang, Junyang Lin, Chang Zhou, and Jingren Zhou · 2023
Cited alongside, same era.
Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Zhe Chen, Jiannan Wu, Wenhai Wang, Weijie Su, Guo Chen, Sen Xing, Zhong Muyan, Qinglong Zhang, Xizhou Zhu, Lewei Lu, et al · 2023
Cited alongside, same era.
Pengi: An audio language model for audio tasks
Soham Deshmukh, Benjamin Elizalde, Rita Singh, and Huaming Wang · 2023
Cited alongside, same era.
Talk2bev: Language-enhanced bird’s-eye view maps for autonomous driving
Vikrant Dewangan, Tushar Choudhary, Shivam Chandhok, Shubham Priyadarshan, Anushka Jain, Arun K Singh, Siddharth Srivastava, Krishna Murthy Jatavallabhula, and K Madhava Krishna · 2023
Cited alongside, same era.
Damages dataset
Capstone2 · 2024
Closest in time.
A survey on evaluation of large language models
Yupeng Chang, Xu Wang, Jindong Wang, Yuan Wu, Linyi Yang, Kaijie Zhu, Hao Chen, Xiaoyuan Yi, Cunxiang Wang, Yidong Wang, et al · 2024
Closest in time.
fire detection dataset
College · 2024
Closest in time.
Worker-safety dataset
computer vision · 2024
Closest in time.
dataset dashboard dataset
Dashboarddataset · 2024
Closest in time.
Vlmevalkit: An open-source toolkit for evaluating large multi-modality models, 2024
Haodong Duan, Junming Yang, Yuxuan Qiao, Xinyu Fang, Lin Chen, Yuan Liu, Xiaoyi Dong, Yuhang Zang, Pan Zhang, Jiaqi Wang, Dahua Lin, and Kai Chen · 2024
Closest in time.
Wheat growth stage challenge
GAURAV DUTTA · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Chaoyou Fu, Renrui Zhang, Haojia Lin, Zihan Wang, Timin Gao, Yongdong Luo, Yubo Huang, Zhengye Zhang, Longtian Qiu, Gaoxiang Ye, et al · 2023
Cited alongside, same era.
Chatgpt for good? on opportunities and challenges of large language models for education
Enkelejda Kasneci, Kathrin Seßler, Stefan Küchemann, Maria Bannert, Daryna Dementieva, Frank Fischer, Urs Gasser, Georg Groh, Stephan Günnemann, Eyke Hüllermeier, et al · 2023
Cited alongside, same era.
Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts
Pan Lu, Hritik Bansal, Tony Xia, Jiacheng Liu, Chunyuan Li, Hannaneh Hajishirzi, Hao Cheng, Kai-Wei Chang, Michel Galley, and Jianfeng Gao · 2023
Cited alongside, same era.
Vehicle damage severity estimation for insurance operations using in-the-wild mobile images
Dimitrios Mallios, Li Xiaofei, Niall McLaughlin, Jesus Martinez Del Rincon, Clare Galbraith, and Rory Garland · 2023
Cited alongside, same era.
Charting new territories: Exploring the geographic and geospatial capabilities of multimodal llms
Jonathan Roberts, Timo Lüddecke, Rehan Sheikh, Kai Han, and Samuel Albanie · 2023
Cited alongside, same era.
Chatgpt and other large language models are double-edged swords, 2023
Yiqiu Shen, Laura Heacock, Jonathan Elias, Keith D Hentel, Beatriu Reig, George Shih, and Linda Moy · 2023
Cited alongside, same era.
Precision viticulture dataset for detailed vineyard mapping composed of geotagged smartphone ground images, phytosanitary status, uav orthomosaics, 3d point clouds, and rtk gnss data - northern spain, july 2022, 2023
Sergio Vélez, Mar Ariza-Sentís, and João Valente · 2023
Cited alongside, same era.
Cardd: A new dataset for vision-based car damage detection
Xinkuang Wang, Wenjing Li, and Zhongcheng Wu · 2023
Cited alongside, same era.
Closest in time.
Tuning car detection dataset
f-rid nagiyev · 2024
Closest in time.
Gemini pro
Google · 2024
Closest in time.
Omnimedvqa: A new large-scale comprehensive evaluation benchmark for medical lvlm
Yutao Hu, Tianbin Li, Quanfeng Lu, Wenqi Shao, Junjun He, Yu Qiao, and Ping Luo · 2024
Closest in time.
Fall detection dataset
UTTEJ KUMAR KANDAGATLA · 2024
Closest in time.
Harnessing gpt-4v (ision) for insurance: A preliminary exploration
Chenwei Lin, Hanjia Lyu, Jiebo Luo, and Xian Xu · 2024
Closest in time.
Visual instruction tuning
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee · 2024
Closest in time.
Hello gpt-4o
OpenAI · 2024
Closest in time.
blood-pressure-monitor-display dataset
Final Project · 2024
Closest in time.
Scifibench: Benchmarking large multimodal models for scientific figure interpretation
Jonathan Roberts, Kai Han, Neil Houlsby, and Samuel Albanie · 2024
Closest in time.
Car dent scratch detection(1) dataset
Sindhu · 2024
Closest in time.
Introducing qwen-vl
Qwen Team · 2024
Closest in time.
mjdfodf-qmbuf dataset
workspace · 2024
Closest in time.
Kaining Ying, Fanqing Meng, Jin Wang, Zhiqian Li, Han Lin, Yue Yang, Hao Zhang, Wenbo Zhang, Yuqi Lin, Shuo Liu, et al · 2024
Closest in time.
SoMeLVLM: A large vision language model for social media processing
Xinnong Zhang, Haoyu Kuang, Xinyi Mou, Hanjia Lyu, Kun Wu, Siming Chen, Jiebo Luo, Xuanjing Huang, and Zhongyu Wei · 2024
Closest in time.
Hackerearth machine learning challenge: Vehicle insurance claim, 2020
HackerEarth · 2025
Closest in time.
Gpt-4v(ision) as a social media analysis engine
Hanjia Lyu, Jinfa Huang, Daoan Zhang, Yongsheng Yu, Xinyi Mou, Jinsheng Pan, Zhengyuan Yang, Zhongyu Wei, and Jiebo Luo · 2025
Closest in time.