Fetching the paper…
Reading the bibliography…
While abundant research has been conducted on improving high-level visual understanding and reasoning capabilities of large multimodal models~(LMMs), their visual quality assessment~(IQA) ability has been relatively under-explored.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, and et al
1901
Earlier work this paper cites.
T. L. Saaty and L. G. Vargas, “Inconsistency and rank preservation,” Journal of Mathematical Psychology , vol. 28, no. 2, pp. 205–214, 1984
1984
Earlier work this paper cites.
VQEG, “Final report from the video quality experts group on the validation of objective models of video quality assessment,” 2000. [Online]. Available: http://www.vqeg.org
2000
Earlier work this paper cites.
B. T. ITU-R, “Methodology for the subjective assessment of the quality of television pictures,” https://www.itu.int/rec/R-REC-BT.500 , 2002
2002
Earlier work this paper cites.
R. Herbrich, T. Minka, and T. Graepel, “Trueskill tm {}^{\textsc{tm}} : A bayesian skill rating system,” in Neural Information Processing Systems , 2007, pp. 569–576
2007
Earlier work this paper cites.
E. C. Larson and D. M. Chandler, “Most apparent distortion: Full-reference image quality assessment and the role of strategy,” Journal of Electronic Imaging , vol. 19, no. 1, pp. 011 006–011 027, 2010
2010
Earlier work this paper cites.
K. Tsukida and M. R. Gupta, “How to analyze paired comparison data,” Technical Report UWEETR-2011-0004, University of Washington, 2011
2011
Earlier work this paper cites.
S. Winkler, “Analysis of public image and video databases for quality assessment,” IEEE Journal of Selected Topics in Signal Processing , vol. 6, no. 6, pp. 616–625, 2012
2012
Earlier work this paper cites.
A. Mittal, R. Soundararajan, and A. C. Bovik, “Making a ‘completely blind’ image quality analyzer,” IEEE Signal Processing Letters , vol. 20, no. 3, pp. 209–212, 2013
2013
Earlier work this paper cites.
D. Ghadiyaram and A. C. Bovik, “Massive online crowdsourced study of subjective and objective picture quality,” IEEE Trans. Image Processing , vol. 25, no. 1, pp. 372–387, 2016
2016
Earlier work this paper cites.
Q. Wu, H. Li, K. N. Ngan, and K. Ma, “Blind image quality assessment using local consistency aware retriever and uncertainty aware evaluator,” IEEE Trans. Circuits and Systems for Video Tech. , vol. 28, no. 9, pp. 2078–2089, 2017
2017
Earlier work this paper cites.
H. Lin, V. Hosu, and D. Saupe, “KADID-10k: A large-scale artificially distorted IQA database,” in IEEE Int. Conf. Multimedia and Expo , 2019, pp. 1–3
2019
Cited alongside, same era.
V. Hosu, H. Lin, T. Sziranyi, and D. Saupe, “KonIQ-10k: An ecologically valid database for deep learning of blind image quality assessment,” IEEE Trans. Image Processing , vol. 29, pp. 4041–4056, 2020
2020
Cited alongside, same era.
Y. Fang, H. Zhu, Y. Zeng, K. Ma, and Z. Wang, “Perceptual quality assessment of smartphone photography,” in IEEE Int. Conf. Computer Vision and Pattern Recognition , 2020, pp. 3677–3686
2020
Cited alongside, same era.
2020
Cited alongside, same era.
OpenAI, “GPT-4V(ision) system card,” https://cdn.openai.com/papers/GPTV_System_Card.pdf/ , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
H. Wu, Z. Zhang, E. Zhang, C. Chen, L. Liao, A. Wang, K. Xu, C. Li, J. Hou, G. Zhai et al. , “Q-Instruct: Improving low-level visual abilities for multi-modality foundation models,” CoRR , vol. abs/:2311.06783, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
W. Zhang, K. Ma, J. Yan, D. Deng, and Z. Wang, “Blind image quality assessment using a deep bilinear convolutional neural network,” IEEE Trans. Circuits and Systems for Video Tech. , vol. 30, no. 1, pp. 36–47, 2020
2020
Cited alongside, same era.
Y. Li, S. Wang, X. Zhang, S. Wang, S. Ma, and Y. Wang, “Quality assessment of end-to-end learned image compression: The benchmark and objective measure,” in ACM Int. Conf. on Multimedia , 2021, pp. 4297–4305
2021
Cited alongside, same era.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, F. Xia, E. Chi, Q. V. Le, D. Zhou, and et al
2022
Cited alongside, same era.
OpenAI, “GPT-4 technical report,” CoRR , vol. abs/2303.08774, 2023
2023
Cited alongside, same era.
H. Face, “Introducing IDEFICS: An open reproduction of state-of-the-art visual language model,” https://huggingface.co/blog/idefics/ , 2023
2023
Cited alongside, same era.
Q. Ye, H. Xu, G. Xu, J. Ye, M. Yan, Y. Zhou, J. Wang, A. Hu, P. Shi, Y. Shi, and et al
2023
Cited alongside, same era.
P. Zhang, X. D. B. Wang, Y. Cao, C. Xu, L. Ouyang, Z. Zhao, S. Ding, S. Zhang, and et al
2023
Cited alongside, same era.
2023
Later among the works it cites.
L. Zhao, W. Zheng, Y. Duan, J. Zhou, and J. Lu, “SPTR: Structure-preserving transformer for unsupervised indoor depth completion,” IEEE Trans. Circuits and Systems for Video Tech. , to appear 2023
2023
Later among the works it cites.
H. Zhu, B. Chen, L. Zhu, and S. Wang, “Learning spatiotemporal interactions for user-generated video quality assessment,” IEEE Trans. Circuits and Systems for Video Tech. , vol. 33, no. 3, pp. 1031–1042, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Z. Fang, A. Ignatov, E. Zamfir, and R. Timofte, “SQAD: Automatic smartphone camera quality assessment and benchmarking,” in IEEE Int. Conf. on Computer Vision , 2023, pp. 20 532–20 542
2023
Later among the works it cites.