Fetching the paper…
Reading the bibliography…
The rapid growth of user-generated content (UGC) videos has produced an urgent need for effective video quality assessment (VQA) algorithms to monitor video quality and guide optimization and recommendation procedures.
Learning generalized spatial-temporal deep feature representation for no-reference video quality assessment
Baoliang Chen, Lingyu Zhu, Guo Li, Fangbo Lu, Hongfei Fan, and Shiqi Wang · 1916
Earlier work this paper cites.
A h. 264/avc video database for the evaluation of quality metrics
Francesca De Simone, Marco Tagliasacchi, Matteo Naccari, Stefano Tubaro, and Touradj Ebrahimi · 2010
Earlier work this paper cites.
Bilibili
Bilibili Inc · 2010
Earlier work this paper cites.
Study of subjective and objective quality assessment of video
Kalpana Seshadrinathan, Rajiv Soundararajan, Alan Conrad Bovik, and Lawrence K Cormack · 2010
Earlier work this paper cites.
Making a “completely blind” image quality analyzer
Anish Mittal, Rajiv Soundararajan, and Alan C Bovik · 2012
Earlier work this paper cites.
Video quality assessment on mobile devices: Subjective, behavioral and objective studies
Anush Krishna Moorthy, Lark Kwon Choi, Alan Conrad Bovik, and Gustavo De Veciana · 2012
Earlier work this paper cites.
Methodology for the subjective assessment of the quality of television pictures
BT Series · 2012
Earlier work this paper cites.
Learning without human scores for blind image quality assessment
Wufeng Xue, Lei Zhang, and Xuanqin Mou · 2013
Earlier work this paper cites.
Blind prediction of natural video quality
Michele A Saad, Alan C Bovik, and Christophe Charrier · 2014
Earlier work this paper cites.
A completely blind video integrity oracle
Anish Mittal, Michele A Saad, and Alan C Bovik · 2015
Earlier work this paper cites.
Cvd2014—a database for evaluating no-reference video quality assessment algorithms
Mikko Nuutinen, Toni Virtanen, Mikko Vaahteranoksa, Tero Vuori, Pirkko Oittinen, and Jukka Häkkinen · 2016
Earlier work this paper cites.
Mcl-jcv: a jnd-based h. 264/avc video quality assessment dataset
Haiqiang Wang, Weihao Gan, Sudeng Hu, Joe Yuchieh Lin, Lina Jin, Longguang Song, Ping Wang, Ioannis Katsavounidis, Anne Aaron, and C-C Jay Kuo · 2016
Earlier work this paper cites.
Blind image quality assessment based on high order statistics aggregation
Jingtao Xu, Peng Ye, Qiaohong Li, Haiqing Du, Yong Liu, and David Doermann · 2016
Earlier work this paper cites.
In-capture mobile video distortions: A study of subjective behavior and objective algorithms
Deepti Ghadiyaram, Janice Pan, Alan C Bovik, Anush Krishna Moorthy, Prasanjit Panda, and Kai-Chieh Yang · 2017
Earlier work this paper cites.
The konstanz natural video database (konvid-1k)
Vlad Hosu, Franz Hahn, Mohsen Jenadeleh, Hanhe Lin, Hui Men, Tamás Szirányi, Shujun Li, and Dietmar Saupe · 2017
Earlier work this paper cites.
Large-scale study of perceptual video quality
Zeina Sinno and Alan Conrad Bovik · 2018
Earlier work this paper cites.
Slowfast networks for video recognition
Christoph Feichtenhofer, Haoqi Fan, Jitendra Malik, and Kaiming He · 2019
Earlier work this paper cites.
Two-level approach for no-reference consumer video quality assessment
Jari Korhonen · 2019
Earlier work this paper cites.
Quality assessment of in-the-wild videos
Dingquan Li, Tingting Jiang, and Ming Jiang · 2019
Earlier work this paper cites.
Youtube ugc dataset for video compression research
Yilin Wang, Sasi Inguva, and Balu Adsumilli · 2019
Earlier work this paper cites.
Ugc-video: Perceptual quality assessment of user-generated videos
Yang Li, Shengbin Meng, Xinfeng Zhang, Shiqi Wang, Yue Wang, and Siwei Ma · 2020
Earlier work this paper cites.
Icme 2021 ugc-vqa challenge
Wang Haiqiang, Li Gary, Liu Shan, and Kuo C.-C. Jay · 2021
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J Hu, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, Weizhu Chen, et al · 2021
Cited alongside, same era.
Completely blind quality assessment of user generated video content
Parimala Kancharla and Sumohana S Channappayya · 2021
Cited alongside, same era.
Spatiotemporal representation learning for blind video quality assessment
Yongxu Liu, Jinjian Wu, Leida Li, Weisheng Dong, Jinpeng Zhang, and Guangming Shi · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
Rich features for perceptual quality assessment of ugc videos
Yilin Wang, Junjie Ke, Hossein Talebi, Joong Gon Yim, Neil Birkbeck, Balu Adsumilli, Peyman Milanfar, and Feng Yang · 2021
Cited alongside, same era.
Video-llava: Learning united visual representation by alignment before projection
Bin Lin, Bin Zhu, Yang Ye, Munan Ning, Peng Jin, and Li Yuan · 2023
Later among the works it cites.
An empirical study of scaling instruct-tuned large multimodal models
Yadong Lu, Chunyuan Li, Haotian Liu, Jianwei Yang, Jianfeng Gao, and Yelong Shen · 2023
Later among the works it cites.
Internlm: A multilingual language model with progressively enhanced capabilities, 2023
InternLM Team · 2023
Later among the works it cites.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Later among the works it cites.
Depicting beyond scores: Advancing image quality assessment through multi-modal language models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
No-reference video quality assessment with heterogeneous knowledge ensemble
Jinjian Wu, Yongxu Liu, Leida Li, Weisheng Dong, and Guangming Shi · 2021
Cited alongside, same era.
Patch-vq:’patching up’the video quality problem
Zhenqiang Ying, Maniratnam Mandal, Deepti Ghadiyaram, and Alan Bovik · 2021
Cited alongside, same era.
Long short-term convolutional transformer for no-reference video quality assessment
Junyong You · 2021
Cited alongside, same era.
Predicting the quality of compressed videos with pre-existing distortions
Xiangxu Yu, Neil Birkbeck, Yilin Wang, Christos G Bampis, Balu Adsumilli, and Alan C Bovik · 2021
Cited alongside, same era.
Flashattention: Fast and memory-efficient exact attention with io-awareness
Tri Dao, Dan Fu, Stefano Ermon, Atri Rudra, and Christopher Ré · 2022
Cited alongside, same era.
Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Junnan Li, Dongxu Li, Caiming Xiong, and Steven Hoi · 2022
Cited alongside, same era.
Exploring the effectiveness of video perceptual representation in blind video quality assessment
Liang Liao, Kangmin Xu, Haoning Wu, Chaofeng Chen, Wenxiu Sun, Qiong Yan, and Weisi Lin · 2022
Cited alongside, same era.
Zhiyuan You, Zheyuan Li, Jinjin Gu, Zhenfei Yin, Tianfan Xue, and Chao Dong · 2023
Later among the works it cites.
Subjective and objective analysis of streamed gaming videos
Xiangxu Yu, Zhenqiang Ying, Neil Birkbeck, Yilin Wang, Balu Adsumilli, and Alan C Bovik · 2023
Later among the works it cites.
Capturing co-existing distortions in user-generated content for no-reference video quality assessment
Kun Yuan, Zishang Kong, Chuanchuan Zheng, Ming Sun, and Xing Wen · 2023
Later among the works it cites.
Md-vqa: Multi-dimensional quality assessment for ugc live videos
Zicheng Zhang, Wei Wu, Wei Sun, Danyang Tu, Wei Lu, Xiongkuo Min, Ying Chen, and Guangtao Zhai · 2023
Later among the works it cites.
Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Zhe Chen, Jiannan Wu, Wenhai Wang, Weijie Su, Guo Chen, Sen Xing, Muyan Zhong, Qinglong Zhang, Xizhou Zhu, Lewei Lu, et al · 2024
Closest in time.
Videollama 2: Advancing spatial-temporal modeling and audio understanding in video-llms
Zesen Cheng, Sicong Leng, Hang Zhang, Yifei Xin, Xin Li, Guanzheng Chen, Yongxin Zhu, Wenqi Zhang, Ziyang Luo, Deli Zhao, and Lidong Bing · 2024
Closest in time.
Uniprocessor: A text-induced unified low-level image processor
Huiyu Duan, Xiongkuo Min, Sijing Wu, Wei Shen, and Guangtao Zhai · 2024
Closest in time.
Lmm-vqa: Advancing video quality assessment with large multimodal models
Qihang Ge, Wei Sun, Yu Zhang, Yunhao Li, Zhongpeng Ji, Fengyu Sun, Shangling Jui, Xiongkuo Min, and Guangtao Zhai · 2024
Closest in time.
Kvq: Kwai video quality assessment for short-form videos
Yiting Lu, Xin Li, Yajing Pei, Kun Yuan, Qizhi Xie, Yunpeng Qu, Ming Sun, Chao Zhou, and Zhibo Chen · 2024
Closest in time.
Video-chatgpt: Towards detailed video understanding via large vision and language models
Muhammad Maaz, Hanoona Rasheed, Salman Khan, and Fahad Shahbaz Khan · 2024
Closest in time.
Perceptual video quality assessment: A survey
Xiongkuo Min, Huiyu Duan, Wei Sun, Yucheng Zhu, and Guangtao Zhai · 2024
Closest in time.
Analysis of video quality datasets via design of minimalistic video quality models
Wei Sun, Wen Wen, Xiongkuo Min, Long Lan, Guangtao Zhai, and Kede Ma · 2024
Closest in time.
Qwen2-vl: Enhancing vision-language model’s perception of the world at any resolution
Peng Wang, Shuai Bai, Sinan Tan, Shijie Wang, Zhihao Fan, Jinze Bai, Keqin Chen, Xuejing Liu, Jialin Wang, Wenbin Ge, Yang Fan, Kai Dang, Mengfei Du, Xuancheng Ren, Rui Men, Dayiheng Liu, Chang Zhou, Jingren Zhou, and Junyang Lin · 2024
Closest in time.
Minicpm-v: A gpt-4v level mllm on your phone
Yuan Yao, Tianyu Yu, Ao Zhang, Chongyi Wang, Junbo Cui, Hongji Zhu, Tianchi Cai, Haoyu Li, Weilin Zhao, Zhihui He, et al · 2024
Closest in time.
mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration
Qinghao Ye, Haiyang Xu, Jiabo Ye, Ming Yan, Anwen Hu, Haowei Liu, Qi Qian, Ji Zhang, and Fei Huang · 2024
Closest in time.
Descriptive image quality assessment in the wild
Zhiyuan You, Jinjin Gu, Zheyuan Li, Xin Cai, Kaiwen Zhu, Chao Dong, and Tianfan Xue · 2024
Closest in time.
Minigpt-4: Enhancing vision-language understanding with advanced large language models
Deyao Zhu, Jun Chen, Xiaoqian Shen, Xiang Li, and Mohamed Elhoseiny · 2024
Closest in time.