Fetching the paper…
Reading the bibliography…
The explosive growth of videos on streaming media platforms has underscored the urgent need for effective video quality assessment (VQA) algorithms to monitor and perceptually optimize the quality of streaming videos.
Reduced-reference video quality assessment using discriminative local harmonic strength with motion consideration
Irwan P Gunawan and Mohammed Ghanbari · 2008
Earlier work this paper cites.
Prediction of transmission distortion for wireless video communication: Analysis
Zhifeng Chen and Dapeng Wu · 2011
Earlier work this paper cites.
No-reference blur assessment of digital pictures based on multifeature classifiers
Alexandre Ciancio, André Luiz N Targino Targino da Costa, Eduardo A. B. da Silva, Amir Said, Ramin Samadani, and Pere Obrador · 2011
Earlier work this paper cites.
Im2text: Describing images using 1 million captioned photographs
Vicente Ordonez, Girish Kulkarni, and Tamara Berg · 2011
Earlier work this paper cites.
No-reference image quality assessment in the spatial domain
Anish Mittal, Anush Krishna Moorthy, and Alan Conrad Bovik · 2012
Earlier work this paper cites.
Making a “completely blind” image quality analyzer
Anish Mittal, Rajiv Soundararajan, and Alan C Bovik · 2012
Earlier work this paper cites.
Blind image quality assessment: A natural scene statistics approach in the dct domain
Michele A Saad, Alan C Bovik, and Christophe Charrier · 2012
Earlier work this paper cites.
Video quality assessment by reduced reference spatio-temporal entropic differencing
Rajiv Soundararajan and Alan C Bovik · 2012
Earlier work this paper cites.
Tv-l1 optical flow estimation
Javier Sánchez Pérez, Enric Meinhardt-Llopis, and Gabriele Facciolo · 2013
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart Van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
Blind prediction of natural video quality
Michele A Saad, Alan C Bovik, and Christophe Charrier · 2014
Earlier work this paper cites.
No-reference video quality assessment based on artifact measurement and statistical analysis
Kongfeng Zhu, Chengqing Li, Vijayan Asari, and Dietmar Saupe · 2014
Earlier work this paper cites.
Massive online crowdsourced study of subjective and objective picture quality
Deepti Ghadiyaram and Alan C Bovik · 2015
Earlier work this paper cites.
The analysis of image contrast: From quality assessment to automatic enhancement
Ke Gu, Guangtao Zhai, Weisi Lin, and Min Liu · 2015
Earlier work this paper cites.
A completely blind video integrity oracle
Anish Mittal, Michele A Saad, and Alan C Bovik · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
An optical flow-based full reference video quality assessment algorithm
K Manasa and Sumohana S Channappayya · 2016
Earlier work this paper cites.
Yfcc100m: The new data in multimedia research
Bart Thomee, David A Shamma, Gerald Friedland, Benjamin Elizalde, Karl Ni, Douglas Poland, Damian Borth, and Li-Jia Li · 2016
Earlier work this paper cites.
In-capture mobile video distortions: A study of subjective behavior and objective algorithms
Deepti Ghadiyaram, Janice Pan, Alan C Bovik, Anush Krishna Moorthy, Prasanjit Panda, and Kai-Chieh Yang · 2017
Earlier work this paper cites.
Spatiotemporal feature integration and model fusion for full reference video quality assessment
Christos G Bampis, Zhi Li, and Alan C Bovik · 2018
Earlier work this paper cites.
End-to-end blind quality assessment of compressed videos using deep neural networks
Wentao Liu, Zhengfang Duanmu, and Zhou Wang · 2018
Earlier work this paper cites.
Color-sensitivity-based combined psnr for objective video quality assessment
Xiwu Shang, Jie Liang, Guozhong Wang, Haiwu Zhao, Chengjia Wu, and Chang Lin · 2018
Earlier work this paper cites.
Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image captioning
Piyush Sharma, Nan Ding, Sebastian Goodman, and Radu Soricut · 2018
Earlier work this paper cites.
Large-scale study of perceptual video quality
Zeina Sinno and Alan Conrad Bovik · 2018
Earlier work this paper cites.
Slowfast networks for video recognition
Christoph Feichtenhofer, Haoqi Fan, Jitendra Malik, and Kaiming He · 2019
Earlier work this paper cites.
Two-level approach for no-reference consumer video quality assessment
Jari Korhonen · 2019
Earlier work this paper cites.
Quality assessment of in-the-wild videos
Dingquan Li, Tingting Jiang, and Ming Jiang · 2019
Earlier work this paper cites.
Quality assessment of in-the-wild videos
Dingquan Li, Tingting Jiang, and Ming Jiang · 2019
Cited alongside, same era.
Youtube ugc dataset for video compression research
Yilin Wang, Sasi Inguva, and Balu Adsumilli · 2019
Cited alongside, same era.
Image quality assessment: Unifying structure and texture similarity
Keyan Ding, Kede Ma, Shiqi Wang, and Eero P Simoncelli · 2020
Cited alongside, same era.
Perceptual quality assessment of smartphone photography
Yuming Fang, Hanwei Zhu, Yan Zeng, Kede Ma, and Zhou Wang · 2020
Cited alongside, same era.
The konstanz natural video database, 2020
V Hosu, F Hahn, M Jenadeleh, H Lin, H Men, T Szirányi, S Li, and D Saupe · 2020
Cited alongside, same era.
Koniq-10k: An ecologically valid database for deep learning of blind image quality assessment
Vlad Hosu, Hanhe Lin, Tamas Sziranyi, and Dietmar Saupe · 2020
Cited alongside, same era.
Llama-vid: An image is worth 2 tokens in large language models
Yanwei Li, Chengyao Wang, and Jiaya Jia · 2023
Later among the works it cites.
Video-llava: Learning united visual representation by alignment before projection
Bin Lin, Bin Zhu, Yang Ye, Munan Ning, Peng Jin, and Li Yuan · 2023
Later among the works it cites.
Improved baselines with visual instruction tuning
Haotian Liu, Chunyuan Li, Yuheng Li, and Yong Jae Lee · 2023
Later among the works it cites.
Ada-dqa: Adaptive diverse quality-aware feature acquisition for video quality assessment
Hongbo Liu, Mingda Wu, Kun Yuan, Ming Sun, Yansong Tang, Chuanchuan Zheng, Xing Wen, and Xiu Li · 2023
Later among the works it cites.
One for all: Video conversation is feasible without video instruction tuning, 2023
Ruyang Liu, Chen Li, Yixiao Ge, Ying Shan, Thomas H. Li, and Ge Li · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Scaling laws for neural language models
Jared Kaplan, Sam McCandlish, Tom Henighan, Tom B Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei · 2020
Cited alongside, same era.
From patches to pictures (paq-2-piq): Mapping the perceptual space of picture quality
Zhenqiang Ying, Haoran Niu, Praful Gupta, Dhruv Mahajan, Deepti Ghadiyaram, and Alan Bovik · 2020
Cited alongside, same era.
Perceiver: General perception with iterative attention
Andrew Jaegle, Felix Gimeno, Andy Brock, Oriol Vinyals, Andrew Zisserman, and Joao Carreira · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
Laion-400m: Open dataset of clip-filtered 400 million image-text pairs
Christoph Schuhmann, Richard Vencu, Romain Beaumont, Robert Kaczmarczyk, Clayton Mullis, Aarush Katta, Theo Coombes, Jenia Jitsev, and Aran Komatsuzaki · 2021
Cited alongside, same era.
Deep learning based full-reference and no-reference quality assessment models for compressed ugc videos
Wei Sun, Tao Wang, Xiongkuo Min, Fuwang Yi, and Guangtao Zhai · 2021
Cited alongside, same era.
Later among the works it cites.
Video-chatgpt: Towards detailed video understanding via large vision and language models
Muhammad Maaz, Hanoona Rasheed, Salman Khan, and Fahad Shahbaz Khan · 2023
Later among the works it cites.
Exploring video quality assessment on user generated contents from aesthetic and technical perspectives
Haoning Wu, Erli Zhang, Liang Liao, Chaofeng Chen, Jingwen Hou, Annan Wang, Wenxiu Sun, Qiong Yan, and Weisi Lin · 2023
Later among the works it cites.
Q-bench: A benchmark for general-purpose foundation models on low-level vision
Haoning Wu, Zicheng Zhang, Erli Zhang, Chaofeng Chen, Liang Liao, Annan Wang, Chunyi Li, Wenxiu Sun, Qiong Yan, Guangtao Zhai, et al · 2023
Later among the works it cites.
Q-instruct: Improving low-level visual abilities for multi-modality foundation models
Haoning Wu, Zicheng Zhang, Erli Zhang, Chaofeng Chen, Liang Liao, Annan Wang, Kaixin Xu, Chunyi Li, Jingwen Hou, Guangtao Zhai, et al · 2023
Later among the works it cites.
Q-align: Teaching lmms for visual scoring via discrete text-defined levels
Haoning Wu, Zicheng Zhang, Weixia Zhang, Chaofeng Chen, Liang Liao, Chunyi Li, Yixuan Gao, Annan Wang, Erli Zhang, Wenxiu Sun, et al · 2023
Later among the works it cites.
mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration
Qinghao Ye, Haiyang Xu, Jiabo Ye, Ming Yan, Haowei Liu, Qi Qian, Ji Zhang, Fei Huang, and Jingren Zhou · 2023
Later among the works it cites.
Video-llama: An instruction-tuned audio-visual language model for video understanding
Hang Zhang, Xin Li, and Lidong Bing · 2023
Later among the works it cites.
Pan Zhang, Xiaoyi Dong Bin Wang, Yuhang Cao, Chao Xu, Linke Ouyang, Zhiyuan Zhao, Shuangrui Ding, Songyang Zhang, Haodong Duan, Hang Yan, et al · 2023
Later among the works it cites.
Llama-adapter: Efficient fine-tuning of language models with zero-init attention
Renrui Zhang, Jiaming Han, Aojun Zhou, Xiangfei Hu, Shilin Yan, Pan Lu, Hongsheng Li, Peng Gao, and Yu Qiao · 2023
Later among the works it cites.
Md-vqa: Multi-dimensional quality assessment for ugc live videos
Zicheng Zhang, Wei Wu, Wei Sun, Danyang Tu, Wei Lu, Xiongkuo Min, Ying Chen, and Guangtao Zhai · 2023
Later among the works it cites.
Videollama 2: Advancing spatial-temporal modeling and audio understanding in video-llms
Zesen Cheng, Sicong Leng, Hang Zhang, Yifei Xin, Xin Li, Guanzheng Chen, Yongxin Zhu, Wenqi Zhang, Ziyang Luo, Deli Zhao, et al · 2024
Closest in time.
Abhimanyu Dubey, Abhinav Jauhri, Abhinav Pandey, Abhishek Kadian, Ahmad Al-Dahle, Aiesha Letman, Akhil Mathur, Alan Schelten, Amy Yang, Angela Fan, et al · 2024
Closest in time.
Llava-next-interleave: Tackling multi-image, video, and 3d in large multimodal models
Feng Li, Renrui Zhang, Hao Zhang, Yuanhan Zhang, Bo Li, Wei Li, Zejun Ma, and Chunyuan Li · 2024
Closest in time.
Videochat: Chat-centric video understanding, 2024
KunChang Li, Yinan He, Yi Wang, Yizhuo Li, Wenhai Wang, Ping Luo, Yali Wang, Limin Wang, and Yu Qiao · 2024
Closest in time.
Vila: On pre-training for visual language models
Ji Lin, Hongxu Yin, Wei Ping, Pavlo Molchanov, Mohammad Shoeybi, and Song Han · 2024
Closest in time.
Visual instruction tuning
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee · 2024
Closest in time.
Perceptual video quality assessment: A survey
Xiongkuo Min, Huiyu Duan, Wei Sun, Yucheng Zhu, and Guangtao Zhai · 2024
Closest in time.
Analysis of video quality datasets via design of minimalistic video quality models
Wei Sun, Wen Wen, Xiongkuo Min, Long Lan, Guangtao Zhai, and Kede Ma · 2024
Closest in time.
Enhancing blind video quality assessment with rich quality-aware features
Wei Sun, Haoning Wu, Zicheng Zhang, Jun Jia, Zhichao Zhang, Linhan Cao, Qiubo Chen, Xiongkuo Min, Weisi Lin, and Guangtao Zhai · 2024
Closest in time.
Ptm-vqa: Efficient video quality assessment leveraging diverse pretrained models from the wild
Kun Yuan, Hongbo Liu, Mading Li, Muyi Sun, Ming Sun, Jiachao Gong, Jinhua Hao, Chao Zhou, and Yansong Tang · 2024
Closest in time.
A benchmark for multi-modal foundation models on low-level vision: from single images to pairs
Zicheng Zhang, Haoning Wu, Erli Zhang, Guangtao Zhai, and Weisi Lin · 2024
Closest in time.