Fetching the paper…
Reading the bibliography…
Ultra-low bitrate image compression is a challenging and demanding topic.
Mpeg: A video compression standard for multimedia applications
Didier Le Gall · 1991
Earlier work this paper cites.
The jpeg 2000 still image compression standard
Athanassios N. Skodras, Charilaos A. Christopoulos, and Touradj Ebrahimi · 2001
Earlier work this paper cites.
Methodology for the subjective assessment of the quality of television pictures
I. T. Union · 2002
Earlier work this paper cites.
Overview of the h.264/avc video coding standard
Thomas Wiegand, Gary J. Sullivan, Gisle Bjøntegaard, and Ajay Luthra · 2003
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Zhou Wang · 2004
Earlier work this paper cites.
Overview of the high efficiency video coding (hevc) standard
Gary J. Sullivan, Jens-Rainer Ohm, Woojin Han, and Thomas Wiegand · 2012
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Webpage saliency
Chengyao Shen and Qi Zhao · 2014
Earlier work this paper cites.
Convolutional neural networks for no-reference image quality assessment
Le Kang, Peng Ye, Yi Li, and David Doermann · 2014
Earlier work this paper cites.
Esim: Edge similarity for screen content image quality assessment
Zhangkai Ni, Lin Ma, Huanqiang Zeng, Jing Chen, Canhui Cai, and Kai-Kuang Ma · 2017
Earlier work this paper cites.
Unified blind quality assessment of compressed natural, graphic, and screen content images
Xiongkuo Min, Kede Ma, Ke Gu, Guangtao Zhai, Zhou Wang, and Weisi Lin · 2017
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2017
Earlier work this paper cites.
The perception-distortion tradeoff
Yochai Blau and Tomer Michaeli · 2018
Earlier work this paper cites.
The unreasonable effectiveness of deep features as a perceptual metric
Richard Zhang, Phillip Isola, Alexei A Efros, Eli Shechtman, and Oliver Wang · 2018
Earlier work this paper cites.
Pieapp: Perceptual image-error assessment through pairwise preference
Ekta Prashnani, Hong Cai, Yasamin Mostofi, and Pradeep Sen · 2018
Earlier work this paper cites.
Blind image quality assessment using a deep bilinear convolutional neural network
Weixia Zhang, Kede Ma, Jia Yan, Dexiang Deng, and Zhou Wang · 2018
Earlier work this paper cites.
Rethinking lossy compression: The rate-distortion-perception tradeoff
Yochai Blau and Tomer Michaeli · 2019
Earlier work this paper cites.
Nonlinear transform coding
Johannes Ballé, Philip A Chou, David Minnen, Saurabh Singh, Nick Johnston, Eirikur Agustsson, Sung Jin Hwang, and George Toderici · 2020
Earlier work this paper cites.
Image quality assessment: Unifying structure and texture similarity
Keyan Ding, Kede Ma, Shiqi Wang, and Eero P Simoncelli · 2020
Earlier work this paper cites.
Blindly assess image quality in the wild guided by a self-adaptive hyper network
Shaolin Su, Qingsen Yan, Yu Zhu, Cheng Zhang, Xin Ge, Jinqiu Sun, and Yanning Zhang · 2020
Earlier work this paper cites.
Overview of the versatile video coding (vvc) standard and its applications
Benjamin Bross, Ye-Kui Wang, Yan Ye, Shan Liu, Jianle Chen, Gary J. Sullivan, and Jens-Rainer Ohm · 2021
Earlier work this paper cites.
Cross modal compression: Towards human-comprehensible semantic compression
Jiguo Li, Chuanmin Jia, Xinfeng Zhang, Siwei Ma, and Wen Gao · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Earlier work this paper cites.
Photorealistic text-to-image diffusion models with deep language understanding
Chitwan Saharia, William Chan, Saurabh Saxena, Lala Li, Jay Whang, Emily L Denton, Kamyar Ghasemipour, Raphael Gontijo Lopes, Burcu Karagol Ayan, Tim Salimans, et al · 2022
Earlier work this paper cites.
Ntire 2022 challenge on perceptual image quality assessment
Jinjin Gu, Haoming Cai, Chao Dong, Jimmy S Ren, Radu Timofte, Yuan Gong, Shanshan Lao, Shuwei Shi, Jiahao Wang, Sidi Yang, et al · 2022
Earlier work this paper cites.
A full-reference quality assessment metric for cartoon images
Chunyi Li, Zicheng Zhang, Wei Sun, Xiongkuo Min, and Guangtao Zhai · 2022
Earlier work this paper cites.
Hierarchical text-conditional image generation with clip latents, 2022
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen · 2022
Earlier work this paper cites.
Text-guided synthesis of artistic images with retrieval-augmented diffusion models, 2022
Robin Rombach, Andreas Blattmann, and Björn Ommer · 2022
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Earlier work this paper cites.
Attentions help cnns see better: Attention-based hybrid image quality assessment network, 2022
Shanshan Lao, Yuan Gong, Shuwei Shi, Sidi Yang, Tianhe Wu, Jiahao Wang, Weihao Xia, and Yujiu Yang · 2022
Cited alongside, same era.
Seed-bench: Benchmarking multimodal llms with generative comprehension, 2023
Bohao Li, Rui Wang, Guangzhi Wang, Yuying Ge, Yixiao Ge, and Ying Shan · 2023
Cited alongside, same era.
Hrs-bench: Holistic, reliable and scalable benchmark for text-to-image models
Eslam Mohamed Bakr, Pengzhan Sun, Xiaogian Shen, Faizan Farooq Khan, Li Erran Li, and Mohamed Elhoseiny · 2023
Cited alongside, same era.
T2i-compbench: A comprehensive benchmark for open-world compositional text-to-image generation
Kaiyi Huang, Kaiyue Sun, Enze Xie, Zhenguo Li, and Xihui Liu · 2023
Cited alongside, same era.
Text + sketch: Image compression at ultra low rates, 2023
Eric Lei, Yiğit Berkay Uslu, Hamed Hassani, and Shirin Saeedi Bidokhti · 2023
Cited alongside, same era.
Blind image quality assessment via vision-language correspondence: A multitask learning perspective
Weixia Zhang, Guangtao Zhai, Ying Wei, Xiaokang Yang, and Kede Ma · 2023
Later among the works it cites.
Mmbench: Is your multi-modal model an all-around player?, 2024
Yuan Liu, Haodong Duan, Yuanhan Zhang, Bo Li, Songyang Zhang, Wangbo Zhao, Yike Yuan, Jiaqi Wang, Conghui He, Ziwei Liu, Kai Chen, and Dahua Lin · 2024
Closest in time.
Cross modal compression with variable rate prompt
Junlong Gao, Jiguo Li, Chuanmin Jia, Shanshe Wang, Siwei Ma, and Wen Gao · 2024
Closest in time.
Rate-distortion optimized cross modal compression with multiple domains
Junlong Gao, Chuanmin Jia, Zhimeng Huang, Shanshe Wang, Siwei Ma, and Wen Gao · 2024
Closest in time.
Extreme image compression using fine-tuned vqgans
Qi Mao, Tinghan Yang, Yinuo Zhang, Zijian Wang, Meng Wang, Shiqi Wang, Libiao Jin, and Siwei Ma · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Adding conditional control to text-to-image diffusion models
Lvmin Zhang, Anyi Rao, and Maneesh Agrawala · 2023
Cited alongside, same era.
Q-instruct: Improving low-level visual abilities for multi-modality foundation models, 2023
Haoning Wu, Zicheng Zhang, Erli Zhang, Chaofeng Chen, Liang Liao, Annan Wang, Kaixin Xu, Chunyi Li, Jingwen Hou, Guangtao Zhai, et al · 2023
Cited alongside, same era.
Dall-eval: Probing the reasoning skills and social biases of text-to-image generation models
Jaemin Cho, Abhay Zala, and Mohit Bansal · 2023
Cited alongside, same era.
Agiqa-3k: An open database for ai-generated image quality assessment
Chunyi Li, Zicheng Zhang, Haoning Wu, Wei Sun, Xiongkuo Min, Xiaohong Liu, Guangtao Zhai, and Weisi Lin · 2023
Cited alongside, same era.
A perceptual quality assessment exploration for aigc images
Zicheng Zhang, Chunyi Li, Wei Sun, Xiaohong Liu, Xiongkuo Min, and Guangtao Zhai · 2023
Cited alongside, same era.
A real-time blind quality-of-experience assessment metric for http adaptive streaming
Chunyi Li, May Lim, Abdelhak Bentaleb, and Roger Zimmermann · 2023
Cited alongside, same era.
Advancing zero-shot digital human quality assessment through text-prompted evaluation, 2023
Zicheng Zhang, Wei Sun, Yingjie Zhou, Haoning Wu, Chunyi Li, Xiongkuo Min, Xiaohong Liu, Guangtao Zhai, and Weisi Lin · 2023
Cited alongside, same era.
Naifu Xue, Qi Mao, Zijian Wang, Yuan Zhang, and Siwei Ma · 2024
Closest in time.
Once-for-all: Controllable generative image compression with dynamic granularity adaption, 2024
Anqi Li, Yuxi Liu, Huihui Bai, Feng Li, Runmin Cong, Meng Wang, and Yao Zhao · 2024
Closest in time.
Misc: Ultra-low bitrate image semantic compression driven by large multimodal model, 2024
Chunyi Li, Guo Lu, Donghui Feng, Haoning Wu, Zicheng Zhang, Xiaohong Liu, Guangtao Zhai, Weisi Lin, and Wenjun Zhang · 2024
Closest in time.
Q-bench: A benchmark for general-purpose foundation models on low-level vision, 2024
Haoning Wu, Zicheng Zhang, Erli Zhang, Chaofeng Chen, Liang Liao, Annan Wang, Chunyi Li, Wenxiu Sun, Qiong Yan, Guangtao Zhai, and Weisi Lin · 2024
Closest in time.
Fakebench: Uncover the achilles’ heels of fake images with large multimodal models, 2024
Yixuan Li, Xuelin Liu, Xiaoyang Wang, Shiqi Wang, and Weisi Lin · 2024
Closest in time.
A-bench: Are lmms masters at evaluating ai-generated images?, 2024
Zicheng Zhang, Haoning Wu, Chunyi Li, Yingjie Zhou, Wei Sun, Xiongkuo Min, Zijian Chen, Xiaohong Liu, Weisi Lin, and Guangtao Zhai · 2024
Closest in time.
Imagereward: Learning and evaluating human preferences for text-to-image generation
Jiazheng Xu, Xiao Liu, Yuchen Wu, Yuxuan Tong, Qinkai Li, Ming Ding, Jie Tang, and Yuxiao Dong · 2024
Closest in time.
A reduced-reference quality assessment metric for textured mesh digital humans
Zicheng Zhang, Yingjie Zhou, Chunyi Li, Kang Fu, Wei Sun, Xiaohong Liu, Xiongkuo Min, and Guangtao Zhai · 2024
Closest in time.
Quality-of-experience evaluation for digital twins in 6g network environments
Zicheng Zhang, Yingjie Zhou, Long Teng, Wei Sun, Chunyi Li, Xiongkuo Min, Xiao-Ping Zhang, and Guangtao Zhai · 2024
Closest in time.
Paps-ovqa: Projection-aware patch sampling for omnidirectional video quality assessment
Chunyi Li, Zicheng Zhang, Haoning Wu, Kaiwei Zhang, Lei Bai, Xiaohong Liu, Guangtao Zhai, and Weisi Lin · 2024
Closest in time.
Lmm-pcqa: Assisting point cloud quality assessment with lmm, 2024
Zicheng Zhang, Haoning Wu, Yingjie Zhou, Chunyi Li, Wei Sun, Chaofeng Chen, Xiongkuo Min, Xiaohong Liu, Weisi Lin, and Guangtao Zhai · 2024
Closest in time.
Playground v2.5: Three insights towards enhancing aesthetic quality in text-to-image generation, 2024
Daiqing Li, Aleks Kamko, Ehsan Akhgari, Ali Sabet, Linmiao Xu, and Suhail Doshi · 2024
Closest in time.
Progressive knowledge distillation of stable diffusion xl using layer level loss, 2024
Yatharth Gupta, Vishnu V. Jaddipal, Harish Prabhala, Sayak Paul, and Patrick Von Platen · 2024
Closest in time.
Aigiqa-20k: A large database for ai-generated image quality assessment, 2024
Chunyi Li, Tengchuan Kou, Yixuan Gao, Yuqin Cao, Wei Sun, Zicheng Zhang, Yingjie Zhou, Zhichao Zhang, Weixia Zhang, Haoning Wu, Xiaohong Liu, Xiongkuo Min, and Guangtao Zhai · 2024
Closest in time.
NTIRE 2024 quality assessment of AI-generated content challenge
Xiaohong Liu, Xiongkuo Min, Guangtao Zhai, Chunyi Li, Tengchuan Kou, Wei Sun, Haoning Wu, Yixuan Gao, Yuqin Cao, Zicheng Zhang, Xiele Wu, Radu Timofte, et al · 2024
Closest in time.
Subjective-aligned dataset and metric for text-to-video quality assessment, 2024
Tengchuan Kou, Xiaohong Liu, Zicheng Zhang, Chunyi Li, Haoning Wu, Xiongkuo Min, Guangtao Zhai, and Ning Liu · 2024
Closest in time.
G-refine: A general quality refiner for text-to-image generation, 2024
Chunyi Li, Haoning Wu, Hongkun Hao, Zicheng Zhang, Tengchaun Kou, Chaofeng Chen, Lei Bai, Xiaohong Liu, Weisi Lin, and Guangtao Zhai · 2024
Closest in time.
Q-refine: A perceptual quality refiner for ai-generated image, 2024
Chunyi Li, Haoning Wu, Zicheng Zhang, Hongkun Hao, Kaiwei Zhang, Lei Bai, Xiaohong Liu, Xiongkuo Min, Weisi Lin, and Guangtao Zhai · 2024
Closest in time.
Realvisxl-v4.0
Civital · 2024
Closest in time.
Animatediff: Animate your personalized text-to-image diffusion models without specific tuning
Yuwei Guo, Ceyuan Yang, Anyi Rao, Zhengyang Liang, Yaohui Wang, Yu Qiao, Maneesh Agrawala, Dahua Lin, and Bo Dai · 2024
Closest in time.
Diffbir: Towards blind image restoration with generative diffusion prior, 2024
Xinqi Lin, Jingwen He, Ziyan Chen, Zhaoyang Lyu, Bo Dai, Fanghua Yu, Wanli Ouyang, Yu Qiao, and Chao Dong · 2024
Closest in time.
Pixel-aware stable diffusion for realistic image super-resolution and personalized stylization, 2024
Tao Yang, Rongyuan Wu, Peiran Ren, Xuansong Xie, and Lei Zhang · 2024
Closest in time.
Topiq: A top-down approach from semantics to distortions for image quality assessment
Chaofeng Chen, Jiadi Mo, Jingwen Hou, Haoning Wu, Liang Liao, Wenxiu Sun, Qiong Yan, and Weisi Lin · 2024
Closest in time.