Fetching the paper…
Reading the bibliography…
Document images are often degraded by various stains, significantly impacting their readability and hindering downstream applications such as document digitization and analysis.
Historical review of ocr research and development
Shunji Mori, Ching Y Suen, and Kazuhiko Yamamoto · 1992
Earlier work this paper cites.
Adaptive document image binarization
Jaakko Sauvola and Matti Pietikäinen · 2000
Earlier work this paper cites.
Show-through cancellation in scans of duplex printed documents
Gaurav Sharma · 2001
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Zhou Wang, Alan C Bovik, Hamid R Sheikh, and Eero P Simoncelli · 2004
Earlier work this paper cites.
Ocr binarization and image pre-processing for searching historical documents
Maya R Gupta, Nathaniel P Jacobson, and Eric K Garcia · 2007
Earlier work this paper cites.
Low quality document image modeling and enhancement
Reza Farrahi Moghaddam and Mohamed Cheriet · 2009
Earlier work this paper cites.
Constrained energy maximization and self-referencing method for invisible ink detection from multispectral historical document images
Rachid Hedjam, Mohamed Cheriet, and Margaret Kalacska · 2014
Earlier work this paper cites.
Convolutional neural networks for direct text deblurring
Michal Hradis, Jan Kotera, Pavel Zemcík, and Filip Sroubek · 2015
Earlier work this paper cites.
Icdar2015 competition on smartphone document capture and ocr (smartdoc)
Jean-Christophe Burie, Joseph Chazalon, Mickaël Coustaty, Sébastien Eskenazi, Muhammad Muzzamil Luqman, Maroua Mehri, Nibal Nayef, Jean-Marc Ogier, Sophea Prum, and Marçal Rusiñol · 2015
Earlier work this paper cites.
Synthetic data for text localisation in natural images
Ankush Gupta, Andrea Vedaldi, and Andrew Zisserman · 2016
Earlier work this paper cites.
SGDR: stochastic gradient descent with warm restarts
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
Document enhancement using visibility detection
Netanel Kligler, Sagi Katz, and Ayellet Tal · 2018
Earlier work this paper cites.
Degraded historical document image binarization using local features and support vector machine (svm)
Wei Xiong, Jingjing Xu, Zijie Xiong, Juan Wang, and Min Liu · 2018
Earlier work this paper cites.
Docunet: Document image unwarping via a stacked u-net
Ke Ma, Zhixin Shu, Xue Bai, Jue Wang, and Dimitris Samaras · 2018
Earlier work this paper cites.
The unreasonable effectiveness of deep features as a perceptual metric
Richard Zhang, Phillip Isola, Alexei A. Efros, Eli Shechtman, and Oliver Wang · 2018
Earlier work this paper cites.
Deepotsu: Document enhancement and binarization using iterative deep learning
Sheng He and Lambert Schomaker · 2019
Earlier work this paper cites.
Document rectification and illumination correction using a patch-based cnn
Xiaoyu Li, Bo Zhang, Jing Liao, and Pedro V Sander · 2019
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2019
Earlier work this paper cites.
Deep learning for historical document analysis and recognition—a survey
Francesco Lombardi and Simone Marinai · 2020
Earlier work this paper cites.
Geometry-aware cell detection with deep learning
Hao Jiang, Sen Li, Weihuang Liu, Hongjin Zheng, Jinghao Liu, and Yang Zhang · 2020
Earlier work this paper cites.
Fine-grained breast cancer classification with bilinear convolutional neural networks (bcnns)
Weihuang Liu, Mario Juhas, and Yang Zhang · 2020
Earlier work this paper cites.
Sienet: Siamese expansion network for image extrapolation
Xiaofeng Zhang, Feng Chen, Cailing Wang, Ming Tao, and Guo-Ping Jiang · 2020
Earlier work this paper cites.
De-gan: A conditional generative adversarial network for document enhancement
Mohamed Ali Souibgui and Yousri Kessentini · 2020
Earlier work this paper cites.
Deep learning for covid-19 chest ct (computed tomography) image analysis: A lesson from lung cancer
Hao Jiang, Shiming Tang, Weihuang Liu, and Yang Zhang · 2021
Earlier work this paper cites.
Scene text image super-resolution via parallelly contextual attention network
Cairong Zhao, Shuyang Feng, Brian Nlong Zhao, Zhijun Ding, Jun Wu, Fumin Shen, and Heng Tao Shen · 2021
Cited alongside, same era.
End-to-end unsupervised document image blind denoising
Mehrdad J. Gangeh, Marcin Plata, Hamid R. Motahari Nezhad, and Nigel P. Duffy · 2021
Cited alongside, same era.
Doctr: Document image transformer for geometric unwarping and illumination correction
Hao Feng, Yuechen Wang, Wengang Zhou, Jiajun Deng, and Houqiang Li · 2021
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby · 2021
Cited alongside, same era.
Swin transformer: Hierarchical vision transformer using shifted windows
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo · 2021
Cited alongside, same era.
Explicit visual prompting for low-level structure segmentations
Weihuang Liu, Xi Shen, Chi-Man Pun, and Xiaodong Cun · 2023
Later among the works it cites.
High-resolution document shadow removal via a large-scale real-world dataset and a frequency-aware shadow erasing net
Zinuo Li, Xuhang Chen, Chi-Man Pun, and Xiaodong Cun · 2023
Later among the works it cites.
Docdeshadower: Frequency-aware transformer for document shadow removal
Ziyang Zhou, Yingtie Lei, Xuhang Chen, Shenghong Luo, Wenjun Zhang, Chi-Man Pun, and Zhen Wang · 2023
Later among the works it cites.
Deep unrestricted document image rectification
Hao Feng, Shaokai Liu, Jiajun Deng, Wengang Zhou, and Houqiang Li · 2023
Later among the works it cites.
A comparative study of image restoration networks for general backbone network design
Xiangyu Chen, Zheyuan Li, Yuandong Pu, Yihao Liu, Jiantao Zhou, Yu Qiao, and Chao Dong · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Swinir: Image restoration using swin transformer
Jingyun Liang, Jiezhang Cao, Guolei Sun, Kai Zhang, Luc Van Gool, and Radu Timofte · 2021
Cited alongside, same era.
Correction of out-of-focus microscopic images by deep learning
Chi Zhang, Hao Jiang, Weihuang Liu, Junyi Li, Shiming Tang, Mario Juhas, and Yang Zhang · 2022
Cited alongside, same era.
Monocular robust 3d human localization by global and body-parts depth awareness
Haolun Li and Chi-Man Pun · 2022
Cited alongside, same era.
Few-shot object detection via high-and-low resolution representation
Haolun Li, Senlin Ge, Chuyi Gao, and Hao Gao · 2022
Cited alongside, same era.
Docentr: An end-to-end document image enhancement transformer
Mohamed Ali Souibgui, Sanket Biswas, Sana Khamekhem Jemni, Yousri Kessentini, Alicia Fornés, Josep Lladós, and Umapada Pal · 2022
Cited alongside, same era.
Enhance to read better: a multi-task adversarial network for handwritten document image enhancement
Sana Khamekhem Jemni, Mohamed Ali Souibgui, Yousri Kessentini, and Alicia Fornés · 2022
Cited alongside, same era.
Udoc-gan: Unpaired document illumination correction with background light prior
Yonghui Wang, Wengang Zhou, Zhenbo Lu, and Houqiang Li · 2022
Cited alongside, same era.
Xiangyu Chen, Xintao Wang, Jiantao Zhou, Yu Qiao, and Chao Dong · 2023
Later among the works it cites.
Template-guided illumination correction for document images with imperfect geometric reconstruction
Felix Hertlein and Alexander Naumann · 2023
Later among the works it cites.
Adaptive spatial-temporal graph-mixer for human motion prediction
Shubo Yang, Haolun Li, Chi-Man Pun, Chun Du, and Hao Gao · 2024
Closest in time.
Deformmlp: Dynamic large-scale receptive field mlp networks for human motion prediction
Haitao Huang, Chi-Man Pun, Haolun Li, Mengqi Liu, Jian Xiong, and Hao Gao · 2024
Closest in time.
Encoding enhanced complex cnn for accurate and highly accelerated mri
Zimeng Li, Sa Xiao, Cheng Wang, Haidong Li, Xiuchao Zhao, Caohui Duan, Qian Zhou, Qiuchen Rao, Yuan Fang, Junshuai Xie, Lei Shi, Fumin Guo, Chaohui Ye, and Xin Zhou · 2024
Closest in time.
Docres: A generalist model toward unifying document image restoration tasks
Jiaxin Zhang, Dezhi Peng, Chongyu Liu, Peirong Zhang, and Lianwen Jin · 2024
Closest in time.
Test-time intensity consistency adaptation for shadow detection
Leyi Zhu, Weihuang Liu, Xinyi Chen, Zimeng Li, Xuhang Chen, Zhen Wang, and Chi-Man Pun · 2024
Closest in time.
Cross-domain visual prompting with spatial proximity knowledge distillation for histological image classification
Xiaohong Li, Guoheng Huang, Lianglun Cheng, Guo Zhong, Weihuang Liu, Xuhang Chen, and Muyan Cai · 2024
Closest in time.
Muraldiff: Diffusion for ancient murals restoration on large-scale pre-training
Zishan Xu, Xiaofeng Zhang, Wei Chen, Jueting Liu, Tingting Xu, and Zehua Wang · 2024
Closest in time.
Shadclips: When parameter-efficient fine-tuning with multimodal meets shadow removal
Xiaofeng Zhang, Zishan Xu, Hao Tang, Chaochen Gu, Shanying Zhu, and Xinping Guan · 2024
Closest in time.
Hierarchical local temporal network for 2d-to-3d human pose estimation
Xin Yan, Jiucheng Xie, Mengqi Liu, Haolun Li, and Hao Gao · 2024
Closest in time.
Local optimization networks for multi-view multi-person human posture estimation
Jucheng Song, Chi-Man Pun, Haolun Li, Rushi Lan, Jiu-Cheng Xie, and Hao Gao · 2024
Closest in time.
Hierarchical local temporal feature enhancing for transformer-based 3d human pose estimation
Xin Yan, Chi-Man Pun, Haolun Li, Mengqi Liu, and Hao Gao · 2024
Closest in time.
Docnlc: A document image enhancement framework with normalized and latent contrastive representation for multiple degradations
Ruilu Wang, Yang Xue, and Lianwen Jin · 2024
Closest in time.
Depth-aware test-time training for zero-shot video object segmentation
Weihuang Liu, Xi Shen, Haolun Li, Xiuli Bi, Bo Liu, Chi-Man Pun, and Xiaodong Cun · 2024
Closest in time.
Smaformer: Synergistic multi-attention transformer for medical image segmentation
Fuchen Zheng, Xuhang Chen, Weihuang Liu, Haolun Li, Yingtie Lei, Jiahui He, Chi-Man Pun, and Shounjun Zhou · 2024
Closest in time.
Dh-gan: Image manipulation localization via a dual homology-aware generative adversarial network
Weihuang Liu, Xiaodong Cun, and Chi-Man Pun · 2024
Closest in time.
Forgeryttt: Zero-shot image manipulation localization with test-time training
Weihuang Liu, Xi Shen, Chi-Man Pun, and Xiaodong Cun · 2024
Closest in time.
Dual-hybrid attention network for specular highlight removal
Xiaojiao Guo, Xuhang Chen, Shenghong Luo, Shuqiang Wang, and Chi-Man Pun · 2024
Closest in time.