Fetching the paper…
Reading the bibliography…
Real-world documents may suffer various forms of degradation, often resulting in lower accuracy in optical character recognition (OCR) systems.
N. Otsu, “A threshold selection method from gray-level histograms,” IEEE Transactions on Systems, Man, and Cybernetics , vol. 9, no. 1, pp. 62–66, 1979
1979
Earlier work this paper cites.
J. Sauvola and M. Pietikäinen, “Adaptive document image binarization,” Pattern Recognition , vol. 33, no. 2, pp. 225–236, 2000. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S0031320399000552
2000
Earlier work this paper cites.
G. Panci, P. Campisi, S. Colonnese, and G. Scarano, “Multichannel blind image deconvolution using the bussgang algorithm: spatial and multiresolution approaches,” IEEE Transactions on Image Processing , vol. 12, no. 11, pp. 1324–1337, 2003
2003
Earlier work this paper cites.
A. Horé and D. Ziou, “Image quality metrics: Psnr vs. ssim,” in 2010 20th International Conference on Pattern Recognition , 2010, pp. 2366–2369
2010
Earlier work this paper cites.
F. Deng, Z. Wu, Z. Lu, and M. Brown, “Binarizationshop: A user-assisted software suite for converting old documents to black-and-white,” in JCDL’10 - Digital Libraries - 10 Years Past, 10 Years Forward, a 2020 Vision , ser. Proceedings of the ACM International Conference on Digital Libraries, 2010, pp. 255–258, copyright: Copyright 2010 Elsevier B.V., All rights reserved.; 10th Annual Joint Conference on Digital Libraries, JCDL 2010 ; Conference date: 21-06-2010 Through 25-06-2010
2010
Earlier work this paper cites.
H. Cho, J. Wang, and S. Lee, “Text image deblurring using text-specific properties,” in Computer Vision – ECCV 2012 , vol. 7576, 10 2012, pp. 524–537
2012
Earlier work this paper cites.
T. Lelore and F. Bouchara, “Fair: A fast algorithm for document image restoration,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 35, no. 8, pp. 2039–2048, 2013
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
H. Nafchi, S. Ayatollahi, R. Farrahi Moghaddam, and M. Cheriet, “An efficient ground truthing tool for binarization of historical manuscripts,” 08 2013, pp. 807–811
2013
Earlier work this paper cites.
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. C. Courville, and Y. Bengio, “Generative adversarial nets,” in Neural Information Processing Systems , 2014. [Online]. Available: https://api.semanticscholar.org/CorpusID:261560300
2014
Earlier work this paper cites.
J. Pan, Z. Hu, Z. Su, and M.-H. Yang, “Deblurring text images via l0-regularized intensity and gradient prior,” in 2014 IEEE Conference on Computer Vision and Pattern Recognition , 2014, pp. 2901–2908
2014
Earlier work this paper cites.
M. Hradiš, J. Kotera, P. Zemčík, and F. Šroubek, “Convolutional neural networks for direct text deblurring,” in Proceedings of BMVC 2015 . The British Machine Vision Association and Society for Pattern Recognition, 2015. [Online]. Available: http://www.fit.vutbr.cz/research/view_pub.php?id=10922
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
B. Shi, X. Bai, and C. Yao, “An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 39, pp. 2298–2304, 2015. [Online]. Available: https://api.semanticscholar.org/CorpusID:24139
2015
Earlier work this paper cites.
R. Hedjam, H. Z. Nafchi, R. F. Moghaddam, M. Kalacska, and M. Cheriet, “Icdar 2015 contest on multispectral text extraction (ms-tex 2015),” in 2015 13th International Conference on Document Analysis and Recognition (ICDAR) , 2015, pp. 1181–1185
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
J.-C. Burie, M. Coustaty, S. Hadi, M. W. A. Kesiman, J.-M. Ogier, E. Paulus, K. Sok, I. M. G. Sunarya, and D. Valy, “Icfhr2016 competition on the analysis of handwritten text in images of balinese palm leaf manuscripts,” 2016 15th International Conference on Frontiers in Handwriting Recognition (ICFHR) , pp. 596–601, 2016. [Online]. Available: https://api.semanticscholar.org/CorpusID:1165182
2016
Earlier work this paper cites.
I. Pratikakis, K. Zagoris, G. Barlas, and B. Gatos, “Icdar2017 competition on document image binarization (dibco 2017),” in 2017 14th IAPR International Conference on Document Analysis and Recognition (ICDAR) , vol. 01, 2017, pp. 1395–1403
2017
Earlier work this paper cites.
P. Isola, J.-Y. Zhu, T. Zhou, and A. A. Efros, “Image-to-image translation with conditional adversarial networks,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 1125–1134
2017
Earlier work this paper cites.
I. Pratikakis, K. Zagori, P. Kaddas, and B. Gatos, “Icfhr 2018 competition on handwritten document image binarization (h-dibco 2018),” in 2018 16th International Conference on Frontiers in Handwriting Recognition (ICFHR) , 2018, pp. 489–493
2018
Earlier work this paper cites.
T. Karras, T. Aila, S. Laine, and J. Lehtinen, “Progressive growing of GANs for improved quality, stability, and variation,” in International Conference on Learning Representations , 2018. [Online]. Available: https://openreview.net/forum?id=Hk99zCeAb
2018
Cited alongside, same era.
O. Kupyn, V. Budzan, M. Mykhailych, D. Mishkin, and J. Matas, “Deblurgan: Blind motion deblurring using conditional adversarial networks,” in 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . Los Alamitos, CA, USA: IEEE Computer Society, jun 2018, pp. 8183–8192. [Online]. Available: https://doi.ieeecomputersociety.org/10.1109/CVPR.2018.00854
2018
Cited alongside, same era.
Y. Blau and T. Michaeli, “The perception-distortion tradeoff,” in 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2018, pp. 6228–6237
2018
Cited alongside, same era.
R. Zhang, P. Isola, A. A. Efros, E. Shechtman, and O. Wang, “The unreasonable effectiveness of deep features as a perceptual metric,” 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition , pp. 586–595, 2018. [Online]. Available: https://api.semanticscholar.org/CorpusID:4766599
2022
Later among the works it cites.
C. Lu, Y. Zhou, F. Bao, J. Chen, C. Li, and J. Zhu, “DPM-solver: A fast ODE solver for diffusion probabilistic model sampling in around 10 steps,” in Advances in Neural Information Processing Systems , A. H. Oh, A. Agarwal, D. Belgrave, and K. Cho, Eds., 2022. [Online]. Available: https://openreview.net/forum?id=2uAaGwlP_V
2022
Later among the works it cites.
S. Gonwirat and O. Surinta, “Deblurgan-cnn: Effective image denoising and recognition for noisy handwritten characters,” IEEE Access , vol. 10, pp. 90 133–90 148, 2022
2022
Later among the works it cites.
H. Li, Y. Yang, M. Chang, S. Chen, H. Feng, Z. Xu, Q. Li, and Y. Chen, “Srdiff: Single image super-resolution with diffusion probabilistic models,” Neurocomputing , vol. 479, pp. 47–59, 2022. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S0925231222000522
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
A. Sulaiman, K. Omar, and M. F. Nasrudin, “Degraded historical document binarization: A review on issues, challenges, techniques, and future directions,” Journal of Imaging , vol. 5, no. 4, 2019. [Online]. Available: https://www.mdpi.com/2313-433X/5/4/48
2019
Cited alongside, same era.
S. Lin, Z. He, and L. Sun, “Defect enhancement generative adversarial network for enlarging data set of microcrack defect,” IEEE Access , vol. 7, pp. 148 413–148 423, 2019
2019
Cited alongside, same era.
I. Pratikakis, K. Zagoris, X. Karagiannis, L. Tsochatzidis, T. Mondal, and I. Marthot-Santaniello, “Icdar 2019 competition on document image binarization (dibco 2019),” in 2019 International Conference on Document Analysis and Recognition (ICDAR) , 2019, pp. 1547–1556
2019
Cited alongside, same era.
P. V. Bezmaternykh, D. A. Ilin, and D. P. Nikolaev, “U-Net-bin: hacking the document image binarization contest,” Computer Optics , vol. 43, no. 5, pp. 825–832, Oct. 2019
2019
Cited alongside, same era.
M. Fritsche, S. Gu, and R. Timofte, “Frequency separation for real-world super-resolution,” 2019 IEEE/CVF International Conference on Computer Vision Workshop (ICCVW) , pp. 3599–3608, 2019. [Online]. Available: https://api.semanticscholar.org/CorpusID:208158302
2019
Cited alongside, same era.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” in Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M. Balcan, and H. Lin, Eds., vol. 33. Curran Associates, Inc., 2020, pp. 6840–6851. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/2020/file/4c5bcfec8584af0d967f1ab10179ca4b-Paper.pdf
2020
Cited alongside, same era.
K. Ding, K. Ma, S. Wang, and E. P. Simoncelli, “Image quality assessment: Unifying structure and texture similarity,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 44, pp. 2567–2581, 2020. [Online]. Available: https://api.semanticscholar.org/CorpusID:215785896
2020
Cited alongside, same era.
2021
Cited alongside, same era.
2022
Later among the works it cites.
J. Whang, M. Delbracio, H. Talebi, C. Saharia, A. G. Dimakis, and P. Milanfar, “Deblurring via stochastic refinement,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2022, pp. 16 293–16 303
2022
Later among the works it cites.
T. Salimans and J. Ho, “Progressive distillation for fast sampling of diffusion models,” in International Conference on Learning Representations , 2022. [Online]. Available: https://openreview.net/forum?id=TIdIXIpzhoI
2022
Later among the works it cites.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” in 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . Los Alamitos, CA, USA: IEEE Computer Society, jun 2022, pp. 10 674–10 685. [Online]. Available: https://doi.ieeecomputersociety.org/10.1109/CVPR52688.2022.01042
2022
Later among the works it cites.
K.-U. Song, D. Shim, K.-W. Kim, J. young Lee, and Y.-G. Kim, “Fs-ncsr: Increasing diversity of the super-resolution space via frequency separation and noise-conditioned normalizing flow,” 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) , pp. 967–976, 2022. [Online]. Available: https://api.semanticscholar.org/CorpusID:248299786
2022
Later among the works it cites.
S. Khamekhem Jemni, M. A. Souibgui, Y. Kessentini, and A. Fornés, “Enhance to read better: A multi-task adversarial network for handwritten document image enhancement,” Pattern Recognition , vol. 123, p. 108370, 2022. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S0031320321005501
2022
Later among the works it cites.
Z. Yang, B. Liu, Y. Xxiong, L. Yi, G. Wu, X. Tang, Z. Liu, J. Zhou, and X. Zhang, “Docdiff: Document enhancement via residual diffusion models,” in Proceedings of the 31st ACM International Conference on Multimedia , ser. MM ’23. New York, NY, USA: Association for Computing Machinery, 2023, p. 2795–2806. [Online]. Available: https://doi.org/10.1145/3581783.3611730
2023
Later among the works it cites.
C. Saharia, J. Ho, W. Chan, T. Salimans, D. J. Fleet, and M. Norouzi, “Image super-resolution via iterative refinement,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 45, no. 4, pp. 4713–4726, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
Q. Zhang and Y. Chen, “Fast sampling of diffusion models with exponential integrator,” in The Eleventh International Conference on Learning Representations , 2023. [Online]. Available: https://openreview.net/forum?id=Loek7hfb46P
2023
Later among the works it cites.
S. Xue, M. Yi, W. Luo, S. Zhang, J. Sun, Z. Li, and Z.-M. Ma, “Sa-solver: Stochastic adams solver for fast sampling of diffusion models,” 2023
2023
Later among the works it cites.
A. Sauer, D. Lorenz, A. Blattmann, and R. Rombach, “Adversarial diffusion distillation,” 2023
2023
Later among the works it cites.
C. Meng, R. Rombach, R. Gao, D. Kingma, S. Ermon, J. Ho, and T. Salimans, “On distillation of guided diffusion models,” in 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . Los Alamitos, CA, USA: IEEE Computer Society, jun 2023, pp. 14 297–14 306. [Online]. Available: https://doi.ieeecomputersociety.org/10.1109/CVPR52729.2023.01374
2023
Later among the works it cites.
L. Yang, Z. Zhang, Y. Song, S. Hong, R. Xu, Y. Zhao, W. Zhang, B. Cui, and M.-H. Yang, “Diffusion models: A comprehensive survey of methods and applications,” ACM Comput. Surv. , vol. 56, no. 4, nov 2023. [Online]. Available: https://doi.org/10.1145/3626235
2023
Later among the works it cites.
Z. Luo, F. K. Gustafsson, Z. Zhao, J. Sjölund, and T. B. Schön, “Refusion: Enabling large-size realistic image restoration with latent-space diffusion models,” in 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) , 2023, pp. 1680–1691
2023
Later among the works it cites.
M. Yang and S. Xu, “A novel degraded document binarization model through vision transformer network,” Information Fusion , vol. 93, pp. 159–173, 2023. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S1566253522002597
2023
Later among the works it cites.
Y. Zhou, S. Zuo, Z. Yang, J. He, J. Shi, and R. Zhang, “A review of document image enhancement based on document degradation problem,” Applied Sciences , vol. 13, no. 13, 2023. [Online]. Available: https://www.mdpi.com/2076-3417/13/13/7855
2076
Closest in time.