Fetching the paper…
Reading the bibliography…
We present a novel approach for disentangling the content of a text image from all aspects of its appearance.
M. Hood, “Pen, ink, & evidence: A study of writing and writing materials for the penman, collector, and document detective,” 1992
1992
Earlier work this paper cites.
C. Mediavilla, Calligraphy: from calligraphy to abstract painting . Scirpus Publ., 1996
1996
Earlier work this paper cites.
U.-V. Marti and H. Bunke, “The IAM-database: an english sentence database for offline handwriting recognition,” Int. J. Doc. Anal. Recog. , vol. 5, no. 1, pp. 39–46, 2002
2002
Earlier work this paper cites.
P. Pérez, M. Gangnet, and A. Blake, “Poisson image editing,” in ACM Trans. Graph. , 2003, pp. 313–318
2003
Earlier work this paper cites.
R. Bringhurst, The elements of typographic style . Hartley & Marks Vancouver, 2004
2004
Earlier work this paper cites.
T. Samara, Typography workbook: A real-world guide to using type in graphic design . Rockport Publishers, 2004
2004
Earlier work this paper cites.
Z. Wang, A. C. Bovik, H. R. Sheikh, and E. P. Simoncelli, “Image quality assessment: from error visibility to structural similarity,” IEEE Trans. Image Process. , vol. 13, no. 4, pp. 600–612, 2004
2004
Earlier work this paper cites.
T. M. Rath and R. Manmatha, “Word spotting for historical documents,” Int. J. Doc. Anal. Recog. , vol. 9, no. 2-4, pp. 139–152, 2007
2007
Earlier work this paper cites.
K. Wang, B. Babenko, and S. Belongie, “End-to-end scene text recognition,” in Int. Conf. Comput. Vis. , 2011, pp. 1457–1464
2011
Earlier work this paper cites.
L. Neumann and J. Matas, “Real-time scene text localization and recognition,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2012, pp. 3538–3545
2012
Earlier work this paper cites.
T. Hassner, M. Rehbein, P. A. Stokes, and L. Wolf, “Computation and palaeography: potentials and limits,” Dagstuhl Reports , vol. 2, no. 9, pp. 184–199, 2012
2012
Earlier work this paper cites.
F. Kleber, S. Fiel, M. Diem, and R. Sablatnig, “CVL-database: An off-line database for writer retrieval, writer identification and word spotting,” in Int. Conf. Document Analysis and Recog. , 2013, pp. 560–564
2013
Earlier work this paper cites.
D. Karatzas, F. Shafait, S. Uchida, M. Iwamura, L. G. i Bigorda, S. R. Mestre, J. Mas, D. F. Mota, J. Almazán, and L. de las Heras, “ICDAR 2013 robust reading competition,” in Int. Conf. Document Analysis and Recog. , 2013, pp. 1484–1493
2013
Earlier work this paper cites.
T. Hassner, R. Sablatnig, D. Stutzmann, and S. Tarte, “Digital palaeography: New machines and old texts (dagstuhl seminar 14302),” Dagstuhl Reports , vol. 4, no. 7, 2014
2014
Earlier work this paper cites.
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in Adv. Neural Inform. Process. Syst. , 2014, pp. 2672––2680
2014
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-encoding variational bayes,” in Int. Conf. Learn. Represent. , 2014
2014
Earlier work this paper cites.
J. A. Sánchez, V. Romero, A. H. Toselli, and E. Vidal, “ICFHR2014 competition on handwritten text recognition on transcriptorium datasets (HTRtS),” in Int. Conf. Front. Hand. Recog. , 2014, pp. 785–790
2014
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,” in Int. Conf. Learn. Represent. , Y. Bengio and Y. LeCun, Eds., 2015
2015
Earlier work this paper cites.
D. Karatzas, L. Gomez-Bigorda, A. Nicolaou, S. K. Ghosh, A. D. Bagdanov, M. Iwamura, J. Matas, L. Neumann, V. R. Chandrasekhar, S. Lu, F. Shafait, S. Uchida, and E. Valveny, “ICDAR 2015 competition on robust reading,” in Int. Conf. Document Analysis and Recog. , 2015, pp. 1156–1160
2015
Earlier work this paper cites.
L. A. Gatys, A. S. Ecker, and M. Bethge, “Image style transfer using convolutional neural networks,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2016, pp. 2414–2423
2016
Earlier work this paper cites.
J. Johnson, A. Alahi, and L. Fei-Fei, “Perceptual losses for real-time style transfer and super-resolution,” in Eur. Conf. Comput. Vis. , 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2016, pp. 770–778
2016
Cited alongside, same era.
A. Gupta, A. Vedaldi, and A. Zisserman, “Synthetic data for text localisation in natural images,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2016, pp. 2315–2324
2016
Cited alongside, same era.
P. Isola, J.-Y. Zhu, T. Zhou, and A. A. Efros, “Image-to-image translation with conditional adversarial networks,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2017
2017
Cited alongside, same era.
J.-Y. Zhu, T. Park, P. Isola, and A. A. Efros, “Unpaired image-to-image translation using cycle-consistent adversarial networks,” in Int. Conf. Comput. Vis. , 2017
2017
Cited alongside, same era.
X. Huang and S. Belongie, “Arbitrary style transfer in real-time with adaptive instance normalization,” in Int. Conf. Comput. Vis. , 2017
R. Gomez, A. F. Biten, L. Gómez, J. Gibert, D. Karatzas, and M. Rusiñol, “Selective style transfer for text,” in Int. Conf. Document Analysis and Recog. , 2019, pp. 805–812
2019
Later among the works it cites.
E. Alonso, B. Moysset, and R. O. Messina, “Adversarial generation of handwritten text images conditioned on sequences,” in Int. Conf. Document Analysis and Recog. , 2019, pp. 481–486
2019
Later among the works it cites.
J. Baek, G. Kim, J. Lee, S. Park, D. Han, S. Yun, S. J. Oh, and H. Lee, “What is wrong with scene text recognition model comparisons? dataset and model analysis,” in Int. Conf. Comput. Vis. , 2019, pp. 4714–4722
2019
Later among the works it cites.
A. Singh, V. Natarjan, M. Shah, Y. Jiang, X. Chen, D. Parikh, and M. Rohrbach, “Towards VQA models that can read,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2019, pp. 8317–8326
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
S. Yang, J. Liu, Z. Lian, and Z. Guo, “Awesome typography: Statistics-based text effects transfer,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2017, pp. 7464–7473
2017
Cited alongside, same era.
K. He, G. Gkioxari, P. Dollár, and R. Girshick, “Mask R-CNN,” in Int. Conf. Comput. Vis. , 2017, pp. 2961–2969
2017
Cited alongside, same era.
Z. Cheng, F. Bai, Y. Xu, G. Zheng, S. Pu, and S. Zhou, “Focusing attention: Towards accurate text recognition in natural images,” in Int. Conf. Comput. Vis. , 2017, pp. 5086–5094
2017
Cited alongside, same era.
I. Krasin, T. Duerig, N. Alldrin, V. Ferrari, S. Abu-El-Haija, A. Kuznetsova, H. Rom, J. Uijlings, S. Popov, S. Kamali, M. Malloci, J. Pont-Tuset, A. Veit, S. Belongie, V. Gomes, A. Gupta, C. Sun, G. Chechik, D. Cai, Z. Feng, D. Narayanan, and K. Murphy, “OpenImages: A public dataset for large-scale multi-label and multi-class image classification.” Dataset available from https://storage.googleapis.com/openimages/web/index.html , 2017
2017
Cited alongside, same era.
M. Heusel, H. Ramsauer, T. Unterthiner, B. Nessler, and S. Hochreiter, “GANs trained by a two time-scale update rule converge to a local nash equilibrium,” in Adv. Neural Inform. Process. Syst. , 2017, pp. 6626–6637
2017
Cited alongside, same era.
I. Masi, T. Hassner, A. T. Tran, and G. Medioni, “Rapid synthesis of massive face sets for improved face recognition,” in Int. Conf. on Automatic Face and Gesture Recognition , 2017, pp. 604–611
2017
Cited alongside, same era.
S. Azadi, M. Fisher, V. G. Kim, Z. Wang, E. Shechtman, and T. Darrell, “Multi-content GAN for few-shot font style transfer,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2018, pp. 7564–7573
2018
Cited alongside, same era.
I. Masi, A. T. Tran, T. Hassner, G. Sahin, and G. Medioni, “Face-specific data augmentation for unconstrained face recognition,” Int. J. Comput. Vis. , vol. 127, no. 6-7, pp. 642–667, 2019
2019
Later among the works it cites.
A. Rossler, D. Cozzolino, L. Verdoliva, C. Riess, J. Thies, and M. Nießner, “Faceforensics++: Learning to detect manipulated facial images,” in Int. Conf. Comput. Vis. , 2019, pp. 1–11
2019
Later among the works it cites.
2019
Later among the works it cites.
Q. Yang, J. Huang, and W. Lin, “SwapText: Image based texts transfer in scenes,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2020, pp. 14 688–14 697
2020
Later among the works it cites.
P. Roy, S. Bhattacharya, S. Ghosh, and U. Pal, “STEFANN: scene text editor using font adaptive neural network,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2020, pp. 13 225–13 234
2020
Later among the works it cites.
S. Fogel, H. Averbuch-Elor, S. Cohen, S. Mazor, and R. Litman, “Scrabblegan: Semi-supervised varying length handwritten text generation,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2020, pp. 4323–4332
2020
Later among the works it cites.
2020
Later among the works it cites.
Y. Choi, Y. Uh, J. Yoo, and J.-W. Ha, “StarGAN v2: Diverse image synthesis for multiple domains,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2020, pp. 8188–8197
2020
Later among the works it cites.
L. Kang, P. Riba, Y. Wang, M. Rusiñol, A. Fornés, and M. Villegas, “GANwriting: Content-conditioned generation of styled handwritten word images,” in Eur. Conf. Comput. Vis. , 2020, pp. 273–289
2020
Later among the works it cites.
T. Karras, S. Laine, M. Aittala, J. Hellsten, J. Lehtinen, and T. Aila, “Analyzing and improving the image quality of StyleGAN,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2020, pp. 8107–8116
2020
Later among the works it cites.
W. Li, Y. He, Y. Qi, Z. Li, and Y. Tang, “FET-GAN: Font and effect transfer via k-shot adaptive instance normalization,” in AAAI , vol. 34, no. 02, 2020, pp. 1717–1724
2020
Later among the works it cites.
M. Liao, G. Pang, J. Huang, T. Hassner, and X. Bai, “Mask TextSpotter v3: Segmentation proposal network for robust scene text spotting,” in Eur. Conf. Comput. Vis. , 2020
2020
Later among the works it cites.
M. Liao, B. Song, S. Long, M. He, C. Yao, and X. Bai, “SynthText3D: synthesizing scene text images from 3D virtual worlds,” Science China Information Sciences , vol. 63, no. 2, pp. 1–14, 2020
2020
Later among the works it cites.
S. Long, X. He, and C. Yao, “Scene text detection and recognition: The deep learning era,” Int. J. Comput. Vis. , vol. 129, no. 1, pp. 161–184, 2021
2021
Closest in time.
A. Singh, G. Pang, M. Toh, J. Huang, T. Hassner, and W. Galuba, “TextOCR: Towards large-scale end-to-end reasoning for arbitrary-shaped scene text,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2021
2021
Closest in time.
J. Huang, G. Pang, R. Kovvuri, M. Toh, K. J. Liang, P. Krishnan, X. Yin, and T. Hassner, “A multiplexed network for end-to-end, multilingual OCR,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2021
2021
Closest in time.