Fetching the paper…
Reading the bibliography…
Diffusion models, known for their impressive image generation abilities, have played a pivotal role in the rise of visual text generation.
ICDAR 2013 Robust Reading Competition
D. Karatzas, F. Shafait, S. Uchida, M. Iwamura, L. G. i Bigorda, S. R. Mestre, et al · 2013
Earlier work this paper cites.
ICDAR 2015 Competition on Robust Reading
D. Karatzas, L. Gomez-Bigorda, A. Nicolaou, S. Ghosh, A. Bagdanov, M. Iwamura, J. Matas, L. Neumann, V. R. Chandrasekhar, S. Lu, et al · 2015
Earlier work this paper cites.
U-Net: Convolutional Networks for Biomedical Image Segmentation
O. Ronneberger, P. Fischer, and T. Brox · 2015
Earlier work this paper cites.
Synthetic Data for Text Localisation in Natural Images
A. Gupta, A. Vedaldi, and A. Zisserman · 2016
Earlier work this paper cites.
ICDAR2017 Robust Reading Challenge on Multi-Lingual Scene Text Detection and Script Identification - RRC-MLT
N. Nayef, F. Yin, et al · 2017
Earlier work this paper cites.
Verisimilar Image Synthesis for Accurate Detection and Recognition of Texts in Scenes
F. Zhan, S. Lu, and C. Xue · 2018
Earlier work this paper cites.
Learning to Draw Text in Natural Images with Conditional Adversarial Networks
S. Fang, H. Xie, J. Chen, J. Tan, and Y. Zhang · 2019
Earlier work this paper cites.
ICDAR2019 Robust Reading Challenge on Multi-lingual Scene Text Detection and Recognition – RRC-MLT-2019
N. Nayef, Y. Patel, M. Busta, et al · 2019
Earlier work this paper cites.
Spatial Fusion GAN for Image Synthesis
F. Zhan, H. Zhu, and S. Lu · 2019
Earlier work this paper cites.
Total-Text: Toward Orientation Robustness in Scene Text Detection
C.-K. Ch’ng, C. S. Chan, and C.-L. Liu · 2020
Earlier work this paper cites.
Denoising Diffusion Probabilistic Models
J. Ho, A. Jain, and P. Abbeel · 2020
Earlier work this paper cites.
EraseNet: End-to-End Text Removal in the Wild
C. Liu, Y. Liu, l. Jin, S. Zhang, C. Luo, and Y. Wang · 2020
Earlier work this paper cites.
UnrealText: Synthesizing Realistic Scene Text Images from the Unreal World
S. Long and C. Yao · 2020
Earlier work this paper cites.
Blindly Assess Image Quality in the Wild Guided by a Self-Adaptive Hyper Network
S. Su, Q. Yan, Y. Zhu, C. Zhang, X. Ge, J. Sun, and Y. Zhang · 2020
Earlier work this paper cites.
Read Like Humans: Autonomous, Bidirectional and Iterative Language Modeling for Scene Text Recognition
S. Fang, H. Xie, Y. Wang, Z. Mao, and Y. Zhang · 2021
Cited alongside, same era.
LayoutTransformer: Layout Generation and Completion with Self-attention
K. Gupta, J. Lazarow, A. Achille, L. S. Davis, V. Mahadevan, and A. Shrivastava · 2021
Cited alongside, same era.
Mask is All You Need: Rethinking Mask R-CNN for Dense and Arbitrary-Shaped Scene Text Detection
X. Qin, Y. Zhou, Y. Guo, D. Wu, Z. Tian, N. Jiang, H. Wang, and W. Wang · 2021
Cited alongside, same era.
PP-OCRv3: More Attempts for the Improvement of Ultra Lightweight OCR System
C. Li, W. Liu, R. Guo, X. Yin, K. Jiang, Y. Du, Y. Du, L. Zhu, B. Lai, X. Hu, et al · 2022
Cited alongside, same era.
Character-Aware Models Improve Visual Text Rendering
R. Liu, D. Garrette, C. Saharia, W. Chan, A. Roberts, S. Narang, I. Blok, R. Mical, M. Norouzi, and N. Constant · 2022
Improving Diffusion Models for Scene Text Editing with Dual Encoders
J. Ji, G. Zhang, Z. Wang, B. Hou, Z. Zhang, B. Price, and S. Chang · 2023
Later among the works it cites.
GlyphDraw: Learning to Draw Chinese Characters in Image Synthesis Models Coherently
J. Ma, M. Zhao, C. Chen, R. Wang, D. Niu, H. Lu, and X. Lin · 2023
Later among the works it cites.
OpenAI: Introducing ChatGPT, 2023
OpenAI · 2023
Later among the works it cites.
Towards Robust Real-Time Scene Text Detection: From Semantic to Instance Representation Learning
X. Qin, P. Lyu, C. Zhang, Y. Zhou, K. Yao, P. Zhang, H. Lin, and W. Wang · 2023
Later among the works it cites.
Perceiving Ambiguity and Semantics without Recognition: An Efficient and Effective Ambiguous Scene Text Detector
Y. Shu, W. Wang, Y. Zhou, S. Liu, et al · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
High-Resolution Image Synthesis with Latent Diffusion Models
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer · 2022
Cited alongside, same era.
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding
C. Saharia, W. Chan, S. Saxena, L. Li, J. Whang, E. L. Denton, K. Ghasemipour, R. Gontijo Lopes, B. Karagol Ayan, T. Salimans, et al · 2022
Cited alongside, same era.
LaMa: Resolution-robust Large Mask Inpainting with Fourier Convolutions
R. Suvorov, E. Logacheva, A. Mashikhin, et al · 2022
Cited alongside, same era.
TPSNet: Reverse Thinking of Thin Plate Splines for Arbitrary Shape Scene Text Representation
W. Wang, Y. Zhou, J. Lv, D. Wu, G. Zhao, N. Jiang, and W. Wang · 2022
Cited alongside, same era.
ByT5: Towards a Token-Free Future with Pre-trained Byte-to-Byte Models
L. Xue, A. Barua, N. Constant, R. Al-Rfou, S. Narang, M. Kale, A. Roberts, and C. Raffel · 2022
Cited alongside, same era.
ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth
S. F. Bhat, R. Birkl, D. Wofk, P. Wonka, and M. Müller · 2023
Cited alongside, same era.
TextDiffuser: Diffusion Models as Text Painters
J. Chen, Y. Huang, T. Lv, L. Cui, Q. Chen, and F. Wei · 2023
Cited alongside, same era.
AnyText: Multilingual Visual Text Generation And Editing
Y. Tuo, W. Xiang, J.-Y. He, Y. Geng, and X. Xie · 2023
Later among the works it cites.
Open-Vocabulary Panoptic Segmentation with Text-to-Image Diffusion Models
J. Xu, S. Liu, A. Vahdat, W. Byeon, X. Wang, and S. De Mello · 2023
Later among the works it cites.
GlyphControl: Glyph Conditional Control for Visual Text Generation
Y. Yang, D. Gui, Y. Yuan, H. Ding, H. Hu, and K. Chen · 2023
Later among the works it cites.
Adding Conditional Control to Text-to-Image Diffusion Models
L. Zhang, A. Rao, and M. Agrawala · 2023
Later among the works it cites.
TextBlockV2: Towards Precise-Detection-Free Scene Text Spotting with Pre-trained Language Model
J. Lyu, J. Wei, G. Zeng, Z. Li, E. Xie, W. Wang, and Y. Zhou · 2024
Closest in time.
Visual Text Meets Low-level Vision: A Comprehensive Survey on Visual Text Processing
Y. Shu, W. Zeng, Z. Li, F. Zhao, and Y. Zhou · 2024
Closest in time.
Masked and Permuted Implicit Context Learning for Scene Text Recognition
X. Yang, Z. Qiao, J. Wei, D. Yang, and Y. Zhou · 2024
Closest in time.
Brush Your Text: Synthesize Any Scene Text on Images via Diffusion Model
L. Zhang, X. Chen, Y. Wang, Y. Lu, and Y. Qiao · 2024
Closest in time.