Fetching the paper…
Reading the bibliography…
Scientific knowledge is predominantly stored in books and scientific journals, often in the form of PDFs.
Cycle-Consistency for Robust Visual Question Answering, February 2019
Meet Shah, Xinlei Chen, Marcus Rohrbach, and Devi Parikh · 1902
Earlier work this paper cites.
The Curious Case of Neural Text Degeneration, February 2020
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi · 1904
Earlier work this paper cites.
Zelun Wang and Jyh-Charn Liu · 1908
Earlier work this paper cites.
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Ves Stoyanov, and Luke Zettlemoyer · 1910
Earlier work this paper cites.
LayoutLM: Pre-training of Text and Layout for Document Image Understanding
Yiheng Xu, Minghao Li, Lei Cui, Shaohan Huang, Furu Wei, and Ming Zhou · 1912
Earlier work this paper cites.
Calculus
Herman W. (Herman William) March and Henry C. (Henry Charles) Wolff · 1917
Earlier work this paper cites.
Distributional Structure
Zellig S. Harris · 1954
Earlier work this paper cites.
Binary codes capable of correcting deletions, insertions, and reversals
V. Levenshtein · 1965
Earlier work this paper cites.
URL https://ntrs.nasa.gov/citations/19700022795
Kinetics and Thermodynamics in High-Temperature Gases, January 1970 · 1970
Earlier work this paper cites.
Bleu: a Method for Automatic Evaluation of Machine Translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu · 2002
Earlier work this paper cites.
Best practices for convolutional neural networks applied to visual document analysis
P.Y. Simard, D. Steinkraus, and J.C. Platt · 2003
Earlier work this paper cites.
METEOR: An Automatic Metric for MT Evaluation with Improved Correlation with Human Judgments
Satanjeev Banerjee and Alon Lavie · 2005
Earlier work this paper cites.
An Overview of the Tesseract OCR Engine
R. Smith · 2007
Earlier work this paper cites.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale, June 2021
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby · 2010
Earlier work this paper cites.
GROBID, February 2023
Patrice Lopez · 2012
Earlier work this paper cites.
Recognition of on-line handwritten mathematical expressions using 2D stochastic context-free grammars and hidden Markov models
Francisco Álvaro, Joan-Andreu Sánchez, and José-Miguel Benedí · 2012
Earlier work this paper cites.
ConvMath: A Convolutional Sequence Network for Mathematical Expression Recognition, December 2020
Zuoyu Yan, Xiaode Zhang, Liangcai Gao, Ke Yuan, and Zhi Tang · 2012
Cited alongside, same era.
LayoutLMv2: Multi-modal Pre-training for Visually-Rich Document Understanding, January 2022
Yang Xu, Yiheng Xu, Tengchao Lv, Lei Cui, Furu Wei, Guoxin Wang, Yijuan Lu, Dinei Florencio, Cha Zhang, Wanxiang Che, Min Zhang, and Lidong Zhou · 2012
Cited alongside, same era.
Statistics of the Common Crawl Corpus 2012, June 2013
Sebastian Spiegler · 2013
Cited alongside, same era.
A new approach for recognizing handwritten mathematics using relational grammars and fuzzy sets
Scott MacLean and George Labahn · 2013
Cited alongside, same era.
Online publishing via pdf2htmlEX, 2013
Lu Wang and Wanmin Liu · 2013
Cited alongside, same era.
S2ORC: The Semantic Scholar Open Research Corpus
Kyle Lo, Lucy Lu Wang, Mark Neumann, Rodney Kinney, and Daniel Weld · 2020
Later among the works it cites.
pix2tex - LaTeX OCR, February 2023
Lukas Blecher · 2020
Later among the works it cites.
Representation Learning for Information Extraction from Form-like Documents
Bodhisattwa Prasad Majumder, Navneet Potti, Sandeep Tata, James Bradley Wendt, Qi Zhao, and Marc Najork · 2020
Later among the works it cites.
Rethinking Text Line Recognition Models, April 2021
Daniel Hernandez Diaz, Siyang Qin, Reeve Ingle, Yasuhisa Fujii, and Alessandro Bissacco · 2021
Later among the works it cites.
Handwritten Mathematical Expression Recognition with Bidirectionally Trained Transformer, May 2021
Wenqi Zhao, Liangcai Gao, Zuoyu Yan, Shuai Peng, Lin Du, and Ziyin Zhang · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A global learning approach for an online handwritten mathematical expression recognition system
Ahmad-Montaser Awal, Harold Mouchre, and Christian Viard-Gaudin · 2014
Cited alongside, same era.
Image-to-Markup Generation with Coarse-to-Fine Attention, September 2016
Yuntian Deng, Anssi Kanervisto, Jeffrey Ling, and Alexander M. Rush · 2016
Cited alongside, same era.
PDFFigures 2.0: Mining Figures from Research Papers
Christopher Clark and Santosh Divvala · 2016
Cited alongside, same era.
Full-Page Text Recognition: Learning Where to Start and When to Stop, April 2017
Bastien Moysset, Christopher Kermorvant, and Christian Wolf · 2017
Cited alongside, same era.
Training an End-to-End System for Handwritten Mathematical Expression Recognition by Generated Patterns
Anh Duc Le and Masaki Nakagawa · 2017
Cited alongside, same era.
Attention Is All You Need, December 2017
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Teaching Machines to Code: Neural Markup Generation with Visual Attention, June 2018
Sumeet S. Singh · 2018
Cited alongside, same era.
Srikar Appalaraju, Bhavan Jasani, Bhargava Urala Kota, Yusheng Xie, and R. Manmatha · 2021
Later among the works it cites.
Swin Transformer: Hierarchical Vision Transformer using Shifted Windows, August 2021
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo · 2021
Later among the works it cites.
Scene Text Recognition with Permuted Autoregressive Sequence Models, July 2022
Darwin Bautista and Rowel Atienza · 2022
Later among the works it cites.
TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models, September 2022
Minghao Li, Tengchao Lv, Jingye Chen, Lei Cui, Yijuan Lu, Dinei Florencio, Cha Zhang, Zhoujun Li, and Furu Wei · 2022
Later among the works it cites.
LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking, July 2022
Yupan Huang, Tengchao Lv, Lei Cui, Yutong Lu, and Furu Wei · 2022
Later among the works it cites.
OCR-free Document Understanding Transformer, October 2022
Geewook Kim, Teakgyu Hong, Moonbin Yim, Jeongyeon Nam, Jinyoung Park, Jinyeong Yim, Wonseok Hwang, Sangdoo Yun, Dongyoon Han, and Seunghyun Park · 2022
Later among the works it cites.
End-to-end Document Recognition and Understanding with Dessurt, June 2022
Brian Davis, Bryan Morse, Bryan Price, Chris Tensmeyer, Curtis Wigington, and Vlad Morariu · 2022
Later among the works it cites.
Galactica: A Large Language Model for Science, November 2022
Ross Taylor, Marcin Kardas, Guillem Cucurull, Thomas Scialom, Anthony Hartshorn, Elvis Saravia, Andrew Poulton, Viktor Kerkez, and Robert Stojnic · 2022
Later among the works it cites.
OCR-IDL: OCR Annotations for Industry Document Library Dataset, February 2022
Ali Furkan Biten, Rubèn Tito, Lluis Gomez, Ernest Valveny, and Dimosthenis Karatzas · 2022
Later among the works it cites.
Beyond neural scaling laws: beating power law scaling via data pruning, November 2022
Ben Sorscher, Robert Geirhos, Shashank Shekhar, Surya Ganguli, and Ari S. Morcos · 2022
Later among the works it cites.
Albumentations: Fast and Flexible Image Augmentations
Alexander Buslaev, Vladimir I. Iglovikov, Eugene Khvedchenya, Alex Parinov, Mikhail Druzhinin, and Alexandr A. Kalinin · 2078
Closest in time.