Fetching the paper…
Reading the bibliography…
Researchers use figures to communicate rich, complex information in scientific papers.
Figure captioning with reasoning and sequence-level training
Charles Chen, Ruiyi Zhang, Eunyee Koh, Sungchul Kim, Scott Cohen, Tong Yu, Ryan Rossi, and Razvan Bunescu. 2019b · 1906
Earlier work this paper cites.
Nltk: The natural language toolkit
Edward Loper and Steven Bird. 2002 · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Scanssd: Scanning single shot detector for mathematical formulas in pdf document images
Parag Mali, Puneeth Kukkadapu, Mahshad Mahdavi, and Richard Zanibbi. 2020 · 2003
Earlier work this paper cites.
Protograph-based low-density parity-check hadamard codes
Peng W Zhang, Francis Lau, and Chiu-W Sham. 2020 · 2010
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Supporting the design and fabrication of physical visualizations
Saiganesh Swaminathan, Conglei Shi, Yvonne Jansen, Pierre Dragicevic, Lora A Oehlberg, and Jean-Daniel Fekete. 2014 · 2014
Earlier work this paper cites.
How transferable are features in deep neural networks?
Jason Yosinski, Jeff Clune, Yoshua Bengio, and Hod Lipson. 2014 · 2014
Earlier work this paper cites.
Building proteins in a day: Efficient 3d molecular reconstruction
Marcus A Brubaker, Ali Punjani, and David J Fleet. 2015 · 2015
Earlier work this paper cites.
Effective approaches to attention-based neural machine translation
Minh-Thang Luong, Hieu Pham, and Christopher D Manning. 2015 · 2015
Earlier work this paper cites.
Pdffigures 2.0: Mining figures from research papers
Christopher Clark and Santosh Divvala. 2016 · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Figureseer: Parsing result-figures in research papers
Noah Siegel, Zachary Horvitz, Roie Levin, Santosh Divvala, and Ali Farhadi. 2016 · 2016
Cited alongside, same era.
Linespace: A sensemaking platform for the blind
Saiganesh Swaminathan, Thijs Roumen, Robert Kovacs, David Stangl, Stefanie Mueller, and Patrick Baudisch. 2016 · 2016
Cited alongside, same era.
Incremental dfs algorithms: a theoretical and experimental study
Surender Baswana, Ayush Goel, and Shahbaz Khan. 2017 · 2017
Cited alongside, same era.
Figureqa: An annotated figure dataset for visual reasoning
Samira Ebrahimi Kahou, Vincent Michalski, Adam Atkinson, Ákos Kádár, Adam Trischler, and Yoshua Bengio. 2017 · 2017
Cited alongside, same era.
Neural caption generation over figures
Charles Chen, Ruiyi Zhang, Sungchul Kim, Scott Cohen, Tong Yu, Ryan Rossi, and Razvan Bunescu. 2019a · 2019
Later among the works it cites.
On the use of arxiv as a dataset
Colin B. Clement, Matthew Bierbaum, Kevin P. O’Keeffe, and Alexander A. Alemi. 2019 · 2019
Later among the works it cites.
Figure captioning with relation maps for reasoning
Charles Chen, Ruiyi Zhang, Eunyee Koh, Sungchul Kim, Scott Cohen, and Ryan Rossi. 2020 · 2020
Later among the works it cites.
Mmm: Multi-stage multi-task learning for multi-choice reading comprehension
Di Jin, Shuyang Gao, Jiun-Yu Kao, Tagyoung Chung, and Dilek Hakkani-tur. 2020 · 2020
Later among the works it cites.
Finding old answers to new math questions: the arqmath lab at clef 2020
Behrooz Mansouri, Anurag Agarwal, Douglas Oard, and Richard Zanibbi. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Toward scalable social alt text: Conversational crowdsourcing as a tool for refining vision-to-language technology for the blind
Elliot Salisbury, Ece Kamar, and Meredith Ringel Morris. 2017 · 2017
Cited alongside, same era.
A data driven approach for compound figure separation using convolutional neural networks
Satoshi Tsutsui and David J Crandall. 2017 · 2017
Cited alongside, same era.
Automatic alt-text: Computer-generated image descriptions for blind users on a social network service
Shaomei Wu, Jeffrey Wieland, Omid Farivar, and Julie Schiller. 2017 · 2017
Cited alongside, same era.
Dvqa: Understanding data visualizations via question answering
Kushal Kafle, Brian Price, Scott Cohen, and Christopher Kanan. 2018 · 2018
Cited alongside, same era.
Scibert: Pretrained language model for scientific text
Iz Beltagy, Kyle Lo, and Arman Cohan. 2019 · 2019
Cited alongside, same era.
μ \mu graph: Haptic exploration and editing of 3d chemical diagrams
Cristian Bernareggi, Dragan Ahmetovic, and Sergio Mascetti. 2019 · 2019
Cited alongside, same era.
Chart-to-text: Generating natural language descriptions for charts by adapting the transformer model
Jason Obeid and Enamul Hoque. 2020 · 2020
Later among the works it cites.
A formative study on designing accurate and natural figure captioning systems
Xin Qian, Eunyee Koh, Fan Du, Sungchul Kim, and Joel Chan. 2020 · 2020
Later among the works it cites.
Neural data-driven captioning of time-series line charts
Andrea Spreafico and Giuseppe Carenini. 2020 · 2020
Later among the works it cites.
Quantifying bias in automatic speech recognition
Siyuan Feng, Olya Kudina, Bence Mark Halpern, and Odette Scharenborg. 2021 · 2021
Closest in time.
Generating accurate caption units for figure captioning
Xin Qian, Eunyee Koh, Fan Du, Sungchul Kim, Joel Chan, Ryan A Rossi, Sana Malik, and Tak Yeon Lee. 2021 · 2021
Closest in time.
Show, attend and tell: Neural image caption generation with visual attention
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio. 2015 · 2057
Closest in time.