Fetching the paper…
Reading the bibliography…
Today, the dominant paradigm for training neural networks involves minimizing task loss on a large dataset.
Dropout: A simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 1958
Earlier work this paper cites.
Horn clause queries and generalizations
Ashok K Chandra and David Harel. 1985 · 1985
Earlier work this paper cites.
Refinement of approximate domain theories by knowledge-based neural networks
Geoffrey G Towell, Jude W Shavlik, and Michiel O Noordewier. 1990 · 1990
Earlier work this paper cites.
A comparison of the computational power of sigmoid and boolean threshold circuits
Wolfgang Maass, Georg Schnitger, and Eduardo D Sontag. 1994 · 1994
Earlier work this paper cites.
Introduction to the CoNLL-2000 shared task: Chunking
Erik F Tjong Kim Sang and Sabine Buchholz. 2000 · 2000
Earlier work this paper cites.
Boolean functions and artificial neural networks
Martin Anthony. 2003 · 2003
Earlier work this paper cites.
ConceptNet – A Practical Commonsense Reasoning Tool-Kit
Hugo Liu and Push Singh. 2004 · 2004
Earlier work this paper cites.
Markov logic networks
Matthew Richardson and Pedro Domingos. 2006 · 2006
Earlier work this paper cites.
Linguistic structure prediction
Noah A Smith. 2011 · 2011
Earlier work this paper cites.
Structured learning with constrained conditional models
Ming-Wei Chang, Lev Ratinov, and Dan Roth. 2012 · 2012
Earlier work this paper cites.
A short introduction to probabilistic soft logic
Angelika Kimmig, Stephen Bach, Matthias Broecheler, Bert Huang, and Lise Getoor. 2012 · 2012
Earlier work this paper cites.
Building high-level features using large scale unsupervised learning
Quoc V Le, Marc’Aurelio Ranzato, Rajat Monga, Matthieu Devin, Kai Chen, Greg S Corrado, Jeff Dean, and Andrew Y Ng. 2012 · 2012
Earlier work this paper cites.
Triangular norms
Erich Peter Klement, Radko Mesiar, and Endre Pap. 2013 · 2013
Earlier work this paper cites.
Fast relational learning using bottom clause propositionalization with artificial neural networks
Manoel VM França, Gerson Zaverucha, and Artur S d’Avila Garcez. 2014 · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
Dropout: A simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2015 · 2015
Cited alongside, same era.
A large annotated corpus for learning natural language inference
Samuel R Bowman, Gabor Angeli, Christopher Potts, and Christopher D Manning. 2015 · 2015
Cited alongside, same era.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean. 2015 · 2015
Cited alongside, same era.
Bidirectional lstm-crf models for sequence tagging
Zhiheng Huang, Wei Xu, and Kai Yu. 2015 · 2015
Cited alongside, same era.
Effective approaches to attention-based neural machine translation
Thang Luong, Hieu Pham, and Christopher D Manning. 2015 · 2015
Cited alongside, same era.
Injecting logical background knowledge into embeddings for relation extraction
Learning to compose words into sentences with reinforcement learning
Dani Yogatama, Phil Blunsom, Chris Dyer, Edward Grefenstette, and Wang Ling. 2017 · 2017
Later among the works it cites.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer. 2017 · 2017
Later among the works it cites.
Neural natural language inference models enhanced with external knowledge
Qian Chen, Xiaodan Zhu, Zhen-Hua Ling, Diana Inkpen, and Si Wei. 2018 · 2018
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Later among the works it cites.
A unified model for extractive and abstractive summarization using inconsistency loss
Wan-Ting Hsu, Chieh-Kai Lin, Ming-Ying Lee, Kerui Min, Jing Tang, and Min Sun. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Tim Rocktäschel, Sameer Singh, and Sebastian Riedel. 2015 · 2015
Cited alongside, same era.
A neural attention model for abstractive sentence summarization
Alexander M Rush, Sumit Chopra, and Jason Weston. 2015 · 2015
Cited alongside, same era.
Harnessing deep neural networks with logic rules
Zhiting Hu, Xuezhe Ma, Zhengzhong Liu, Eduard Hovy, and Eric Xing. 2016 · 2016
Cited alongside, same era.
Expressiveness of rectifier networks
Xingyuan Pan and Vivek Srikumar. 2016 · 2016
Cited alongside, same era.
A decomposable attention model for natural language inference
Ankur Parikh, Oscar Täckström, Dipanjan Das, and Jakob Uszkoreit. 2016 · 2016
Cited alongside, same era.
Squad: 100,000+ questions for machine comprehension of text
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Cited alongside, same era.
Artificial Intelligence: A Modern Approach
Stuart J Russell and Peter Norvig. 2016 · 2016
Cited alongside, same era.
Towards semi-supervised learning for deep semantic role labeling
Sanket Vaibhav Mehta, Jay Yoon Lee, and Jaime Carbonell. 2018 · 2018
Later among the works it cites.
Adversarially regularising neural nli models to integrate logical background knowledge
Pasquale Minervini and Sebastian Riedel. 2018 · 2018
Later among the works it cites.
SparseMAP: Differentiable sparse structured inference
Vlad Niculae, André FT Martins, Mathieu Blondel, and Claire Cardie. 2018 · 2018
Later among the works it cites.
Backpropagating through Structured Argmax using a SPIGOT
Hao Peng, Sam Thomson, and Noah A Smith. 2018 · 2018
Later among the works it cites.
Deep contextualized word representations
Matthew Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Later among the works it cites.
Deep probabilistic logic: A unifying framework for indirect supervision
Hai Wang and Hoifung Poon. 2018 · 2018
Later among the works it cites.
A semantic loss function for deep learning with symbolic knowledge
Jingyi Xu, Zilu Zhang, Tal Friedman, Yitao Liang, and Guy Van den Broeck. 2018 · 2018
Later among the works it cites.
Deep contextualized word representations
Matthew Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Later among the works it cites.
Deeplogic: End-to-end logical reasoning
Nuri Cingillioglu and Alessandra Russo. 2019 · 2019
Closest in time.
Be consistent! improving procedural text comprehension using label consistency
Xinya Du, Bhavana Dalvi, Niket Tandon, Antoine Bosselut, Wen tau Yih, Peter Clark, and Claire Cardie. 2019 · 2019
Closest in time.
Dl2: Training and querying neural networks with logic
Marc Fischer, Mislav Balunovic, Dana Drachsler-Cohen, Timon Gehr, Ce Zhang, and Martin Vechev. 2019 · 2019
Closest in time.