Fetching the paper…
Reading the bibliography…
Complex textual information extraction tasks are often posed as sequence labeling or \emph{shallow parsing}, where fields are extracted using local labels made consistent through probabilistic inference in a graphical model with constrained transitions.
Acceleration of stochastic approximation by averaging
Boris T Polyak and Anatoli B Juditsky. 1992 · 1992
Earlier work this paper cites.
Latent-dynamic discriminative models for continuous gesture recognition
Louis-Philippe Morency, Ariadna Quattoni, and Trevor Darrell. 2007 · 2007
Earlier work this paper cites.
Hidden conditional random fields
Ariadna Quattoni, Sybor Wang, Louis-Philippe Morency, Morency Collins, and Trevor Darrell. 2007 · 2007
Earlier work this paper cites.
Dynamic conditional random fields: Factorized probabilistic models for labeling and segmenting sequence data
Charles Sutton, Andrew McCallum, and Khashayar Rohanimanesh. 2007 · 2007
Earlier work this paper cites.
Visualizing data using t-sne
Laurens van der Maaten and Geoffrey Hinton. 2008 · 2008
Earlier work this paper cites.
Learning structural svms with latent variables
Chun-Nam John Yu and Thorsten Joachims. 2009 · 2009
Earlier work this paper cites.
Fast high-dimensional filtering using the permutohedral lattice
Andrew Adams, Jongmin Baek, and Myers Abraham Davis. 2010 · 2010
Earlier work this paper cites.
Dual decomposition for parsing with non-projective head automata
Terry Koo, Alexander M Rush, Michael Collins, Tommi Jaakkola, and David Sontag. 2010 · 2010
Earlier work this paper cites.
Natural language processing (almost) from scratch
Ronan Collobert, Jason Weston, Léon Bottou, Michael Karlen, Koray Kavukcuoglu, and Pavel Kuksa. 2011 · 2011
Earlier work this paper cites.
Efficient inference in fully connected crfs with gaussian edge potentials
Philipp Krähenbühl and Vladlen Koltun. 2011 · 2011
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
Geoffrey Hinton, Li Deng, Dong Yu, George E Dahl, Abdel-rahman Mohamed, Navdeep Jaitly, Andrew Senior, Vincent Vanhoucke, Patrick Nguyen, Tara N Sainath, et al. 2012 · 2012
Cited alongside, same era.
A tutorial on dual decomposition and lagrangian relaxation for inference in natural language processing
Alexander M Rush and MJ Collins. 2012 · 2012
Cited alongside, same era.
A new dataset for fine-grained citation field extraction
Sam Anzaroot and Andrew McCallum. 2013 · 2013
Cited alongside, same era.
Learning soft linear constraints with application to citation field extraction
Sam Anzaroot, Alexandre Passos, David Belanger, and Andrew McCallum. 2014 · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Cited alongside, same era.
Bethe projections for non-local inference
Luke Vilnis, David Belanger, Daniel Sheldon, and Andrew McCallum. 2015 · 2015
Later among the works it cites.
Structured prediction energy networks
David Belanger and Andrew McCallum. 2016 · 2016
Later among the works it cites.
Inside-outside and forward-backward algorithms are just backprop (tutorial paper)
Jason Eisner. 2016 · 2016
Later among the works it cites.
Neural architectures for named entity recognition
Guillaume Lample, Miguel Ballesteros, Sandeep Subramanian, Kazuya Kawakami, and Chris Dyer. 2016 · 2016
Later among the works it cites.
End-to-end learning for structured prediction energy networks
David Belanger, Bishan Yang, and Andrew McCallum. 2017 · 2017
Later among the works it cites.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer. 2017 · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning distributed representations for structured output prediction
Vivek Srikumar and Christopher D Manning. 2014 · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2015 · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2015 · 2015
Cited alongside, same era.
Not all contexts are created equal: Better word representations with variable attention
Wang Ling, Yulia Tsvetkov, Silvio Amir, Ramon Fermandez, Chris Dyer, Alan W Black, Isabel Trancoso, and Chu-Cheng Lin. 2015 · 2015
Cited alongside, same era.
Benchmarking clinical speech recognition and information extraction: new data, methods, and evaluations
Hanna Suominen, Liyuan Zhou, Leif Hanlen, and Gabriela Ferraro. 2015 · 2015
Cited alongside, same era.
Later among the works it cites.
Yellowfin and the art of momentum tuning
Jian Zhang, Ioannis Mitliagkas, and Christopher Ré. 2017 · 2017
Later among the works it cites.
Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs
Liang-Chieh Chen, George Papandreou, Iasonas Kokkinos, Kevin Murphy, and Alan L Yuille. 2018 · 2018
Closest in time.
Framewise phoneme classification with bidirectional lstm networks
Alex Graves and Jürgen Schmidhuber. 2005 · 2052
Closest in time.