Fetching the paper…
Reading the bibliography…
Entity resolution (ER) is one of the fundamental problems in data integration, where machine learning (ML) based classifiers often provide the state-of-the-art results.
A theory for record linkage
I. Fellegi and A. Sunter · 1969
Earlier work this paper cites.
Transfer of learning
D. N. Perkins, G. Salomon, et al · 1992
Earlier work this paper cites.
Assessing the accuracy of prediction algorithms for classification: an overview
P. Baldi, S. Brunak, Y. Chauvin, C. A. Andersen, and H. Nielsen · 2000
Earlier work this paper cites.
Logistic regression in rare events data
G. King and L. Zeng · 2001
Earlier work this paper cites.
Learning to match and cluster large high-dimensional data sets for data integration
W. W. Cohen and J. Richman · 2002
Earlier work this paper cites.
Interactive deduplication using active learning
S. Sarawagi and A. Bhamidipaty · 2002
Earlier work this paper cites.
Adaptive duplicate detection using learnable string similarity measures
M. Bilenko and R. J. Mooney · 2003
Earlier work this paper cites.
To transfer or not to transfer
M. T. Rosenstein, Z. Marx, L. P. Kaelbling, and T. G. Dietterich · 2005
Earlier work this paper cites.
Generic entity resolution in the serf project
O. Benjelloun, H. Garcia-molina, H. Kawai, T. E. Larson, D. Menestrina, Q. Su, S. Thavisomboon, and J. Widom · 2006
Earlier work this paper cites.
Domain adaptation with structural correspondence learning
J. Blitzer, R. McDonald, and F. Pereira · 2006
Earlier work this paper cites.
Domain adaptation for statistical classifiers
H. Daume III and D. Marcu · 2006
Earlier work this paper cites.
Entity resolution with markov logic
P. Singla and P. Domingos · 2006
Earlier work this paper cites.
Analysis of representations for domain adaptation
S. Ben-David, J. Blitzer, K. Crammer, and F. Pereira · 2007
Earlier work this paper cites.
Discriminative learning for differing training and test distributions
S. Bickel, M. Brückner, and T. Scheffer · 2007
Earlier work this paper cites.
Frustratingly easy domain adaptation
H. Daumé III · 2007
Earlier work this paper cites.
Duplicate record detection: A survey
A. K. Elmagarmid, P. G. Ipeirotis, and V. S. Verykios · 2007
Earlier work this paper cites.
Correcting sample selection bias by unlabeled data
J. Huang, A. Gretton, K. M. Borgwardt, B. Schölkopf, and A. J. Smola · 2007
Earlier work this paper cites.
Direct importance estimation with model selection and its application to covariate shift adaptation
M. Sugiyama, S. Nakajima, H. Kashima, P. V. Buenau, and M. Kawanabe · 2008
Earlier work this paper cites.
On active learning of record matching packages
A. Arasu, M. Götz, and R. Kaushik · 2010
Cited alongside, same era.
Frustratingly easy semi-supervised domain adaptation
H. Daumé III, A. Kumar, and A. Saha · 2010
Cited alongside, same era.
Evaluation of entity resolution approaches on real-world match problems
H. Köpcke, A. Thor, and E. Rahm · 2010
Cited alongside, same era.
Co-regularization based semi-supervised domain adaptation
A. Kumar, A. Saha, and H. Daume · 2010
Cited alongside, same era.
An introduction to duplicate detection
F. Naumann and M. Herschel · 2010
Cited alongside, same era.
A survey on transfer learning
S. J. Pan, Q. Yang, et al · 2010
Cited alongside, same era.
The million song dataset
Learning with augmented features for supervised and semi-supervised heterogeneous domain adaptation
W. Li, L. Duan, D. Xu, and I. W. Tsang · 2014
Later among the works it cites.
A simple machine learning method to detect covariate shift
F. J. Martin · 2014
Later among the works it cites.
Glove: Global vectors for word representation
J. Pennington, R. Socher, and C. D. Manning · 2014
Later among the works it cites.
Enriching word vectors with subword information
P. Bojanowski, E. Grave, A. Joulin, and T. Mikolov · 2016
Later among the works it cites.
Deep Learning
I. Goodfellow, Y. Bengio, and A. Courville · 2016
Later among the works it cites.
Magellan: Toward building entity matching management systems
P. Konda, S. Das, P. Suganthan GC, A. Doan, A. Ardalan, J. R. Ballard, H. Li, F. Panahi, H. Zhang, J. Naughton, et al · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Bertin-Mahieux, D. P. Ellis, B. Whitman, and P. Lamere · 2011
Cited alongside, same era.
Scikit-learn: Machine learning in python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, et al · 2011
Cited alongside, same era.
A course in machine learning
H. Daumé III · 2012
Cited alongside, same era.
Dedoop: Efficient deduplication with hadoop
L. Kolb, A. Thor, and E. Rahm · 2012
Cited alongside, same era.
Scaling multiple-source entity resolution using statistically efficient transfer learning
S. N. Negahban, B. I. Rubinstein, and J. G. Gemmell · 2012
Cited alongside, same era.
Crowder: Crowdsourcing entity resolution
J. Wang, T. Kraska, M. J. Franklin, and J. Feng · 2012
Cited alongside, same era.
Later among the works it cites.
Data programming: Creating large training sets, quickly
A. J. Ratner, C. M. De Sa, S. Wu, D. Selsam, and C. Ré · 2016
Later among the works it cites.
Towards universal paraphrastic sentence embeddings
J. Wieting, M. Bansal, K. Gimpel, and K. Livescu · 2016
Later among the works it cites.
A simple but tough-to-beat baseline for sentence embeddings
S. Arora, Y. Liang, and T. Ma · 2017
Later among the works it cites.
Representations for language: From word embeddings to sentence meanings
C. Manning · 2017
Later among the works it cites.
Generating concise entity matching rules
R. Singh, V. Meduri, A. K. Elmagarmid, S. Madden, P. Papotti, J. Quiané-Ruiz, A. Solar-Lezama, and N. Tang · 2017
Later among the works it cites.
Dataset shift in machine learning
M. Sugiyama, N. D. Lawrence, A. Schwaighofer, et al · 2017
Later among the works it cites.
Data integration and machine learning: A natural synergy
X. L. Dong and T. Rekatsinas · 2018
Closest in time.
Distributed representations of tuples for entity resolution
M. Ebraheem, S. Thirumuruganathan, S. Joty, M. Ouzzani, and N. Tang · 2018
Closest in time.
A la carte embedding: Cheap but effective induction of semantic feature vectors
M. Khodak, N. Saunshi, Y. Liang, T. Ma, B. Stewart, and S. Arora · 2018
Closest in time.
Deep learning for entity matching: A design space exploration
S. Mudgal, H. Li, T. Rekatsinas, A. Doan, Y. Park, G. Krishnan, R. Deep, E. Arcaute, and V. Raghavendra · 2018
Closest in time.
Deep contextualized word representations
M. Peters, M. Neumann, M. Iyyer, M. Gardner, C. Clark, K. Lee, and L. Zettlemoyer · 2018
Closest in time.
Data integration: The current status and the way forward
M. Stonebraker and I. F. Ilyas · 2018
Closest in time.