Fetching the paper…
Reading the bibliography…
Detecting the semantic types of data columns in relational tables is important for various data preparation and information retrieval tasks such as data cleaning, schema matching, data discovery, and semantic search.
Error bounds for convolutional codes and an asymptotically optimum decoding algorithm
A. Viterbi · 1967
Earlier work this paper cites.
Markov random field image models and their applications to computer vision
S. Geman and C. Graffigne · 1986
Earlier work this paper cites.
Learning representations by back-propagating errors
D. E. Rumelhart, G. E. Hinton, and R. J. Williams · 1986
Earlier work this paper cites.
A tutorial on hidden markov models and selected applications in speech recognition
L. R. Rabiner · 1989
Earlier work this paper cites.
Semantic integration in heterogeneous databases using neural networks
W.-S. Li and C. Clifton · 1994
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Conditional random fields: Probabilistic models for segmenting and labeling sequence data
J. D. Lafferty, A. McCallum, and F. C. N. Pereira · 2001
Earlier work this paper cites.
A survey of approaches to automatic schema matching
E. Rahm and P. A. Bernstein · 2001
Earlier work this paper cites.
Potter’s wheel: An interactive data cleaning system
V. Raman and J. M. Hellerstein · 2001
Earlier work this paper cites.
Modeling annotated data
D. M. Blei and M. I. Jordan · 2003
Earlier work this paper cites.
Latent dirichlet allocation
D. M. Blei, A. Y. Ng, and M. I. Jordan · 2003
Earlier work this paper cites.
A bayesian hierarchical model for learning natural scene categories
L. Fei-Fei and P. Perona · 2005
Earlier work this paper cites.
DBpedia: a nucleus for a web of open data
S. Auer, C. Bizer, G. Kobilarov, J. Lehmann, R. Cyganiak, and Z. Ives · 2007
Earlier work this paper cites.
Predicting structured data
G. Bakır, T. Hofmann, B. Schölkopf, A. J. Smola, and B. Taskar · 2007
Earlier work this paper cites.
Freebase: A collaboratively created graph database for structuring human knowledge
K. Bollacker, C. Evans, P. Paritosh, T. Sturge, and J. Taylor · 2008
Earlier work this paper cites.
Webtables: Exploring the power of tables on the web
M. J. Cafarella, A. Halevy, D. Z. Wang, E. Wu, and Y. Zhang · 2008
Earlier work this paper cites.
Visualizing data using t-SNE
L. van der Maaten and G. Hinton · 2008
Earlier work this paper cites.
The unreasonable effectiveness of data
A. Halevy, P. Norvig, and F. Pereira · 2009
Cited alongside, same era.
Probabilistic graphical models: principles and techniques
D. Koller and N. Friedman · 2009
Cited alongside, same era.
Sparse higher order conditional random fields for improved sequence labeling
X. Qian, X. Jiang, Q. Zhang, X. Huang, and L. Wu · 2009
Cited alongside, same era.
Permutation importance: a corrected feature importance measure
A. Altmann, L. Toloşi, O. Sander, and T. Lengauer · 2010
Cited alongside, same era.
Annotating and searching web tables using entities, types and relationships
G. Limaye, S. Sarawagi, and S. Chakrabarti · 2010
Cited alongside, same era.
Software Framework for Topic Modelling with Large Corpora
R. Řehůřek and P. Sojka · 2010
Cited alongside, same era.
Semantic labeling: a domain-independent approach
M. Pham, S. Alse, C. A. Knoblock, and P. Szekely · 2016
Later among the works it cites.
Learning the structure of variable-order CRFs: a finite-state perspective
T. Lavergne and F. Yvon · 2017
Later among the works it cites.
Automatic differentiation in pytorch
A. Paszke, S. Gross, S. Chintala, G. Chanan, E. Yang, Z. DeVito, Z. Lin, A. Desmaison, L. Antiga, and A. Lerer · 2017
Later among the works it cites.
Aurum: A data discovery system
R. Castro Fernandez, Z. Abedjan, F. Koko, G. Yuan, S. Madden, and M. Stonebraker · 2018
Later among the works it cites.
Seeping semantics: Linking datasets using word embeddings for data discovery
R. Castro Fernandez, E. Mansour, A. Qahtan, A. Elmagarmid, I. Ilyas, S. Madden, M. Ouzzani, M. Stonebraker, and N. Tang · 2018
Later among the works it cites.
Synthesizing type-detection logic for rich semantic data types using open-source code
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Exploiting a web of semantic data for interpreting tables
Z. Syed, T. Finin, V. Mulwad, A. Joshi, et al · 2010
Cited alongside, same era.
Wrangler: Interactive visual specification of data transformation scripts
S. Kandel, A. Paepcke, J. Hellerstein, and J. Heer · 2011
Cited alongside, same era.
Recovering semantics of tables on the web
P. Venetis, A. Halevy, J. Madhavan, M. Paşca, W. Shen, F. Wu, G. Miao, and C. Wu · 2011
Cited alongside, same era.
Probabilistic topic models
D. M. Blei · 2012
Cited alongside, same era.
Exploiting structure within data for accurate labeling using conditional random fields
A. Goel, C. A. Knoblock, and K. Lerman · 2012
Cited alongside, same era.
A specialist approach for classification of column data
N. W. Puranik · 2012
Cited alongside, same era.
C. Yan and Y. He · 2018
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2019
Closest in time.
Messytables ⋅ \cdot pypi, 2019
O. K. Foundation · 2019
Closest in time.
Google Data Studio, 2019
Google · 2019
Closest in time.
Viznet: Towards a large-scale visualization learning and benchmarking repository
K. Hu, N. Gaikwad, M. Bakker, M. Hulsebos, E. Zgraggen, C. Hidalgo, T. Kraska, G. Li, A. Satyanarayan, and Ç. Demiralp · 2019
Closest in time.
Sherlock: A deep learning approach to semantic data type detection
M. Hulsebos, K. Z. Hu, M. A. Bakker, E. Zgraggen, A. Satyanarayan, T. Kraska, Ç. Demiralp, and C. A. Hidalgo · 2019
Closest in time.
Fine-tune BERT for extractive summarization
Y. Liu · 2019
Closest in time.
RoBERTa: A robustly optimized BERT pretraining approach
Y. Liu, M. Ott, N. Goyal, J. Du, M. Joshi, D. Chen, O. Levy, M. Lewis, L. Zettlemoyer, and V. Stoyanov · 2019
Closest in time.
Power BI — Interactive Data Visualization BI, 2019
Microsoft · 2019
Closest in time.
Tableau Desktop, 2019
Tableau · 2019
Closest in time.
Meimei: An efficient probabilistic approach for semantically annotating tables
K. Takeoka, M. Oyamada, S. Nakadai, and T. Okadome · 2019
Closest in time.
Data Wrangling Tools & Software, 2019
Trifacta · 2019
Closest in time.