Fetching the paper…
Reading the bibliography…
The success of deep learning has sparked interest in improving relational table tasks, like data preparation and search, with table representation models trained on large table corpora.
DBpedia: A nucleus for a web of open data
Sören Auer, Christian Bizer, Georgi Kobilarov, Jens Lehmann, Richard Cyganiak, and Zachary Ives. 2007 · 2007
Earlier work this paper cites.
Manyeyes: a site for visualization at internet scale
Fernanda B Viegas, Martin Wattenberg, Frank Van Ham, Jesse Kriss, and Matt McKeon. 2007 · 2007
Earlier work this paper cites.
WebTables: exploring the power of tables on the web
Michael J Cafarella, Alon Halevy, Daisy Zhe Wang, Eugene Wu, and Yang Zhang. 2008a · 2008
Earlier work this paper cites.
WebTables: Exploring the Power of Tables on the Web
Michael J. Cafarella, Alon Halevy, Daisy Zhe Wang, Eugene Wu, and Yang Zhang. 2008b · 2008
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database. In 2009 IEEE conference on computer vision and pattern recognition . Ieee, 248–255
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009 · 2009
Earlier work this paper cites.
About WordNet
Princeton University. 2010 · 2010
Earlier work this paper cites.
pandas: a foundational Python library for data analysis and statistics
Wes McKinney et al · 2011
Earlier work this paper cites.
Web Data Commons - extracting structured data from two large web corpora. In LDOW
Hannes Mühleisen and Christian Bizer. 2012 · 2012
Earlier work this paper cites.
WDC Web Table Corpus 2012
WebDataCommons. 2021 · 2012
Earlier work this paper cites.
Methods for exploring and mining tables on wikipedia. In Proceedings of the ACM SIGKDD workshop on interactive data exploration and analytics . 18–26
Chandra Sekhar Bhagavatula, Thanapon Noraset, and Doug Downey. 2013 · 2013
Earlier work this paper cites.
Building the dresden web table corpus: A classification approach. In 2015 IEEE/ACM 2nd International Symposium on Big Data Computing (BDC) . IEEE, 41–50
Julian Eberius, Katrin Braunschweig, Markus Hentsch, Maik Thiele, Ahmad Ahmadov, and Wolfgang Lehner. 2015 · 2015
Earlier work this paper cites.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton. 2015 · 2015
Earlier work this paper cites.
The CTU prague relational learning repository
Jan Motl and Oliver Schulte. 2015 · 2015
Earlier work this paper cites.
Schema. org: evolution of structured data on the web
Ramanathan V Guha, Dan Brickley, and Steve Macbeth. 2016 · 2016
Earlier work this paper cites.
A Large Public Corpus of Web Tables Containing Time and Context Metadata. In WWW Companion . 75–76
Oliver Lehmberg, Dominique Ritze, Robert Meusel, and Christian Bizer. 2016 · 2016
Earlier work this paper cites.
Characteristics of open data CSV files. In 2016 2nd International Conference on Open and Big Data (OBD) . IEEE, 72–79
Johann Mitlöhner, Sebastian Neumaier, Jürgen Umbrich, and Axel Polleres. 2016 · 2016
Earlier work this paper cites.
Automated quality assessment of metadata across open data portals
Sebastian Neumaier, Jürgen Umbrich, and Axel Polleres. 2016 · 2016
Cited alongside, same era.
Democratic databases: science on GitHub
Jeffrey Perkel. 2016 · 2016
Cited alongside, same era.
CSV on the web: A primer
Jeni Tenneson. 2016 · 2016
Cited alongside, same era.
Database meets deep learning: Challenges and opportunities
Wei Wang, Meihui Zhang, Gang Chen, HV Jagadish, Beng Chin Ooi, and Kian-Lee Tan. 2016 · 2016
Cited alongside, same era.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov. 2017 · 2017
Cited alongside, same era.
Discovering enterprise concepts using spreadsheet tables. In Proceedings of the 23rd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining . 1873–1882
Keqian Li, Yeye He, and Kris Ganjam. 2017 · 2017
TURL: Table Understanding through Representation Learning
Xiang Deng, Huan Sun, Alyssa Lees, You Wu, and Cong Yu. 2020 · 2020
Later among the works it cites.
Results of SemTab 2020. In CEUR Workshop Proceedings , Vol. 2775. 1–8
Ernesto Jimenez-Ruiz, Oktie Hassanzadeh, Vasilis Efthymiou, Jiaoyan Chen, Kavitha Srinivas, and Vincenzo Cutrona. 2020 · 2020
Later among the works it cites.
Dataset Reuse: Toward Translating Principles to Practice
Laura Koesten, Pavlos Vougiouklis, Elena Simperl, and Paul Groth. 2020 · 2020
Later among the works it cites.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Later among the works it cites.
REVISE: A tool for measuring and mitigating bias in visual datasets. In European Conference on Computer Vision . Springer, 733–751
Angelina Wang, Arvind Narayanan, and Olga Russakovsky. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Matching web tables to DBpedia – a feature utility study
Dominique Ritze and Christian Bizer. 2017 · 2017
Cited alongside, same era.
Ten years of webtables
Michael Cafarella, Alon Halevy, Hongrae Lee, Jayant Madhavan, Cong Yu, Daisy Zhe Wang, and Eugene Wu. 2018 · 2018
Cited alongside, same era.
Plotly Community Feed
Plotly. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 . 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
VizNet: Towards a large-scale visualization learning and benchmarking repository. In CHI . ACM
Kevin Hu, Neil Gaikwad, Michiel Bakker, Madelon Hulsebos, Emanuel Zgraggen, César Hidalgo, Tim Kraska, Guoliang Li, Arvind Satyanarayan, and Çağatay Demiralp. 2019 · 2019
Cited alongside, same era.
Sherlock: A deep learning approach to semantic data type detection. In Proceedings of the 25th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . 1500–1508
Madelon Hulsebos, Kevin Hu, Michiel Bakker, Emanuel Zgraggen, Arvind Satyanarayan, Tim Kraska, Çagatay Demiralp, and César Hidalgo. 2019 · 2019
Cited alongside, same era.
TaBERT: Pretraining for Joint Understanding of Textual and Tabular Data. In ACL
Pengcheng Yin, Graham Neubig, Wen-tau Yih, and Sebastian Riedel. 2020 · 2020
Later among the works it cites.
Web table extraction, retrieval, and augmentation: A survey
Shuo Zhang and Krisztian Balog. 2020 · 2020
Later among the works it cites.
On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?. In Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency . 610–623
Emily M Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Closest in time.
Large image datasets: A pyrrhic win for computer vision?. In 2021 IEEE Winter Conference on Applications of Computer Vision (WACV) . IEEE, 1536–1546
Abeba Birhane and Vinay Uday Prabhu. 2021 · 2021
Closest in time.
Results of SemTab 2021. In Proceedings of the Semantic Web Challenge on Tabular Data to Knowledge Graph Matching co-located with the 20th International Semantic Web Conference (ISWC 2021), Virtual conference, October 27, 2021 (CEUR Workshop Proceedings, Vol. 3103) . CEUR-WS.org, 1–12
Vincenzo Cutrona, Jiaoyan Chen, Vasilis Efthymiou, Oktie Hassanzadeh, Ernesto Jiménez-Ruiz, Juan Sequeda, Kavitha Srinivas, Nora Abdelmageed, Madelon Hulsebos, Daniela Oliveira, and Catia Pesquita. 2021 · 2021
Closest in time.
Towards Learned Metadata Extraction for Data Lakes
Sven Langenecker, Christoph Sturm, Christian Schalles, and Carsten Binnig. 2021 · 2021
Closest in time.
T2Dv2 Gold Standard for Matching Web Tables to DBpedia
Dominique Ritze, Oliver Lehmberg, and Christian Bizer. 2021 · 2021
Closest in time.
TCN: Table Convolutional Network for Web Table Interpretation
Daheng Wang, Prashant Shiralkar, Colin Lockard, Binxuan Huang, Xin Luna Dong, and Meng Jiang. 2021 · 2021
Closest in time.
Knowledge graphs 2021: a data odyssey
Gerhard Weikum. 2021 · 2021
Closest in time.
Universal Sentence Encoder for English. In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing: System Demonstrations . Association for Computational Linguistics, Brussels, Belgium, 169–174
Daniel Cer, Yinfei Yang, Sheng-yi Kong, Nan Hua, Nicole Limtiaco, Rhomni St. John, Noah Constant, Mario Guajardo-Cespedes, Steve Yuan, Chris Tar, Brian Strope, and Ray Kurzweil. 2018 · 2029
Closest in time.