Fetching the paper…
Reading the bibliography…
Tabular data, a prevalent data type across various domains, presents unique challenges due to its heterogeneous nature and complex structural relationships.
Smith, J.W., Everhart, J.E., Dickson, W., Knowler, W.C., Johannes, R.S.: Using the adap learning algorithm to forecast the onset of diabetes mellitus. In: Proceedings of the Annual Symposium on Computer Application in Medical Care, p. 261 (1988). American Medical Informatics Association
1988
Earlier work this paper cites.
Marzal, A., Vidal, E.: Computation of normalized edit distance and applications. IEEE Trans. Pattern Anal. Mach. Intell. 15
1993
Earlier work this paper cites.
Hofmann, H.: Statlog (German Credit Data). UCI Machine Learning Repository. DOI: https://doi.org/10.24432/C5NC77 (1994)
1994
Earlier work this paper cites.
Becker, B., Kohavi, R.: Adult. UCI Machine Learning Repository. DOI: https://doi.org/10.24432/C5XW20 (1996)
1996
Earlier work this paper cites.
Pace, R.K., Barry, R.: Sparse spatial autoregressions. Statistics & Probability Letters 33
1997
Earlier work this paper cites.
Bohanec, M.: Car Evaluation. UCI Machine Learning Repository. DOI: https://doi.org/10.24432/C5JP48 (1997)
1997
Earlier work this paper cites.
Blackard, J.: Covertype. UCI Machine Learning Repository. DOI: https://doi.org/10.24432/C50K5N (1998)
1998
Earlier work this paper cites.
Papineni, K., Roukos, S., Ward, T., Zhu, W.: Bleu: a method for automatic evaluation of machine translation. In: Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics, July 6-12, 2002, Philadelphia, PA, USA, pp. 311–318. ACL, ??? (2002). https://doi.org/10.3115/1073083.1073135 . https://aclanthology.org/P02-1040/
2002
Earlier work this paper cites.
Webber, W., Moffat, A., Zobel, J.: A similarity measure for indefinite rankings. ACM Trans. Inf. Syst. 28
2010
Earlier work this paper cites.
Deng, L.: The mnist database of handwritten digit images for machine learning research [best of the web]. IEEE signal processing magazine 29
2012
Earlier work this paper cites.
Moro, S., Rita, P., , Cortez, P.: Bank Marketing. UCI Machine Learning Repository. DOI: https://doi.org/10.24432/C5K306 (2012)
2012
Earlier work this paper cites.
Mansouri, K., Ringsted, T., Ballabio, D., Todeschini, R., Consonni, V.: QSAR biodegradation. UCI Machine Learning Repository. DOI: https://doi.org/10.24432/C5H60M (2013)
2013
Earlier work this paper cites.
Buza, K.: BlogFeedback. UCI Machine Learning Repository. DOI: https://doi.org/10.24432/C58S3F (2014)
2014
Earlier work this paper cites.
Whiteson, D.: HIGGS. UCI Machine Learning Repository. DOI: https://doi.org/10.24432/C5V312 (2014)
2014
Earlier work this paper cites.
Harper, F.M., Konstan, J.A.: The movielens datasets: History and context. Acm transactions on interactive intelligent systems (tiis) 5
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
Lehmberg, O., Ritze, D., Meusel, R., Bizer, C.: A large public corpus of web tables containing time and context metadata. In: Proceedings of the 25th International Conference Companion on World Wide Web, pp. 75–76 (2016)
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Iyyer, M., Yih, W.-t., Chang, M.-W.: Search-based neural structured learning for sequential question answering. In: Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pp. 1821–1831 (2017)
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., Sutskever, I., et al
2019
Earlier work this paper cites.
Zhang, L., Zhang, S., Balog, K.: Table2vec: Neural word and entity embeddings for table population and retrieval. In: Proceedings of the 42nd International ACM SIGIR Conference on Research and Development in Information Retrieval, pp. 1029–1032 (2019)
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Reimers, N., Gurevych, I.: Sentence-bert: Sentence embeddings using siamese bert-networks. In: Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) (2019). Association for Computational Linguistics
2019
Earlier work this paper cites.
Houlsby, N., Giurgiu, A., Jastrzebski, S., Morrone, B., De Laroussilhe, Q., Gesmundo, A., Attariyan, M., Gelly, S.: Parameter-efficient transfer learning for nlp. In: International Conference on Machine Learning, pp. 2790–2799 (2019). PMLR
2019
Earlier work this paper cites.
Ansuini, A., Laio, A., Macke, J.H., Zoccolan, D.: Intrinsic dimension of data representations in deep neural networks. Advances in Neural Information Processing Systems 32
2019
Earlier work this paper cites.
2020
Earlier work this paper cites.
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J.D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al
2020
Earlier work this paper cites.
Yin, P., Neubig, G., Yih, W.-t., Riedel, S.: TaBERT: Pretraining for joint understanding of textual and tabular data. In: Jurafsky, D., Chai, J., Schluter, N., Tetreault, J. (eds.) Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, pp. 8413–8426. Association for Computational Linguistics, Online (2020). https://doi.org/10.18653/v1/2020.acl-main.745 . https://aclanthology.org/2020.acl-main.745
2020
Earlier work this paper cites.
Yoon, J., Zhang, Y., Jordon, J., Schaar, M.: Vime: Extending the success of self-and semi-supervised learning to tabular domain. Advances in Neural Information Processing Systems 33
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Herzig, J., Nowak, P.K., Mueller, T., Piccinno, F., Eisenschlos, J.: Tapas: Weakly supervised table parsing via pre-training. In: Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, pp. 4320–4333 (2020)
2020
Earlier work this paper cites.
Johnson, A., Bulgarelli, L., Pollard, T., Horng, S., Celi, L.A., Mark, R.: Mimic-iv. PhysioNet. Available online at: https://physionet. org/content/mimiciv/1.0/(accessed August 23, 2021), 49–55 (2020)
2020
Earlier work this paper cites.
Somani, S., Russak, A.J., Richter, F., Zhao, S., Vaid, A., Chaudhry, F., De Freitas, J.K., Naik, N., Miotto, R., Nadkarni, G.N., et al
2021
Earlier work this paper cites.
Gianfrancesco, M.A., Goldstein, N.D.: A narrative review on the validity of electronic health record-based research in epidemiology. BMC medical research methodology 21
2021
Earlier work this paper cites.
Qin, J., Zhang, W., Su, R., Liu, Z., Liu, W., Tang, R., He, X., Yu, Y.: Retrieval & interaction machine for tabular data prediction. In: Proceedings of the 27th ACM SIGKDD Conference on Knowledge Discovery & Data Mining, pp. 1379–1389 (2021)
2021
Earlier work this paper cites.
Liu, Q., Chen, B., Guo, J., Ziyadi, M., Lin, Z., Chen, W., Lou, J.-G.: Tapex: Table pre-training via learning a neural sql executor. In: International Conference on Learning Representations (2021)
2021
Cited alongside, same era.
Iida, H., Thai, D., Manjunatha, V., Iyyer, M.: Tabbie: Pretrained representations of tabular data. In: Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, pp. 3446–3456 (2021)
2021
Cited alongside, same era.
Arik, S.Ö., Pfister, T.: Tabnet: Attentive interpretable tabular learning. In: Proceedings of the AAAI Conference on Artificial Intelligence, vol. 35, pp. 6679–6687 (2021)
2021
Cited alongside, same era.
2021
Cited alongside, same era.
Zhang, T., Xu, H., Genabith, J., Xiong, D., Zan, H.: Napg: Non-autoregressive program generation for hybrid tabular-textual question answering. In: CCF International Conference on Natural Language Processing and Chinese Computing, pp. 591–603 (2023). Springer
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Sui, Y., Zhou, M., Zhou, M., Han, S., Zhang, D.: GPT4Table: Can Large Language Models Understand Structured Table Data? A Benchmark and Empirical Study (2023)
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ucar, T., Hajiramezanali, E., Edwards, L.: Subtab: Subsetting features of tabular data for self-supervised representation learning. Advances in Neural Information Processing Systems 34
2021
Cited alongside, same era.
2021
Cited alongside, same era.
Zhu, F., Lei, W., Huang, Y., Wang, C., Zhang, S., Lv, J., Feng, F., Chua, T.-S.: Tat-qa: A question answering benchmark on a hybrid of tabular and textual content in finance. In: Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers), pp. 3277–3287 (2021)
2021
Cited alongside, same era.
Wang, Z., Dong, H., Jia, R., Li, J., Fu, Z., Han, S., Zhang, D.: Tuta: Tree-based transformers for generally structured table pre-training. In: Proceedings of the 27th ACM SIGKDD Conference on Knowledge Discovery & Data Mining, pp. 1780–1790 (2021)
2021
Cited alongside, same era.
Padhi, I., Schiff, Y., Melnyk, I., Rigotti, M., Mroueh, Y., Dognin, P., Ross, J., Nair, R., Altman, E.: Tabular transformers for modeling multivariate time series. In: ICASSP 2021-2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp. 3565–3569 (2021). IEEE
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
Ye, Y., Hui, B., Yang, M., Li, B., Huang, F., Li, Y.: Large language models are versatile decomposers: Decomposing evidence and questions for table-based reasoning. In: Proceedings of the 46th International ACM SIGIR Conference on Research and Development in Information Retrieval, pp. 174–184 (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Wang, Z., Gao, C., Xiao, C., Sun, J.: MediTab: Scaling Medical Tabular Data Predictors via Data Consolidation, Enrichment, and Refinement (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Nam, J., Song, W., Park, S.H., Tack, J., Yun, S., Kim, J., Shin, J.: Semi-supervised tabular classification via in-context learning of large language models. In: Workshop on Efficient Systems for Foundation Models@ ICML2023 (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Onishi, S., Oono, K., Hayashi, K.: Tabret: Pre-training transformer-based tabular models for unseen columns. In: ICLR 2023 Workshop on Mathematical and Empirical Understanding of Foundation Models (2023)
2023
Later among the works it cites.
Schambach, M., Paul, D., Otterbach, J.: Scaling experiments in self-supervised cross-table representation learning. In: NeurIPS 2023 Second Table Representation Learning Workshop (2023)
2023
Later among the works it cites.
Liu, Y., Gautam, S., Ma, J., Lakkaraju, H.: Investigating the fairness of large language models for predictions on tabular data. In: Socially Responsible Language Modelling Research (2023)
2023
Later among the works it cites.
Hegselmann, S., Buendia, A., Lang, H., Agrawal, M., Jiang, X., Sontag, D.: Tabllm: Few-shot classification of tabular data with large language models. In: International Conference on Artificial Intelligence and Statistics, pp. 5549–5581 (2023). PMLR
2023
Later among the works it cites.
Sarkar, S., Lausen, L.: Testing the limits of unified sequence to sequence llm pretraining on diverse table data tasks. In: NeurIPS 2023 Second Table Representation Learning Workshop (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Carballo, K.V., Na, L., Ma, Y., Boussioux, L., Zeng, C., Soenksen, L.R., Bertsimas, D.: TabText: A Flexible and Contextual Approach to Tabular Data Representation. Jul (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Belyaeva, A., Cosentino, J., Hormozdiari, F., Eswaran, K., Shetty, S., Corrado, G., Carroll, A., McLean, C.Y., Furlotte, N.A.: Multimodal llms for health grounded in individual-specific data. In: Workshop on Machine Learning for Multimodal Healthcare Data, pp. 86–102 (2023). Springer
2023
Later among the works it cites.
2023
Later among the works it cites.
Zhang, D., Wang, L., Dai, X., Jain, S., Wang, J., Fan, Y., Yeh, C.-C.M., Zheng, Y., Zhuang, Z., Zhang, W.: Fata-trans: Field and time-aware transformer for sequential tabular data. In: Proceedings of the 32nd ACM International Conference on Information and Knowledge Management, pp. 3247–3256 (2023)
2023
Later among the works it cites.
Zhang, T., Wang, S., Yan, S., Jian, L., Liu, Q.: Generative table pre-training empowers models for tabular prediction. In: Bouamor, H., Pino, J., Bali, K. (eds.) Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, pp. 14836–14854. Association for Computational Linguistics, Singapore (2023). https://doi.org/10.18653/v1/2023.emnlp-main.917 . https://aclanthology.org/2023.emnlp-main.917
2023
Later among the works it cites.
Nam, J., Song, W., Park, S.H., Tack, J., Yun, S., Kim, J., Shin, J.: Semi-supervised tabular classification via in-context learning of large language models. In: Workshop on Efficient Systems for Foundation Models @ ICML2023 (2023). https://openreview.net/forum?id=r77CeOBO0L
2023
Later among the works it cites.
Ye, Y., Hui, B., Yang, M., Li, B., Huang, F., Li, Y.: Large language models are versatile decomposers: Decomposing evidence and questions for table-based reasoning. In: Proceedings of the 46th International ACM SIGIR Conference on Research and Development in Information Retrieval. SIGIR ’23, pp. 174–184. Association for Computing Machinery, New York, NY, USA (2023). https://doi.org/10.1145/3539618.3591708 . https://doi.org/10.1145/3539618.3591708
2023
Later among the works it cites.
2023
Later among the works it cites.
He, K., Huang, Y., Mao, R., Gong, T., Li, C., Cambria, E.: Virtual prompt pre-training for prototype-based few-shot relation extraction. Expert Systems with Applications 213
2023
Later among the works it cites.
Hsieh, C.-Y., Li, C.-L., YEH, C.-K., Nakhost, H., Fujii, Y., Ratner, A.J., Krishna, R., Lee, C.-Y., Pfister, T.: Distilling step-by-step! outperforming larger language models with less training data and smaller model sizes. In: The 61st Annual Meeting Of The Association For Computational Linguistics (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Bonacin, R., Vechi, S.M., Dametto, M., Ruppert, G.C.S.: Exploring deep learning techniques in the prediction of cancer relapse using an open brazilian tabular database. In: International Conference on Information Technology-New Generations, pp. 331–338 (2024). Springer
2024
Closest in time.
Gandhar, A., Gupta, K., Pandey, A.K., Raj, D.: Fraud detection using machine learning and deep learning. SN Computer Science 5
2024
Closest in time.
Ghebrehiwet, I., Zaki, N., Damseh, R., Mohamad, M.S.: Revolutionizing personalized medicine with generative ai: a systematic review. Artificial Intelligence Review 57
2024
Closest in time.
Ruan, Y., Lan, X., Tan, D.J., Abdullah, H.R., Feng, M.: P-Transformer: A Prompt-based Multimodal Transformer Architecture For Medical Tabular Data (2024)
2024
Closest in time.
Wang, Z., Zhang, H., Li, C.-L., Eisenschlos, J.M., Perot, V., Wang, Z., Miculicich, L., Fujii, Y., Shang, J., Lee, C.-Y., Pfister, T.: Chain-of-table: Evolving tables in the reasoning chain for table understanding. In: The Twelfth International Conference on Learning Representations (2024). https://openreview.net/forum?id=4L0xnS4GQM
2024
Closest in time.
2024
Closest in time.
Schmidhuber, M., Kruschwitz, U.: Llm-based synthetic datasets: Applications and limitations in toxicity detection. LREC-COLING 2024, 37 (2024)
2024
Closest in time.
Liu, N.F., Lin, K., Hewitt, J., Paranjape, A., Bevilacqua, M., Petroni, F., Liang, P.: Lost in the middle: How language models use long contexts. Transactions of the Association for Computational Linguistics 12
2024
Closest in time.