Fetching the paper…
Reading the bibliography…
Answering real-world complex queries, such as complex product search, often requires accurate retrieval from semi-structured knowledge bases that involve blend of unstructured (e.g., textual descriptions of products) and structured (e.g., entity relations of products) information.
A Mathematical Theory of Communication
C. E. Shannon. 1948 · 1948
Earlier work this paper cites.
Certain Language Skills in Children: Their Development and Interrelationships . Vol. 26
MILDRED C. TEMPLIN. 1957 · 1957
Earlier work this paper cites.
Natural language question answering: the view from here
Lynette Hirschman and Robert Gaizauskas. 2001 · 2001
Earlier work this paper cites.
Open Graph Benchmark: Datasets for Machine Learning on Graphs
Weihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong, Hongyu Ren, Bowen Liu, Michele Catasta, and Jure Leskovec. 2021 · 2005
Earlier work this paper cites.
Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen tau Yih, Tim Rocktäschel, Sebastian Riedel, and Douwe Kiela. 2021 · 2005
Earlier work this paper cites.
The Probabilistic Relevance Framework: BM25 and Beyond
Stephen E. Robertson and Hugo Zaragoza. 2009 · 2009
Earlier work this paper cites.
Evaluating the usability of natural language query languages and interfaces to Semantic Web knowledge bases
Esther Kaufmann and Abraham Bernstein. 2010 · 2010
Earlier work this paper cites.
Semantic Parsing on Freebase from Question-Answer Pairs. In EMNLP
Jonathan Berant, Andrew Chou, Roy Frostig, and Percy Liang. 2013 · 2013
Earlier work this paper cites.
Question Answering with Subgraph Embeddings. In Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP) , Alessandro Moschitti, Bo Pang, and Walter Daelemans (Eds.). Association for Computational Linguistics, Doha, Qatar
Antoine Bordes, Sumit Chopra, and Jason Weston. 2014 · 2014
Earlier work this paper cites.
Discovering OLAP dimensions in semi-structured data
Svetlana Mansmann, Nafees Ur Rehman, Andreas Weiler, and Marc H. Scholl. 2014 · 2014
Earlier work this paper cites.
Open domain question answering using Wikipedia-based knowledge model
Pum-Mo Ryu, Myung-Gil Jang, and Hyunki Kim. 2014 · 2014
Earlier work this paper cites.
Large-scale Simple Question Answering with Memory Networks
Antoine Bordes, Nicolas Usunier, Sumit Chopra, and Jason Weston. 2015 · 2015
Earlier work this paper cites.
Image-Based Recommendations on Styles and Substitutes. In SIGIR . ACM
Julian J. McAuley, Christopher Targett, Qinfeng Shi, and Anton van den Hengel. 2015 · 2015
Earlier work this paper cites.
Compositional Semantic Parsing on Semi-Structured Tables. Association for Computational Linguistics
Panupong Pasupat and Percy Liang. 2015 · 2015
Earlier work this paper cites.
An Overview of Microsoft Academic Service (MAS) and Applications. In WWW
Arnab Sinha, Zhihong Shen, Yang Song, Hao Ma, Darrin Eide, Bo-June Paul Hsu, and Kuansan Wang. 2015 · 2015
Earlier work this paper cites.
Semantic Parsing via Staged Query Graph Generation: Question Answering with Knowledge Base. In ACL
Wen-tau Yih, Ming-Wei Chang, Xiaodong He, and Jianfeng Gao. 2015 · 2015
Earlier work this paper cites.
Ups and Downs: Modeling the Visual Evolution of Fashion Trends with One-Class Collaborative Filtering. In WWW . ACM
Ruining He and Julian J. McAuley. 2016 · 2016
Earlier work this paper cites.
TabMCQ: A Dataset of General Knowledge Tables and Multiple-choice Questions
Sujay Kumar Jauhar, Peter D. Turney, and Eduard H. Hovy. 2016 · 2016
Earlier work this paper cites.
MS MARCO: A Human Generated MAchine Reading COmprehension Dataset. In Proceedings of the Workshop on Cognitive Computation: Integrating neural and symbolic approaches 2016 co-located with the 30th Annual Conference on Neural Information Processing Systems (NIPS 2016), Barcelona, Spain, December 9, 2016 (CEUR Workshop Proceedings)
Tri Nguyen, Mir Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, and Li Deng. 2016 · 2016
Earlier work this paper cites.
SQuAD: 100,000+ Questions for Machine Comprehension of Text. In Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Austin, Texas
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Cited alongside, same era.
The Value of Semantic Parse Labeling for Knowledge Base Question Answering. Association for Computational Linguistics, Berlin, Germany
Wen-tau Yih, Matthew Richardson, Chris Meek, Ming-Wei Chang, and Jina Suh. 2016 · 2016
Cited alongside, same era.
SearchQA: A New QA Dataset Augmented with Context from a Search Engine
Matthew Dunn, Levent Sagun, Mike Higgins, V. Ugur Guney, Volkan Cirik, and Kyunghyun Cho. 2017 · 2017
Cited alongside, same era.
Knowledge Rich Natural Language Queries over Structured Biological Databases. In Proceedings of the 8th ACM International Conference on Bioinformatics, Computational Biology, and Health Informatics
Hasan M. Jamil. 2017 · 2017
Cited alongside, same era.
Leveraging Passage Retrieval with Generative Models for Open Domain Question Answering. In EACL
Gautier Izacard and Edouard Grave. 2021 · 2021
Later among the works it cites.
Precision Medicine, AI, and the Future of Personalized Health Care
Kevin B. Johnson, Wei-Qi Wei, Dharini Weeraratne, Mark E. Frisse, Kevin Misulis, Kevin Rhee, Jie Zhao, and J. L. Snowdon. 2021 · 2021
Later among the works it cites.
Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval. In ICLR
Lee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang, Jialin Liu, Paul N. Bennett, Junaid Ahmed, and Arnold Overwijk. 2021 · 2021
Later among the works it cites.
QA-GNN: Reasoning with Language Models and Knowledge Graphs for Question Answering
Michihiro Yasunaga, Hongyu Ren, Antoine Bosselut, Percy Liang, and Jure Leskovec. 2021 · 2021
Later among the works it cites.
KQA Pro: A Dataset with Explicit Compositional Programs for Complex Question Answering over Knowledge Base. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension. In ACL
Mandar Joshi, Eunsol Choi, Daniel S. Weld, and Luke Zettlemoyer. 2017 · 2017
Cited alongside, same era.
NewsQA: A Machine Comprehension Dataset. In Proceedings of the 2nd Workshop on Representation Learning for NLP . Association for Computational Linguistics
Adam Trischler, Tong Wang, Xingdi Yuan, Justin Harris, Alessandro Sordoni, Philip Bachman, and Kaheer Suleman. 2017 · 2017
Cited alongside, same era.
Seq2SQL: Generating Structured Queries from Natural Language using Reinforcement Learning
Victor Zhong, Caiming Xiong, and Richard Socher. 2017 · 2017
Cited alongside, same era.
Constructing Datasets for Multi-hop Reading Comprehension Across Documents
Johannes Welbl, Pontus Stenetorp, and Sebastian Riedel. 2018 · 2018
Cited alongside, same era.
HotpotQA: A dataset for diverse, explainable multi-hop question answering
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William W Cohen, Ruslan Salakhutdinov, and Christopher D Manning. 2018b · 2018
Cited alongside, same era.
Spider: A Large-Scale Human-Labeled Dataset for Complex and Cross-Domain Semantic Parsing and Text-to-SQL Task. Association for Computational Linguistics, Brussels, Belgium
Tao Yu, Rui Zhang, Kai Yang, Michihiro Yasunaga, Dongxu Wang, Zifan Li, James Ma, Irene Li, Qingning Yao, Shanelle Roman, Zilin Zhang, and Dragomir Radev. 2018 · 2018
Cited alongside, same era.
Querying Knowledge via Multi-Hop English Questions
Tiantian Gao, Paul Fodor, and Michael Kifer. 2019 · 2019
Cited alongside, same era.
Natural Questions: A Benchmark for Question Answering Research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, Kristina Toutanova, Llion Jones, Matthew Kelcey, Ming-Wei Chang, Andrew M. Dai, Jakob Uszkoreit, Quoc Le, and Slav Petrov. 2019 · 2019
Cited alongside, same era.
Shulin Cao, Jiaxin Shi, Liangming Pan, Lunyiu Nie, Yutong Xiang, Lei Hou, Juanzi Li, Bin He, and Hanwang Zhang. 2022 · 2022
Later among the works it cites.
UniK-QA: Unified Representations of Structured and Unstructured Knowledge for Open-Domain Question Answering. In ACL Findings
Barlas Oguz, Xilun Chen, Vladimir Karpukhin, Stan Peshterliev, Dmytro Okhonko, Michael Sejr Schlichtkrull, Sonal Gupta, Yashar Mehdad, and Scott Yih. 2022 · 2022
Later among the works it cites.
ColBERTv2: Effective and Efficient Retrieval via Lightweight Late Interaction. In NAACL
Keshav Santhanam, Omar Khattab, Jon Saad-Falcon, Christopher Potts, and Matei Zaharia. 2022 · 2022
Later among the works it cites.
Voyage AI Embeddings API
Voyage AI. 2023 · 2023
Later among the works it cites.
Building a knowledge graph to enable precision medicine
Payal Chandak, Kexin Huang, and Marinka Zitnik. 2023 · 2023
Later among the works it cites.
INSTRUCTEVAL: Towards Holistic Evaluation of Instruction-Tuned Large Language Models
Yew Ken Chia, Pengfei Hong, Lidong Bing, and Soujanya Poria. 2023 · 2023
Later among the works it cites.
OpenAI Embeddings API
OpenAI. 2023 · 2023
Later among the works it cites.
REPLUG: Retrieval-Augmented Black-Box Language Models
Weijia Shi, Sewon Min, Michihiro Yasunaga, Minjoon Seo, Rich James, Mike Lewis, Luke Zettlemoyer, and Wen-tau Yih. 2023 · 2023
Later among the works it cites.
Head-to-Tail: How Knowledgeable are Large Language Models (LLM)? A.K.A. Will LLMs Replace Knowledge Graphs?
Kai Sun, Yifan Ethan Xu, Hanwen Zha, Yue Liu, and Xin Luna Dong. 2023 · 2023
Later among the works it cites.
Large language models for information retrieval: A survey
Yutao Zhu, Huaying Yuan, Shuting Wang, Jiongnan Liu, Wenhan Liu, Chenlong Deng, Zhicheng Dou, and Ji-Rong Wen. 2023 · 2023
Later among the works it cites.
Beyond Yes and No: Improving Zero-Shot LLM Rankers via Scoring Fine-Grained Relevance Labels
Honglei Zhuang, Zhen Qin, Kai Hui, Junru Wu, Le Yan, Xuanhui Wang, and Michael Bendersky. 2023 · 2023
Later among the works it cites.
LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders
Parishad BehnamGhader, Vaibhav Adlakha, Marius Mosbach, Dzmitry Bahdanau, Nicolas Chapados, and Siva Reddy. 2024 · 2024
Closest in time.
G-Retriever: Retrieval-Augmented Generation for Textual Graph Understanding and Question Answering
Xiaoxin He, Yijun Tian, Yifei Sun, Nitesh V. Chawla, Thomas Laurent, Yann LeCun, Xavier Bresson, and Bryan Hooi. 2024 · 2024
Closest in time.
Generative Representational Instruction Tuning
Niklas Muennighoff, Hongjin Su, Liang Wang, Nan Yang, Furu Wei, Tao Yu, Amanpreet Singh, and Douwe Kiela. 2024 · 2024
Closest in time.