Fetching the paper…
Reading the bibliography…
Domain-specific question answering remains challenging for language models, given the deep technical knowledge required to answer questions correctly.
S. Robertson, S. Walker, S. Jones, M. M. Hancock-Beaulieu, and M. Gatford, “Okapi at trec-3,” in Overview of the Third Text REtrieval Conference (TREC-3) , January 1995, pp. 109–126
1995
Earlier work this paper cites.
2005
Earlier work this paper cites.
I. Dagan, O. Glickman, and B. Magnini, “The PASCAL Recognising Textual Entailment Challenge,” in Machine Learning Challenges. Evaluating Predictive Uncertainty, Visual Object Classification, and Recognising Tectual Entailment , 2006, pp. 177–190
2006
Earlier work this paper cites.
2007
Earlier work this paper cites.
2009
Earlier work this paper cites.
H. Shima, H. Kanayama, C.-W. Lee, C.-J. Lin, T. Mitamura, Y. Miyao, S. Shi, and K. Takeda, “Overview of NTCIR-9 RITE: Recognizing Inference in TExt,” in Proceedings of the 9th NTCIR Workshop Meeting on Evaluation of Information Access Technologies: Information Retrieval, Question Answering and Cross-Lingual Information Access, NTCIR-9, National Center of Sciences, Tokyo, Japan, December 6-9, 2011 , Dec. 2011
2011
Earlier work this paper cites.
A. Williams, N. Nangia, and S. Bowman, “A broad-coverage challenge corpus for sentence understanding through inference,” in Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers) , Jun. 2018, pp. 1112–1122
2018
Earlier work this paper cites.
2021
Earlier work this paper cites.
J. Johnson, M. Douze, and H. Jegou, “Billion-scale similarity search with gpus,” IEEE Transactions on Big Data , vol. 7, no. 3, pp. 535–547, Jul. 2021
2021
Cited alongside, same era.
2022
Cited alongside, same era.
Y. Alharahseheh, R. Obeidat, M. Al-Ayoub, and M. Gharaibeh, “A Survey on Textual Entailment: Benchmarks, Approaches and Applications,” in 2022 13th International Conference on Information and Communication Systems (ICICS) , Jun. 2022, pp. 328–336, iSSN: 2573-3346
2022
Cited alongside, same era.
A. Pal, L. K. Umapathi, and M. Sankarasubbu, “MedMCQA: A Large-scale Multi-Subject Multi-Choice Dataset for Medical domain Question Answering,” in Proceedings of the Conference on Health, Inference, and Learning , Apr. 2022, pp. 248–260, iSSN: 2640-3498
2022
Cited alongside, same era.
2023
Later among the works it cites.
A. Maatouk, F. Ayed, N. Piovesan, A. De Domenico, M. Debbah, and Z.-Q. Luo, “TeleQnA: A Benchmark Dataset to Assess Large Language Models Telecommunications Knowledge,” Oct. 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
OpenAI, “GPT-4 Technical Report,” Mar. 2024, arXiv:2303.08774 [cs]
2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
F. Yang, P. Zhao, Z. Wang, L. Wang, B. Qiao, J. Zhang, M. Garg, Q. Lin, S. Rajmohan, and D. Zhang, “Empower Large Language Model to Perform Better on Industrial Domain-Specific Question Answering,” in Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing: Industry Track , M. Wang and I. Zitouni, Eds. Singapore: Association for Computational Linguistics, Dec. 2023, pp. 294–312
2023
Cited alongside, same era.
P. Lee, S. Bubeck, and J. Petro, “Benefits, Limits, and Risks of GPT-4 as an AI Chatbot for Medicine,” New England Journal of Medicine , vol. 388, no. 13, pp. 1233–1239, Mar. 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
“Phi-2: The surprising power of small language models - Microsoft Research.” [Online]. Available: https://www.microsoft.com/en-us/research/blog/phi-2-the-surprising-power-of-small-language-models/
Cited in the paper.
OpenAI, “Gpt-4o mini: advancing cost-efficient intelligence.” [Online]. Available: https://openai.com/index/gpt-4o-mini-advancing-cost-efficient-intelligence/
Cited in the paper.
Anthropic, “Introducing Claude 3.5 Sonnet.” [Online]. Available: https://www.anthropic.com/news/claude-3-5-sonnet
Cited in the paper.
2024
Closest in time.
2024
Closest in time.
L. Team, “The Llama 3 Herd of Models,” Jul. 2024, arXiv:2407.21783 [cs]
2024
Closest in time.