Fetching the paper…
Reading the bibliography…
Retrieval-Augmented Generation (RAG) systems are showing promising potential, and are becoming increasingly relevant in AI-powered legal applications.
Question answering for privacy policies: Combining computational and legal perspectives, 2019
Abhilasha Ravichander, Alan W Black, Shomir Wilson, Thomas Norton, and Norman Sadeh · 1911
Earlier work this paper cites.
Dense passage retrieval for open-domain question answering, 2020
Vladimir Karpukhin, Barlas Oğuz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen tau Yih · 2004
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive NLP tasks
Patrick S. H. Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, Sebastian Riedel, and Douwe Kiela · 2005
Earlier work this paper cites.
Hotpotqa: A dataset for diverse, explainable multi-hop question answering, 2018
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William W. Cohen, Ruslan Salakhutdinov, and Christopher D. Manning · 2018
Earlier work this paper cites.
Cuad: An expert-annotated nlp dataset for legal contract review, 2021
Dan Hendrycks, Collin Burns, Anya Chen, and Spencer Ball · 2021
Earlier work this paper cites.
Contractnli: A dataset for document-level natural language inference for contracts, 2021
Yuta Koreeda and Christopher D. Manning · 2021
Earlier work this paper cites.
LexGLUE: A benchmark dataset for legal language understanding in English
Ilias Chalkidis, Abhik Jana, Dirk Hartung, Michael Bommarito, Ion Androutsopoulos, Daniel Katz, and Nikolaos Aletras · 2022
Earlier work this paper cites.
Stelios Maroudas, Sotiris Legkas, Prodromos Malakasiotis, and Ilias Chalkidis · 2022
Earlier work this paper cites.
Multi-lexsum: Real-world summaries of civil rights lawsuits at multiple granularities, 2022
Zejiang Shen, Kyle Lo, Lauren Yu, Nathan Dahlberg, Margo Schlanger, and Doug Downey · 2022
Cited alongside, same era.
Legal case document summarization: Extractive and abstractive methods and their evaluation, 2022
Abhay Shukla, Paheli Bhattacharya, Soham Poddar, Rajdeep Mukherjee, Kripabandhu Ghosh, Pawan Goyal, and Saptarshi Ghosh · 2022
Cited alongside, same era.
Benchmarking large language models in retrieval-augmented generation, 2023
Jiawei Chen, Hongyu Lin, Xianpei Han, and Le Sun · 2023
Cited alongside, same era.
Neel Guha, Julian Nyarko, Daniel E. Ho, Christopher Ré, Adam Chilton, Aditya Narayana, Alex Chohlas-Wood, Austin Peters, Brandon Waldon, Daniel N. Rockmore, Diego Zambrano, Dmitry Talisman, Enam Hoque, Faiz Surani, Frank Fagan, Galit Sarfaty, Gregory M. Dickinson, Haggai Porat, Jason Hegland, Jessica Wu, Joe Nudell, Joel Niklaus, John Nay, Jonathan H. Choi, Kevin Tobia, Margaret Hagan, Megan Ma, Michael Livermore, Nikon Rasumov-Rahe, Nils Holzenberger, Noam Kolt, Peter Henderson, Sean Rehaag, Sharad Goel, Shang Gao, Spencer Williams, Sunny Gandhi, Tom Zur, Varun Iyer, and Zehua Li · 2023
Rerank, 2024
Cohere · 2024
Closest in time.
Levels of text splitting
FullStackRetrieval-com · 2024
Closest in time.
Recursive text splitter, 2024
LangChain · 2024
Closest in time.
Cheng Niu, Yuanhao Wu, Juno Zhu, Siliang Xu, Kashun Shum, Randy Zhong, Juntong Song, and Tong Zhang · 2024
Closest in time.
Embedding models, 2024
OpenAI · 2024
Closest in time.
Chunking strategies, 2024
Pinecone · 2024
Closest in time.
Multihop-rag: Benchmarking retrieval-augmented generation for multi-hop queries, 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Natural language processing in the legal domain, 2023
Daniel Martin Katz, Dirk Hartung, Lauritz Gerlach, Abhik Jana, and Michael J. Bommarito II au2 · 2023
Cited alongside, same era.
Don’t use a cannon to kill a fly: An efficient cascading pipeline for long documents
Zehua Li, Neel Guha, and Julian Nyarko · 2023
Cited alongside, same era.
Recall: A benchmark for llms robustness against external counterfactual knowledge, 2023
Yi Liu, Lianzhe Huang, Shicheng Li, Sishuo Chen, Hao Zhou, Fandong Meng, Jie Zhou, and Xu Sun · 2023
Cited alongside, same era.
Maud: An expert-annotated legal nlp dataset for merger agreement understanding, 2023
Steven H. Wang, Antoine Scardigli, Leonard Tang, Wei Chen, Dimitry Levkin, Anya Chen, Spencer Ball, Thomas Woodside, Oliver Zhang, and Dan Hendrycks · 2023
Cited alongside, same era.
Yixuan Tang and Yi Yang · 2024
Closest in time.
Evaluation of retrieval-augmented generation: A survey, 2024
Hao Yu, Aoran Gan, Kai Zhang, Shiwei Tong, Qi Liu, and Zhaofeng Liu · 2024
Closest in time.