Fetching the paper…
Reading the bibliography…
This paper presents an analysis of open-source large language models (LLMs) and their application in Retrieval-Augmented Generation (RAG) tasks, specific for enterprise-specific data sets scraped from their websites.
“Dense Passage Retrieval for Open-Domain Question Answering”
Vladimir Karpukhin et al · 2004
Earlier work this paper cites.
“ColBERT: Efficient and Effective Passage Search via Contextualized Late Interaction over BERT”
Omar Khattab and Matei Zaharia · 2004
Earlier work this paper cites.
“Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks”
Patrick Lewis et al · 2020
Earlier work this paper cites.
URL: https://huggingface.co/BAAI/bge-large-en-v1.5
BAAI embedding documentation, 2023 · 2023
Earlier work this paper cites.
“RAGAS: Automated Evaluation of Retrieval Augmented Generation”
Shahul Es, Jithin James, Luis Espinosa-Anke and Steven Schockaert · 2023
Earlier work this paper cites.
“MTEB: Massive Text Embedding Benchmark”
Niklas Muennighoff1, Nouamane andLoïc Magne1 and Nils Reimers · 2023
Earlier work this paper cites.
“A Cost Analysis of Generative Language Models and Influence Operations”
Micah Musser · 2023
Earlier work this paper cites.
Anupam Purwar and Rahul Sundar · 2023
Earlier work this paper cites.
“UL2: Unifying Language Learning Paradigms”
Yi Tay et al · 2023
Earlier work this paper cites.
“C-Pack: Packaged Resources To Advance General Chinese Embedding”, 2023
Shitao Xiao, Zheng Liu, Peitian Zhang and Niklas Muennighoff · 2023
Earlier work this paper cites.
URL: https://docs.confident-ai.com/
deepeval documentation, 2024 · 2024
Cited alongside, same era.
URL: https://python.langchain.com/v0.1/docs/get_started/introduction
langchain documentation, 2024 · 2024
Cited alongside, same era.
URL: https://api.python.langchain.com/en/latest/nltk/langchain_text_splitters.nltk.NLTKTextSplitter.html
NLTKsplitter documentation, 2024 · 2024
Cited alongside, same era.
URL: https://python.langchain.com/v0.1/docs/modules/data_connection/document_transformers/recursive_text_splitter/
RecursiveCharacterTextSplitter langchain documentation, 2024 · 2024
Cited alongside, same era.
URL: https://huggingface.co/models
huggingface documentation, 2024 · 2024
Cited alongside, same era.
URL: https://python.langchain.com/v0.2/docs/integrations/vectorstores/faiss/
FAISS documentation, 2024 · 2024
URL: https://openai.com/api/pricing/
OpenAI API pricing documentation, 2024 · 2024
Closest in time.
URL: https://docs.perplexity.ai/docs/pricing
Perplexity pricing documentation, 2024 · 2024
Closest in time.
URL: https://docs.neuralmagic.com/get-started/finetune/
neuralmagic documentation, 2024 · 2024
Closest in time.
URL: https://blog.lancedb.com/optimizing-llms-a-step-by-step-guide-to-fine-tuning-with-peft-and-qlora-22eddd13d25b/
lancedb documentation, 2024 · 2024
Closest in time.
URL: https://python.langchain.com/v0.2/docs/integrations/retrievers/bm25/
BM25 documentation, 2024 · 2024
Closest in time.
“COS-Mix: Cosine Similarity and Distance Fusion for Improved Information Retrieval”, 2024
Kush Juvekar and Anupam Purwar · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
URL: https://github.com/amaze18/dlabs_hybrid_search/blob/main/pplx.py
Perplexity wrapper code, 2024 · 2024
Cited alongside, same era.
URL: https://docs.perplexity.ai/docs/getting-started
perplexity api documentation, 2024 · 2024
Cited alongside, same era.
URL: https://docs.perplexity.ai/docs/model-cards
Perplexity models documentation, 2024 · 2024
Cited alongside, same era.
Closest in time.
“A FINE-TUNING ENHANCED RAG SYSTEM WITH QUANTIZED INFLUENCE MEASURE AS AI JUDGE”
Keshav Rangan and Yiqiao Yin · 2024
Closest in time.
“Evaluation of Retrieval-Augmented Generation: A Survey”
Hao Yu et al · 2024
Closest in time.
“Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena”, 2024
Lianmin Zheng et al · 2024
Closest in time.