Fetching the paper…
Reading the bibliography…
Large language models demonstrate remarkable proficiency in various linguistic tasks and have extensive knowledge across various domains.
“Measuring the semantic similarity of texts”
Courtney Corley and Rada Mihalcea · 2005
Earlier work this paper cites.
“Neural Machine Translation of Rare Words with Subword Units”
Rico Sennrich, Barry Haddow and Alexandra Birch · 2016
Earlier work this paper cites.
“Attention is all you need”
Ashish Vaswani et al · 2017
Earlier work this paper cites.
“A Call for Clarity in Reporting BLEU Scores”
Matt Post · 2018
Earlier work this paper cites.
“BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding”
Jacob Devlin, Ming-Wei Chang, Kenton Lee and Kristina Toutanova · 2019
Earlier work this paper cites.
“Parameter-efficient transfer learning for NLP”
Neil Houlsby et al · 2019
Earlier work this paper cites.
“Unsupervised Cross-lingual Representation Learning at Scale”
Alexis Conneau et al · 2020
Earlier work this paper cites.
“Making Monolingual Sentence Embeddings Multilingual using Knowledge Distillation”
Nils Reimers and Iryna Gurevych · 2020
Earlier work this paper cites.
“ParSQuAD: Persian Question Answering Dataset based on Machine Translation of SQuAD 2.0”
Negin Abadani, Jamshid Mozafari, Afsaneh Fatemi, Mohamadali Nematbakhsh and Arefeh Kazemi · 2021
Earlier work this paper cites.
“PersianQA: a dataset for Persian Question Answering”
Sajjad Ayoubi and Mohammad Davoodeh · 2021
Earlier work this paper cites.
“ParsBERT: Transformer-based Model for Persian Language Understanding”
Mehrdad Farahani, Mohammad Gharachorloo, Marzieh Farahani and Mohammad Manthouri · 2021
Earlier work this paper cites.
“Leveraging ParsBERT and Pretrained mT5 for Persian Abstractive Text Summarization”
Mehrdad Farahani, Mohammad Gharachorloo and M. Manthouri · 2021
Earlier work this paper cites.
“FarSick: A Persian Semantic Textual Similarity And Natural Language Inference Dataset”
Zahra Ghasemi and Mohammad Keyvanrad · 2021
Earlier work this paper cites.
“Measuring Massive Multitask Language Understanding”
Dan Hendrycks et al · 2021
Earlier work this paper cites.
“ParsiNLU: A Suite of Language Understanding Challenges for Persian”
Daniel Khashabi et al · 2021
Earlier work this paper cites.
“WikiMatrix: Mining 135M Parallel Sentences in 1620 Language Pairs from Wikipedia”
Holger Schwenk, Vishrav Chaudhary, Shuo Sun, Hongyu Gong and Francisco Guzmán · 2021
Earlier work this paper cites.
“Training neural networks with fixed sparse masks”
Yi-Lin Sung, Varun Nair and Colin Raffel · 2021
Earlier work this paper cites.
“mT5: A Massively Multilingual Pre-trained Text-to-Text Transformer”
Linting Xue et al · 2021
Cited alongside, same era.
“BitFit: Simple Parameter-efficient Fine-tuning for Transformer-based Masked Language-models”
Elad Ben, Yoav Goldberg and Shauli Ravfogel · 2022
Cited alongside, same era.
“No language left behind: Scaling human-centered machine translation”
Marta Costa-jussà et al · 2022
Cited alongside, same era.
“Language-agnostic BERT Sentence Embedding”
Fangxiaoyu Feng, Yinfei Yang, Daniel Cer, Naveen Arivazhagan and Wei Wang · 2022
Cited alongside, same era.
“SparseAdapter: An Easy Approach for Improving the Parameter-Efficiency of Adapters”
Shwai He, Liang Ding, Daize Dong, Jeremy Zhang and Dacheng Tao · 2022
Cited alongside, same era.
“Bactrian-X: A Multilingual Replicable Instruction-Following Model with Low-Rank Adaptation”
Haonan Li, Fajri Koto, Minghao Wu, Alham Aji and Timothy Baldwin · 2023
Later among the works it cites.
“AnglE-optimized Text Embeddings”
Xianming Li and Jing Li · 2023
Later among the works it cites.
“Scaling down to scale up: A guide to parameter-efficient fine-tuning”
Vladislav Lialin, Vijeta Deshpande and Anna Rumshisky · 2023
Later among the works it cites.
“XLM-V: Overcoming the Vocabulary Bottleneck in Multilingual Masked Language Models”
Davis Liang et al · 2023
Later among the works it cites.
“MPT” https://www.mosaicml.com/blog/mpt-7b [Accessed: 01-07-2024], 2023
MosaicML · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kevin Heffernan, Onur Çelebi and Holger Schwenk · 2022
Cited alongside, same era.
“LoRA: Low-Rank Adaptation of Large Language Models”
Edward Hu et al · 2022
Cited alongside, same era.
“ChatGPT” https://openai.com/blog/chatgpt [Accessed: 01-07-2024], 2022
OpenAI · 2022
Cited alongside, same era.
“COMET-22: Unbabel-IST 2022 Submission for the Metrics Shared Task”
Ricardo Rei et al · 2022
Cited alongside, same era.
“Efficient fine-tuning of bert models on the edge”
Danilo Vucetic et al · 2022
Cited alongside, same era.
“AdaMix: Mixture-of-Adaptations for Parameter-efficient Model Tuning”
Yaqing Wang et al · 2022
Cited alongside, same era.
“Sustainable AI: Environmental implications, challenges and opportunities”
Carole-Jean Wu et al · 2022
Cited alongside, same era.
“MTEB: Massive Text Embedding Benchmark”
Niklas Muennighoff, Nouamane Tazi, Loic Magne and Nils Reimers · 2023
Later among the works it cites.
“Orca: Progressive learning from complex explanation traces of gpt-4”
Subhabrata Mukherjee et al · 2023
Later among the works it cites.
“GPT-4 Technical Report”, 2023
OpenAI et al · 2023
Later among the works it cites.
Guilherme Penedo et al · 2023
Later among the works it cites.
“Stanford Alpaca: An Instruction-following LLaMA Model”
Rohan Taori et al · 2023
Later among the works it cites.
“Falcon” https://falconllm.tii.ae/falcon.html [Accessed: 01-07-2024], 2023
Tii · 2023
Later among the works it cites.
“Llama: Open and efficient foundation language models”
Hugo Touvron et al · 2023
Later among the works it cites.
“Llama 2: Open foundation and fine-tuned chat models”
Hugo Touvron et al · 2023
Later among the works it cites.
“Visual chatgpt: Talking, drawing and editing with visual foundation models”
Chenfei Wu et al · 2023
Later among the works it cites.
“01” https://01.ai/ [Accessed: 01-07-2024], 2023
Yi · 2023
Later among the works it cites.
“A survey of large language models”
Wayne Zhao et al · 2023
Later among the works it cites.