Fetching the paper…
Reading the bibliography…
We introduce Aligner, a novel Parameter-Efficient Fine-Tuning (PEFT) method for aligning multi-billion-parameter-sized Large Language Models (LLMs).
J. Tenenbaum and W. Freeman, “Separating style and content,” in Advances in Neural Information Processing Systems , M. Mozer, M. Jordan, and T. Petsche, Eds., vol. 9. MIT Press, 1996
1996
Earlier work this paper cites.
M. A. O. Vasilescu and D. Terzopoulos, “Multilinear (tensor) image synthesis, analysis, and recognition [exploratory DSP],” IEEE Signal Processing Magazine , vol. 24, no. 6, pp. 118–123, 2007
2007
Earlier work this paper cites.
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in Neural Information Processing Systems , vol. 30, 2017
2017
Earlier work this paper cites.
M. Gazzaniga, R. Ivry, and G. Mangun, Cognitive Neuroscience: The Biology of the Mind . W.W. Norton, 2019
2019
Earlier work this paper cites.
N. Houlsby, A. Giurgiu, S. Jastrzebski, B. Morrone, Q. De Laroussilhe, A. Gesmundo, M. Attariyan, and S. Gelly, “Parameter-efficient transfer learning for NLP,” in International Conference on Machine Learning , 2019, pp. 2790–2799
2019
Earlier work this paper cites.
2021
Earlier work this paper cites.
D. Hendrycks, C. Burns, S. Basart, A. Zou, M. Mazeika, D. Song, and J. Steinhardt, “Measuring massive multitask language understanding,” Proceedings of the International Conference on Learning Representations (ICLR) , 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
X. L. Li and P. Liang, “Prefix-Tuning: Optimizing continuous prompts for generation,” in Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) . Association for Computational Linguistics, Aug. 2021, pp. 4582–4597
2021
Earlier work this paper cites.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
L. Yao, X. Wan, J. Xiao, B. Peng, and M. Zhang, “LoRA: Low-rank adaptation of large language models,” in Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing , 2021
2021
Cited alongside, same era.
2023
Closest in time.
OpenAI Forum, “Fine-tuning for domain knowledge and questions,” 2023, accessed: 2023-09-27. [Online]. Available: https://community.openai.com/t/finetuning-for-domain-knowledge-and-questions/24817
2023
Closest in time.
C. Qian, H. Tang, Z. Yang, H. Liang, and Y. Liu, “Can large language models empower molecular property prediction?” 2023
2023
Closest in time.
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
AnyScale, “Fine-tuning LLMs: LoRA or full parameter? an in-depth analysis with LLAMA,” 2023, accessed: 2023-09-28. [Online]. Available: https://www.anyscale.com/blog/fine-tuning-llms-lora-or-full-parameter-an-in-depth-analysis-with-llama-2
2023
Cited alongside, same era.
Anyscale, “Fine-tuning is for form, not facts,” 2023, accessed: 2023-09-27. [Online]. Available: https://www.anyscale.com/blog/fine-tuning-is-for-form-not-facts
2023
Cited alongside, same era.
W.-L. Chiang, Z. Li, Z. Lin, Y. Sheng, Z. Wu, H. Zhang, L. Zheng, S. Zhuang, Y. Zhuang, J. E. Gonzalez, I. Stoica, and E. P. Xing, “Vicuna: An open-source chatbot impressing GPT-4 with 90% ChatGPT quality,” March 2023. [Online]. Available: https://lmsys.org/blog/2023-03-30-vicuna/
2023
Cited alongside, same era.
J. Dai, X. Pan, J. Ji, R. Sun, Y. Wang, and Y. Yang, “PKU-Beaver: Constrained value-aligned LLM via safe RLHF,” 2023. [Online]. Available: https://github.com/PKU-Alignment/safe-rlhf
2023
Cited alongside, same era.
2023
Cited alongside, same era.
R. Taori, I. Gulrajani, T. Zhang, Y. Dubois, X. Li, C. Guestrin, P. Liang, and T. B. Hashimoto, “Stanford Alpaca: An instruction-following LLaMA model,” 2023. [Online]. Available: https://github.com/tatsu-lab/stanford_alpaca
2023
Closest in time.
H. Touvron, L. Martin, K. Stone, P. Albert, A. Almahairi, Y. Babaei, N. Bashlykov, S. Batra, P. Bhargava, S. Bhosale et al. , “Introducing LLaMA: A foundational, 65-billion-parameter language model,” Meta AI Blog , 2023. [Online]. Available: https://ai.meta.com/llama/
2023
Closest in time.
2023
Closest in time.
L. Weng, “Prompt engineering,” lilianweng.github.io , Mar 2023. [Online]. Available: https://lilianweng.github.io/posts/2023-03-15-prompt-engineering/
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.