Fetching the paper…
Reading the bibliography…
While large language models (LLMs) have exhibited impressive instruction-following capabilities, it is still unclear whether and to what extent they can respond to explicit constraints that might be entailed in various instructions.
Language models are few-shot learners
Brown, T.; Mann, B.; Ryder, N.; Subbiah, M.; Kaplan, J. D.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; et al. 2020 · 1901
Earlier work this paper cites.
Plug and play language models: A simple approach to controlled text generation
Dathathri, S.; Madotto, A.; Lan, J.; Hung, J.; Frank, E.; Molino, P.; Yosinski, J.; and Liu, R. 2019 · 1912
Earlier work this paper cites.
An argument for basic emotions
Ekman, P. 1992 · 1992
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, K.; Roukos, S.; Ward, T.; and Zhu, W.-J. 2002 · 2002
Earlier work this paper cites.
Automatic evaluation of machine translation quality using longest common subsequence and skip-bigram statistics
Lin, C.-Y.; and Och, F. J. 2004 · 2004
Earlier work this paper cites.
Bounding the Capabilities of Large Language Models in Open Text Generation with Prompt Constraints
Lu, A.; Zhang, H.; Zhang, Y.; Wang, X.; and Yang, D. 2023 · 2008
Earlier work this paper cites.
The Curious Case of Neural Text Degeneration
Holtzman, A.; Buys, J.; Du, L.; Forbes, M.; and Choi, Y. 2019 · 2019
Earlier work this paper cites.
Positional Encoding to Control Output Sequence Length
Takase, S.; and Okazaki, N. 2019 · 2019
Earlier work this paper cites.
RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models
Gehman, S.; Gururangan, S.; Sap, M.; Choi, Y.; and Smith, N. A. 2020 · 2020
Earlier work this paper cites.
Measuring Massive Multitask Language Understanding
Hendrycks, D.; Burns, C.; Basart, S.; Zou, A.; Mazeika, M.; Song, D.; and Steinhardt, J. 2020 · 2020
Earlier work this paper cites.
CommonGen: A Constrained Text Generation Challenge for Generative Commonsense Reasoning
Lin, B. Y.; Zhou, W.; Shen, M.; Zhou, P.; Bhagavatula, C.; Choi, Y.; and Ren, X. 2020 · 2020
Earlier work this paper cites.
GeDi: Generative Discriminator Guided Sequence Generation
Krause, B.; Gotmare, A. D.; McCann, B.; Keskar, N. S.; Joty, S.; Socher, R.; and Rajani, N. F. 2021 · 2021
Earlier work this paper cites.
DExperts: Decoding-Time Controlled Text Generation with Experts and Anti-Experts
Liu, A.; Sap, M.; Lu, X.; Swayamdipta, S.; Bhagavatula, C.; Smith, N. A.; and Choi, Y. 2021 · 2021
Earlier work this paper cites.
Finetuned Language Models are Zero-Shot Learners
Wei, J.; Bosma, M.; Zhao, V.; Guu, K.; Yu, A. W.; Lester, B.; Du, N.; Dai, A. M.; and Le, Q. V. 2021 · 2021
Earlier work this paper cites.
FUDGE: Controlled Text Generation With Future Discriminators
Yang, K.; and Klein, D. 2021 · 2021
Cited alongside, same era.
Twitter Topic Classification
Antypas, D.; Ushio, A.; Camacho-Collados, J.; Silva, V.; Neves, L.; and Barbieri, F. 2022 · 2022
Cited alongside, same era.
Fine-grained controllable text generation using non-residual prompting
Carlsson, F.; Öhman, J.; Liu, F.; Verlinden, S.; Nivre, J.; and Sahlgren, M. 2022 · 2022
Cited alongside, same era.
Scaling Instruction-Finetuned Language Models
Chung, H. W.; Hou, L.; Longpre, S.; Zoph, B.; Tay, Y.; Fedus, W.; Li, Y.; Wang, X.; Dehghani, M.; Brahma, S.; Webson, A.; Gu, S. S.; Dai, Z.; Suzgun, M.; Chen, X.; Chowdhery, A.; Castro-Ros, A.; Pellat, M.; Robinson, K.; Valter, D.; Narang, S.; Mishra, G.; Yu, A.; Zhao, V.; Huang, Y.; Dai, A.; Yu, H.; Petrov, S.; Chi, E. H.; Dean, J.; Devlin, J.; Roberts, A.; Zhou, D.; Le, Q. V.; and Wei, J. 2022 · 2022
Cited alongside, same era.
GLM: General Language Model Pretraining with Autoregressive Blank Infilling
Du, Z.; Qian, Y.; Liu, X.; Ding, M.; Qiu, J.; Yang, Z.; and Tang, J. 2022 · 2022
GPT4All: Training an Assistant-style Chatbot with Large Scale Data Distillation from GPT-3.5-Turbo
Anand, Y.; Nussbaum, Z.; Duderstadt, B.; Schmidt, B.; and Mulyar, A. 2023 · 2023
Later among the works it cites.
Bai, J.; Bai, S.; Chu, Y.; Cui, Z.; Dang, K.; Deng, X.; Fan, Y.; Ge, W.; Han, Y.; Huang, F.; et al. 2023 · 2023
Later among the works it cites.
Vicuna: An Open-Source Chatbot Impressing GPT-4 with 90%* ChatGPT Quality
Chiang, W.-L.; Li, Z.; Lin, Z.; Sheng, Y.; Wu, Z.; Zhang, H.; Zheng, L.; Zhuang, S.; Zhuang, Y.; Gonzalez, J. E.; Stoica, I.; and Xing, E. P. 2023 · 2023
Later among the works it cites.
The Vendi Score: A Diversity Evaluation Metric for Machine Learning
Friedman, D.; and Dieng, A. B. 2023 · 2023
Later among the works it cites.
C-Eval: A Multi-Level Multi-Discipline Chinese Evaluation Suite for Foundation Models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A Distributional Lens for Multi-Aspect Controllable Text Generation
Gu, Y.; Feng, X.; Ma, S.; Zhang, L.; Gong, H.; and Qin, B. 2022 · 2022
Cited alongside, same era.
Unnatural Instructions: Tuning Language Models with (Almost) No Human Labor
Honovich, O.; Scialom, T.; Levy, O.; and Schick, T. 2022 · 2022
Cited alongside, same era.
CTRLEval: An Unsupervised Reference-Free Metric for Evaluating Controlled Text Generation
Ke, P.; Zhou, H.; Lin, Y.; Li, P.; Zhou, J.; Zhu, X.; and Huang, M. 2022 · 2022
Cited alongside, same era.
Controllable natural language generation with contrastive prefixes
Qian, J.; Dong, L.; Shen, Y.; Wei, F.; and Chen, W. 2022 · 2022
Cited alongside, same era.
Cold decoding: Energy-based constrained text generation with langevin dynamics
Qin, L.; Welleck, S.; Khashabi, D.; and Choi, Y. 2022 · 2022
Cited alongside, same era.
Self-instruct: Aligning language model with self generated instructions
Wang, Y.; Kordi, Y.; Mishra, S.; Liu, A.; Smith, N. A.; Khashabi, D.; and Hajishirzi, H. 2022 · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Wei, J.; Wang, X.; Schuurmans, D.; Bosma, M.; Xia, F.; Chi, E.; Le, Q. V.; Zhou, D.; et al. 2022 · 2022
Cited alongside, same era.
Huang, Y.; Bai, Y.; Zhu, Z.; Zhang, J.; Zhang, J.; Su, T.; Liu, J.; Lv, C.; Zhang, Y.; Lei, J.; Fu, Y.; Sun, M.; and He, J. 2023 · 2023
Later among the works it cites.
LongForm: Optimizing Instruction Tuning for Long Text Generation with Corpus Extraction
Köksal, A.; Schick, T.; Korhonen, A.; and Schütze, H. 2023 · 2023
Later among the works it cites.
Instruction-following Evaluation through Verbalizer Manipulation
Li, S.; Yan, J.; Wang, H.; Tang, Z.; Ren, X.; Srinivasan, V.; and Jin, H. 2023 · 2023
Later among the works it cites.
OpenAI. 2023 · 2023
Later among the works it cites.
Is ChatGPT a General-Purpose Natural Language Processing Task Solver?
Qin, C.; Zhang, A.; Zhang, Z.; Chen, J.; Yasunaga, M.; and Yang, D. 2023 · 2023
Later among the works it cites.
RewriteLM: An Instruction-Tuned Large Language Model for Text Rewriting
Shu, L.; Luo, L.; Hoskere, J.; Zhu, Y.; Liu, C.; Tong, S.; Chen, J.; and Meng, L. 2023 · 2023
Later among the works it cites.
Stanford Alpaca: An Instruction-following LLaMA model
Taori, R.; Gulrajani, I.; Zhang, T.; Dubois, Y.; Li, X.; Guestrin, C.; Liang, P.; and Hashimoto, T. B. 2023 · 2023
Later among the works it cites.
A survey of large language models
Zhao, W. X.; Zhou, K.; Li, J.; Tang, T.; Wang, X.; Hou, Y.; Min, Y.; Zhang, B.; Zhang, J.; Dong, Z.; et al. 2023 · 2023
Later among the works it cites.
Controlled text generation with natural language instructions
Zhou, W.; Jiang, Y. E.; Wilcox, E.; Cotterell, R.; and Sachan, M. 2023 · 2023
Later among the works it cites.