Roberta: A robustly optimized BERT pretraining approach
Original
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 1907
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
Richard S. Sutton, David A. McAllester, Satinder P. Singh, and Yishay Mansour · 1999
Earlier work this paper cites.
Helping skills: Facilitating, exploration, insight, and action
Clara E Hill · 2009
Earlier work this paper cites.
Decoupling strategy and generation in negotiation dialogues
He He, Derek Chen, Anusha Balakrishnan, and Percy Liang · 2018
Earlier work this paper cites.
A dynamic strategy coach for effective negotiation
Yiheng Zhou, He He, Alan W. Black, and Yulia Tsvetkov · 2019
Earlier work this paper cites.
CIMA: A large open access dialogue dataset for tutoring
Katherine Stasaski, Kimberly Kao, and Marti A. Hearst · 2020
Earlier work this paper cites.
DIALOGPT : Large-scale generative pre-training for conversational response generation
Yizhe Zhang, Siqi Sun, Michel Galley, Yen-Chun Chen, Chris Brockett, Xiang Gao, Jianfeng Gao, Jingjing Liu, and Bill Dolan · 2020
Earlier work this paper cites.
Augmenting non-collaborative dialog systems with explicit semantic and strategic dialog history
Yiheng Zhou, Yulia Tsvetkov, Alan W. Black, and Zhou Yu · 2020
Earlier work this paper cites.
Unified conversational recommendation policy learning via graph-based reinforcement learning
Yang Deng, Yaliang Li, Fei Sun, Bolin Ding, and Wai Lam · 2021
Earlier work this paper cites.
Advances and challenges in conversational recommender systems: A survey
Chongming Gao, Wenqiang Lei, Xiangnan He, Maarten de Rijke, and Tat-Seng Chua · 2021
Earlier work this paper cites.
Dialograph: Incorporating interpretable strategy-graph networks into negotiation dialogues
Rishabh Joshi, Vidhisha Balachandran, Shikhar Vashishth, Alan W. Black, and Yulia Tsvetkov · 2021
Earlier work this paper cites.
Towards emotional support dialog systems
Siyang Liu, Chujie Zheng, Orianna Demasi, Sahand Sabour, Yu Li, Zhou Yu, Yong Jiang, and Minlie Huang · 2021
Earlier work this paper cites.
Improving dialog systems for negotiation with personality modeling
Runzhe Yang, Jingxiao Chen, and Karthik Narasimhan · 2021
Earlier work this paper cites.
Constitutional AI: harmlessness from AI feedback
Original
Yuntao Bai, Saurav Kadavath, Sandipan Kundu, Amanda Askell, Jackson Kernion, Andy Jones, Anna Chen, Anna Goldie, Azalia Mirhoseini, Cameron McKinnon, Carol Chen, Catherine Olsson, Christopher Olah, Danny Hernandez, Dawn Drain, Deep Ganguli, Dustin Li, Eli Tran-Johnson, Ethan Perez, Jamie Kerr, Jared Mueller, Jeffrey Ladish, Joshua Landau, Kamal Ndousse, Kamile Lukosiute, Liane Lovitt, Michael Sellitto, Nelson Elhage, Nicholas Schiefer, Noemí Mercado, Nova DasSarma, Robert Lasenby, Robin Larson, Sam Ringer, Scott Johnston, Shauna Kravec, Sheer El Showk, Stanislav Fort, Tamera Lanham, Timothy Telleen-Lawton, Tom Conerly, Tom Henighan, Tristan Hume, Samuel R. Bowman, Zac Hatfield-Dodds, Ben Mann, Dario Amodei, Nicholas Joseph, Sam McCandlish, Tom Brown, and Jared Kaplan · 2022
Earlier work this paper cites.
Improving multi-turn emotional support dialogue generation with lookahead strategy planning
Original
Yi Cheng, Wenge Liu, Wenjie Li, Jiashuo Wang, Ruihui Zhao, Bang Liu, Xiaodan Liang, and Yefeng Zheng · 2022
Earlier work this paper cites.
PACIFIC: towards proactive conversational question answering over tabular and textual data in finance
Yang Deng, Wenqiang Lei, Wenxuan Zhang, Wai Lam, and Tat-Seng Chua · 2022
Earlier work this paper cites.
Gpt-critic: Offline reinforcement learning for end-to-end task-oriented dialogue systems
Youngsoo Jang, Jongmin Lee, and Kee-Eung Kim · 2022
Earlier work this paper cites.