Fetching the paper…
Reading the bibliography…
In supervised learning, the question of data quality and curation has been over-shadowed in recent years by increasingly more powerful and expressive models that can ingest internet-scale data.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
A model of inductive bias learning
Jonathan Baxter · 2000
Earlier work this paper cites.
Efficient reductions for imitation learning
Stéphane Ross and Drew Bagnell · 2010
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Stéphane Ross, Geoffrey Gordon, and Drew Bagnell · 2011
Earlier work this paper cites.
Probabilistic model-based imitation learning
Peter Englert, Alexandros Paraschos, Marc Peter Deisenroth, and Jan Peters · 2013
Earlier work this paper cites.
A survey on data quality: classifying poor data
Nuno Laranjeiro, Seyma Nur Soydemir, and Jorge Bernardino · 2015
Earlier work this paper cites.
Dart: Noise injection for robust imitation learning
Michael Laskey, Jonathan Lee, Roy Fox, Anca Dragan, and Ken Goldberg · 2017
Earlier work this paper cites.
Active reward learning from critiques
Yuchen Cui and Scott Niekum · 2018
Earlier work this paper cites.
Asking easy questions: A user-friendly approach to active reward learning
Erdem Bıyık, Malayandi Palan, Nicholas C Landolfi, Dylan P Losey, and Dorsa Sadigh · 2019
Earlier work this paper cites.
An overview of data quality frameworks
Corinna Cichy and Stefan Rass · 2019
Earlier work this paper cites.
Uncertainty-aware data aggregation for deep imitation learning
Yuchen Cui, David Isele, Scott Niekum, and Kikuo Fujimura · 2019
Earlier work this paper cites.
Hg-dagger: Interactive imitation learning with human experts
Michael Kelly, Chelsea Sidrane, Katherine Driggs-Campbell, and Mykel J Kochenderfer · 2019
Earlier work this paper cites.
Teacher-aware active robot learning
Mattia Racca, Antti Oulasvirta, and Ville Kyrki · 2019
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Earlier work this paper cites.
Reward-rational (implicit) choice: A unifying formalism for reward learning
Hong Jun Jeon, Smitha Milli, and Anca D Dragan · 2020
Earlier work this paper cites.
Learning latent plans from play
Corey Lynch, Mohi Khansari, Ted Xiao, Vikash Kumar, Jonathan Tompson, Sergey Levine, and Pierre Sermanet · 2020
Earlier work this paper cites.
Human-in-the-loop imitation learning using remote teleoperation
Ajay Mandlekar, Danfei Xu, Roberto Martín-Martín, Yuke Zhu, Li Fei-Fei, and Silvio Savarese · 2020
Earlier work this paper cites.
Large language models associate muslims with violence
Abubakar Abid, Maheen Farooqi, and James Zou · 2021
Cited alongside, same era.
Learning from suboptimal demonstration via self-supervised reward regression
Letian Chen, Rohan Paleja, and Matthew Gombolay · 2021
Cited alongside, same era.
What matters in learning from offline human demonstrations for robot manipulation
Ajay Mandlekar, Danfei Xu, Josiah Wong, Soroush Nasiriany, Chen Wang, Rohun Kulkarni, Li Fei-Fei, Silvio Savarese, Yuke Zhu, and Roberto Martín-Martín · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
Feedback in imitation learning: The three regimes of covariate shift
Jonathan Spencer, Sanjiban Choudhury, Arun Venkatraman, Brian Ziebart, and J Andrew Bagnell · 2021
Learning and retrieval from prior data for skill-based imitation learning
Soroush Nasiriany, Tian Gao, Ajay Mandlekar, and Yuke Zhu · 2022
Later among the works it cites.
Efficient model finetuning for text classification via data filtering, 2022
Xu Ouyang, Shahina Mohd Azam Ansari, Felix Xiaozhu Lin, and Yangfeng Ji · 2022
Later among the works it cites.
Imitating, fast and slow: Robust learning from demonstrations via decision-time planning
Carl Qi, Pieter Abbeel, and Aditya Grover · 2022
Later among the works it cites.
Scott Reed, Konrad Zolna, Emilio Parisotto, Sergio Gomez Colmenarejo, Alexander Novikov, Gabriel Barth-Maron, Mai Gimenez, Yury Sulsky, Jackie Kay, Jost Tobias Springenberg, et al · 2022
Later among the works it cites.
Mind meld: Personalized meta-learning for robot-centric imitation learning
Mariah L Schrum, Erin Hedlund-Botti, Nina Moorman, and Matthew C Gombolay · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Imitation learning by estimating expertise of demonstrators
Mark Beliaev, Andy Shih, Stefano Ermon, Dorsa Sadigh, and Ramtin Pedarsani · 2022
Cited alongside, same era.
Plato: Predicting latent affordances through object-centric play
Suneel Belkhale and Dorsa Sadigh · 2022
Cited alongside, same era.
Learning from imperfect demonstrations via adversarial confidence transfer
Zhangjie Cao, Zihan Wang, and Dorsa Sadigh · 2022
Cited alongside, same era.
Careful data curation stabilizes in-context learning, 2022
Ting-Yun Chang and Robin Jia · 2022
Cited alongside, same era.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al · 2022
Cited alongside, same era.
A survey of data quality measurement and monitoring tools
Lisa Ehrlinger and Wolfram Wöß · 2022
Cited alongside, same era.
Data augmentation for efficient learning from parametric experts
Alexandre Galashov, Josh S Merel, and Nicolas Heess · 2022
Cited alongside, same era.
Later among the works it cites.
Perceiver-actor: A multi-task transformer for robotic manipulation
Mohit Shridhar, Lucas Manuelli, and Dieter Fox · 2022
Later among the works it cites.
Lamda: Language models for dialog applications
Romal Thoppilan, Daniel De Freitas, Jamie Hall, Noam Shazeer, Apoorv Kulshreshtha, Heng-Tze Cheng, Alicia Jin, Taylor Bos, Leslie Baker, Yu Du, et al · 2022
Later among the works it cites.
Viola: Imitation learning for vision-based manipulation with object proposal priors
Yifeng Zhu, Abhishek Joshi, Peter Stone, and Yuke Zhu · 2022
Later among the works it cites.
Easily accessible text-to-image generation amplifies demographic stereotypes at large scale
Federico Bianchi, Pratyusha Kalluri, Esin Durmus, Faisal Ladhak, Myra Cheng, Debora Nozza, Tatsunori Hashimoto, Dan Jurafsky, James Zou, and Aylin Caliskan · 2023
Closest in time.
Diffusion policy: Visuomotor policy learning via action diffusion
Cheng Chi, Siyuan Feng, Yilun Du, Zhenjia Xu, Eric Cousineau, Benjamin Burchfiel, and Shuran Song · 2023
Closest in time.
From play to policy: Conditional behavior generation from uncurated robot data
Zichen Jeff Cui, Yibin Wang, Nur Muhammad Mahi Shafiullah, and Lerrel Pinto · 2023
Closest in time.
Behavior retrieval: Few-shot imitation learning by querying unlabeled datasets
Maximilian Du, Suraj Nair, Dorsa Sadigh, and Chelsea Finn · 2023
Closest in time.
Starcoder
Hugging Face · 2023
Closest in time.
Instruction-driven history-aware policies for robotic manipulations
Pierre-Louis Guhur, Shizhe Chen, Ricardo Garcia Pinel, Makarand Tapaswi, Ivan Laptev, and Cordelia Schmid · 2023
Closest in time.
Language-driven representation learning for robotics
Siddharth Karamcheti, Suraj Nair, Annie S Chen, Thomas Kollar, Chelsea Finn, Dorsa Sadigh, and Percy Liang · 2023
Closest in time.
Learning fine-grained bimanual manipulation with low-cost hardware
Tony Z. Zhao, Vikash Kumar, Sergey Levine, and Chelsea Finn · 2023
Closest in time.