Fetching the paper…
Reading the bibliography…
Machine learning (ML) systems are rapidly increasing in size, are acquiring new capabilities, and are increasingly deployed in high-stakes settings.
“The Methods of Ethics”, 1907
Henry Sidgwick · 1907
Earlier work this paper cites.
“Map-Colour Theorem”
P.. Heawood · 1949
Earlier work this paper cites.
“Dysfunctional Consequences of Performance Measurements”
V. Ridgway · 1956
Earlier work this paper cites.
“Hedonic relativism and planning the good society”, 1971
Philip Brickman and Donald Campbell · 1971
Earlier work this paper cites.
“Theory of the firm: Managerial behavior, agency costs and ownership structure”
Michael Jensen and William Meckling · 1976
Earlier work this paper cites.
“Systemantics: How Systems Work and Especially How They Fail”, 1977
John Gall · 1977
Earlier work this paper cites.
“Utilitarianism, Economics, and Legal Theory”
Richard. Posner · 1979
Earlier work this paper cites.
“Integrative Complexity Theory and Forecasting International Crises: Berlin 1946-1962”
Theodore. Raphael · 1982
Earlier work this paper cites.
“System Safety in Aircraft Acquisition”, 1984
F.. Frola and C.. Miller · 1984
Earlier work this paper cites.
“Problems of Monetary Management: The UK Experience”, 1984
Charles Goodhart · 1984
Earlier work this paper cites.
“Smoking behavior of adolescents exposed to cigarette advertising”
G. Botvin, C. Goldberg, E.. Botvin and L. Dusenbury · 1993
Earlier work this paper cites.
“Programming Satan’s Computer”
Ross. Anderson and Roger Needham · 1995
Earlier work this paper cites.
“Failure mode and effect analysis : FMEA from theory to execution”
D.. Stamatis · 1996
Earlier work this paper cites.
“The third industrial revolution: Technology, productivity, and income inequality”
Jeremy Greenwood · 1997
Earlier work this paper cites.
“An application of machine learning to anomaly detection”
Terran Lane and Carla Brodley · 1997
Earlier work this paper cites.
“‘Improving ratings’: audit in the British University system”
Marilyn Strathern · 1997
Earlier work this paper cites.
“Planning for seismic rehabilitation: societal issues”, 1998
Building Seismic Safety US · 1998
Earlier work this paper cites.
“A Theory of Justice”
John Rawls · 1999
Earlier work this paper cites.
“Risky business: safety regulations, risk compensation, and individual behavior”
James Hedlund · 2000
Earlier work this paper cites.
“Quadrennial Defense Review Report”, 2001
Department of Defense · 2001
Earlier work this paper cites.
“What drives tropical deforestation?: a meta-analysis of proximate and underlying causes of deforestation based on subnational case study evidence”, 2001
Helmut Geist and Eric Lambin · 2001
Earlier work this paper cites.
“Inequality and Violent Crime”
Pablo Fajnzylber, Daniel Lederman and Norman. Loayza · 2002
Earlier work this paper cites.
“A Brief History of Generative Models for Power Law and Lognormal Distributions”
Michael Mitzenmacher · 2003
Earlier work this paper cites.
“CAPABILITIES AS FUNDAMENTAL ENTITLEMENTS: SEN AND SOCIAL JUSTICE”
Martha Nussbaum · 2003
Earlier work this paper cites.
“The Misbehavior of Markets: A Fractal View of Risk, Ruin, and Reward”, 2004
Benoit Mandelbrot and Richard. Hudson · 2004
Earlier work this paper cites.
“Affective Forecasting”
Timothy Wilson and Daniel Gilbert · 2005
Earlier work this paper cites.
“A history of internet security”
Laura DeNardis · 2007
Earlier work this paper cites.
“The Black Swan: The Impact of the Highly Improbable”, 2007
Nassim Taleb · 2007
Earlier work this paper cites.
“Analysis of the 2007 Cyber Attacks Against Estonia from the Information Warfare Perspective”, 2008
Rain Ottis · 2008
Earlier work this paper cites.
“What the GDP Gets Wrong (Why Managers Should Care)”
Erik Brynjolfsson and Adam Saunders · 2009
Earlier work this paper cites.
“High income improves evaluation of life but not emotional well-being”
Daniel Kahneman and Angus Deaton · 2010
Earlier work this paper cites.
“Outside the closed world: On using machine learning for network intrusion detection”
Robin Sommer and Vern Paxson · 2010
Earlier work this paper cites.
“Chromium-6 in US tap water”
Rebecca Sutton · 2010
Earlier work this paper cites.
“The Flash Crash: The Impact of High Frequency Trading on an Electronic Market”, 2011
A. Kirilenko, Mehrdad Samadi, A. Kyle and Tugkan Tuzun · 2011
Earlier work this paper cites.
“Stuxnet: Dissecting a Cyberwarfare Weapon”
Ralph Langner · 2011
Earlier work this paper cites.
“Monitor alarm fatigue: an integrative review”
Maria Cvach · 2012
Earlier work this paper cites.
“Engineering a Safer World: Systems Thinking Applied to Safety”, 2012
Nancy Leveson · 2012
Earlier work this paper cites.
“Antifragile: Things That Gain from Disorder”, 2012
Nassim Taleb · 2012
Earlier work this paper cites.
“Evasion attacks against machine learning at test time”
Battista Biggio, Igino Corona, Davide Maiorca, Blaine Nelson, Nedim Srndi\’c, Pavel Laskov, Giorgio Giacinto and Fabio Roli · 2013
Earlier work this paper cites.
“Intriguing properties of neural networks”
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow and Rob Fergus · 2013
Earlier work this paper cites.
“On the Difference between Binary Prediction and True Exposure with Implications for Forecasting Tournaments and Decision Making Research”, 2013
Nassim Taleb and Philip Tetlock · 2013
Earlier work this paper cites.
“The Point of View of the Universe: Sidgwick and Contemporary Ethics”, 2014
Katarzyna de Lazari-Radek and Peter Singer · 2014
Earlier work this paper cites.
“Autonomous Weapons: An Open Letter from AI and Robotics Researchers”, 2015
Signed by 30000+ · 2015
Earlier work this paper cites.
“Posterior calibration and exploratory analysis for natural language processing models”
Khanh Nguyen and Brendan. O’Connor · 2015
Earlier work this paper cites.
In Managing the Unexpected
“Principle 1: Preoccupation with Failure” · 2015
Earlier work this paper cites.
“Research Priorities for Robust and Beneficial Artificial Intelligence”
Stuart. Russell, Daniel Dewey and Max Tegmark · 2015
Earlier work this paper cites.
“Hidden technical debt in machine learning systems”
David Sculley, Gary Holt, Daniel Golovin, Eugene Davydov, Todd Phillips, Dietmar Ebner, Vinay Chaudhary, Michael Young, Jean-Francois Crespo and Dan Dennison · 2015
Earlier work this paper cites.
“Recognizing Functions in Binaries with Neural Networks”
E.. Shin, D. Song and R. Moazzezi · 2015
Earlier work this paper cites.
“Superforecasting: The Art and Science of Prediction”, 2015
Philip Tetlock and Dan Gardner · 2015
Earlier work this paper cites.
“The Possibility of an Ongoing Moral Catastrophe”
E.. Williams · 2015
Earlier work this paper cites.
“Deep Learning with Differential Privacy”
Mart\’in Abadi, Andy Chu, I. Goodfellow, H.. McMahan, Ilya Mironov, Kunal Talwar and L. Zhang · 2016
Earlier work this paper cites.
“Concrete Problems in AI Safety”
Dario Amodei, Christopher Olah, Jacob Steinhardt, Paul Christiano, John Schulman and Dandelion Man\’e · 2016
Earlier work this paper cites.
“Towards Open Set Deep Networks”
Abhijit Bendale and Terrance Boult · 2016
Earlier work this paper cites.
“Faulty Reward Functions in the Wild”
Jack Clark and Dario Amodei · 2016
Earlier work this paper cites.
“Cooperative Inverse Reinforcement Learning”
Dylan Hadfield-Menell, Stuart. Russell, P. Abbeel and A. Dragan · 2016
Earlier work this paper cites.
“Equality of Opportunity in Supervised Learning”
Moritz Hardt, Eric Price and Nathan Srebro · 2016
Earlier work this paper cites.
URL: https://blogs.microsoft.com/blog/2016/03/25/learning-tays-introductioverbn/
Microsoft · 2016
Earlier work this paper cites.
“Accessorize to a crime: Real and stealthy attacks on state-of-the-art face recognition”
Mahmood Sharif, Sruti Bhagavatula, Lujo Bauer and Michael Reiter · 2016
Earlier work this paper cites.
“The Rationality Quotient: Toward a Test of Rational Thinking”, 2016
Keith. Stanovich, Richard. West and Maggie. Toplak · 2016
Earlier work this paper cites.
“Alignment for Advanced Machine Learning Systems”, 2016
Jessica Taylor, Eliezer Yudkowsky, Patrick LaVictoire and Andrew Critch · 2016
Earlier work this paper cites.
“Asilomar AI Principles”, 2017
Signed by 2000 AI · 2017
Earlier work this paper cites.
“Decision-based adversarial attacks: Reliable attacks against black-box machine learning models”
Wieland Brendel, Jonas Rauber and Matthias Bethge · 2017
Earlier work this paper cites.
“Towards evaluating the robustness of neural networks”
Nicholas Carlini and David Wagner · 2017
Earlier work this paper cites.
“Badnets: Identifying vulnerabilities in the machine learning model supply chain”
Tianyu Gu, Brendan Dolan-Gavitt and Siddharth Garg · 2017
Earlier work this paper cites.
“On Calibration of Modern Neural Networks”
Chuan Guo, Geoff Pleiss, Yu Sun and Kilian. Weinberger · 2017
Earlier work this paper cites.
“The Off-Switch Game”
Dylan Hadfield-Menell, A. Dragan, P. Abbeel and Stuart. Russell · 2017
Earlier work this paper cites.
“A Baseline for Detecting Misclassified and Out-of-Distribution Examples in Neural Networks”
Dan Hendrycks and Kevin Gimpel · 2017
Earlier work this paper cites.
“Deep Learning Scaling is Predictable, Empirically”
J. Hestness, Sharan Narang, Newsha Ardalani, G. Diamos, Heewoo Jun, Hassan Kianinejad, Md. Mostofa Patwary, Y. Yang and Yanqi Zhou · 2017
Earlier work this paper cites.
“Simple and Scalable Predictive Uncertainty Estimation using Deep Ensembles”
Balaji Lakshminarayanan, A. Pritzel and C. Blundell · 2017
Earlier work this paper cites.
“CheXNet: Radiologist-Level Pneumonia Detection on Chest X-Rays with Deep Learning”
Pranav Rajpurkar, Jeremy. Irvin, Kaylie Zhu, Brandon Yang, Hershel Mehta, T. Duan, D. Ding, Aarti Bagul, C. Langlotz, K. Shpanskaya, M. Lungren and A. Ng · 2017
Cited alongside, same era.
“Membership inference attacks against machine learning models”
Reza Shokri, Marco Stronati, Congzheng Song and Vitaly Shmatikov · 2017
Cited alongside, same era.
“Obfuscated Gradients Give a False Sense of Security: Circumventing Defenses to Adversarial Examples”
Anish Athalye, Nicholas Carlini and David. Wagner · 2018
Cited alongside, same era.
“The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation”
Miles Brundage, Shahar Avin, Jack Clark, H. Toner, P. Eckersley, Ben Garfinkel, A. Dafoe, P. Scharre, T. Zeitzoff, Bobby Filar, H. Anderson, Heather Roff, Gregory. Allen, J. Steinhardt, Carrick Flynn, Se\’an\’O h\’Eigeartaigh, S. Beard, Haydn Belfield, Sebastian Farquhar, Clare Lyle, Rebecca Crootof, Owain Evans, Michael Page, Joanna Bryson, Roman Yampolskiy and Dario Amodei · 2018
Cited alongside, same era.
“AI governance: a research agenda”
“It Takes Two to Lie: One to Lie, and One to Listen”
Denis Peskov, Benny Cheng, Ahmed Elgohary, Joe Barrow, Cristian Danescu-Niculescu-Mizil and Jordan. Boyd-Graber · 2020
Later among the works it cites.
“Intriguing properties of adversarial ml attacks in the problem space”
Fabio Pierazzi, Feargus Pendlebury, Jacopo Cortellazzi and Lorenzo Cavallaro · 2020
Later among the works it cites.
“Aligning AI Optimization to Community Well-Being”
Jonathan Stray · 2020
Later among the works it cites.
“Confidence-Calibrated Adversarial Training: Generalizing to Unseen Attacks”
David Stutz, Matthias Hein and B. Schiele · 2020
Later among the works it cites.
“CSI: Novelty Detection via Contrastive Learning on Distributionally Shifted Instances”
Jihoon Tack, Sangwoo Mo, Jongheon Jeong and Jinwoo Shin · 2020
Later among the works it cites.
“Statistical Consequences of Fat Tails: Real World Preasymptotics, Epistemology, and Applications”, 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Allan Dafoe · 2018
Cited alongside, same era.
“Robust artificial intelligence and robust human organizations”
Thomas. Dietterich · 2018
Cited alongside, same era.
“A rotation and a translation suffice: Fooling cnns with simple transformations”
Logan Engstrom, Brandon Tran, Dimitris Tsipras, Ludwig Schmidt and Aleksander Madry · 2018
Cited alongside, same era.
“Bringing People Closer Together”
Facebook · 2018
Cited alongside, same era.
“Motivating the Rules of the Game for Adversarial Example Research”
J. Gilmer, Ryan. Adams, I. Goodfellow, David. Andersen and George. Dahl · 2018
Cited alongside, same era.
“Explaining Explanations: An Overview of Interpretability of Machine Learning”, 2018
Leilani. Gilpin, David Bau, Ben. Yuan, Ayesha Bajwa, Michael. Specter and Lalana Kagal · 2018
Cited alongside, same era.
“AI safety via debate”
Geoffrey Irving, Paul Christiano and Dario Amodei · 2018
Cited alongside, same era.
“Accurate uncertainties for deep learning using calibrated regression”
Volodymyr Kuleshov, Nathan Fenner and Stefano Ermon · 2018
Cited alongside, same era.
Nassim Taleb · 2020
Later among the works it cites.
“On Adaptive Attacks to Adversarial Example Defenses”
Florian Tram\‘er, Nicholas Carlini, Wieland Brendel and A. Madry · 2020
Later among the works it cites.
“Avoiding Side Effects in Complex Environments”
A.. Turner, Neale Ratzlaff and Prasad Tadepalli · 2020
Later among the works it cites.
“SafeLife 1.0: Exploring Side Effects in Complex Environments”
Carroll. Wainwright and P. Eckersley · 2020
Later among the works it cites.
“Stop-and-Go: Exploring Backdoor Attacks on Deep Reinforcement Learning-based Traffic Congestion Control Systems”
Yue Wang, Esha Sarkar, Wenqing Li, M. Maniatakos and S.. Jabari · 2020
Later among the works it cites.
“Adversarial Weight Perturbation Helps Robust Generalization”
Dongxian Wu, Shutao Xia and Yisen Wang · 2020
Later among the works it cites.
Cihang Xie, Mingxing Tan, Boqing Gong, A. Yuille and Quoc. Le · 2020
Later among the works it cites.
“Dreaming to Distill: Data-Free Knowledge Transfer via DeepInversion”
Hongxu Yin, Pavlo Molchanov, Zhizhong Li, J. \’Alvarez, Arun Mallya, Derek Hoiem, N. Jha and J. Kautz · 2020
Later among the works it cites.
“Neural Ensemble Search for Uncertainty Estimation and Dataset Shift”, 2020
Sheheryar Zaidi, Arber Zela, T. Elsken, Chris. Holmes, F. Hutter and Y. Teh · 2020
Later among the works it cites.
“Trojaning Language Models for Fun and Profit”
Xinyang Zhang, Zheng Zhang and Tianying Wang · 2020
Later among the works it cites.
“Network intrusion detection system: A systematic study of machine learning and deep learning approaches”
Zeeshan Ahmad, A. Khan, W. Cheah, J. Abdullah and Farhan Ahmad · 2021
Closest in time.
“Program Synthesis with Large Language Models”
Jacob Austin, Augustus Odena, Maxwell Nye, Maarten Bosma, Henryk Michalewski, David Dohan, Ellen Jiang, Carrie Cai, Michael Terry, Quoc. Le and Charles Sutton · 2021
Closest in time.
“Blind Backdoors in Deep Learning Models”
Eugene Bagdasaryan and Vitaly Shmatikov · 2021
Closest in time.
“On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?”
Emily. Bender, Timnit Gebru, Angelina McMillan-Major and Shmargaret Shmitchell · 2021
Closest in time.
Victor Besnier, Andrei Bursuc, David Picard and Alexandre Briot · 2021
Closest in time.
“The Values Encoded in Machine Learning Research”
Abeba Birhane, Pratyusha Kalluri, D. Card, William Agnew, Ravit Dotan and Michelle Bao · 2021
Closest in time.
“On the Opportunities and Risks of Foundation Models”
Rishi Bommasani et al · 2021
Closest in time.
“Automating Cyber Attacks”, 2021
Ben Buchanan, John Bansemer, Dakota Cary, Jack Lucas and Micah Musser · 2021
Closest in time.
“Truth, Lies, and Automation”, 2021
Ben Buchanan, Andrew Lohn, Micah Musser and Katerina Sedova · 2021
Closest in time.
“Poisoning and Backdooring Contrastive Learning”
Nicholas Carlini and A. Terzis · 2021
Closest in time.
“Emerging Properties in Self-Supervised Vision Transformers”
Mathilde Caron, Hugo Touvron, Ishan Misra, Herv\’e J\’egou, Julien Mairal, Piotr Bojanowski and Armand Joulin · 2021
Closest in time.
“Evaluating Large Language Models Trained on Code”
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde, J. Kaplan, Harrison Edwards, Yura Burda, Nicholas Joseph, Greg Brockman, Alex Ray, Raul Puri, Gretchen Krueger, Michael Petrov, Heidy Khlaaf, Girish Sastry, Pamela Mishkin, Brooke Chan, Scott Gray, Nick Ryder, Mikhail Pavlov, Alethea. Power, Lukasz Kaiser, Mohammad Bavarian, Clemens Winter, Philippe Tillet, F. Such, D. Cummings, Matthias Plappert, Fotios Chantzis, Elizabeth Barnes, Ariel Herbert-Voss, William. Guss, Alex Nichol, I. Babuschkin, S. Balaji, Shantanu Jain, A. Carr, J. Leike, Joshua Achiam, Vedant Misra, Evan Morikawa, Alec Radford, M. Knight, Miles Brundage, Mira Murati, Katie Mayer, P. Welinder, Bob McGrew, Dario Amodei, Sam McCandlish, Ilya Sutskever and Wojciech Zaremba · 2021
Closest in time.
URL: https://www.af.mil/News/Article-Display/Article/2703548/norad-usnorthverbcom-lead-3rd-global-information-dominance-experiment/
North American Aerospace Command and U.S. Northern Command Affairs, 2021 · 2021
Closest in time.
“Out-of-Distribution Dynamics Detection: RL-Relevant Benchmarks and Results”
Mohamad. Danesh and Alan Fern · 2021
Closest in time.
“Reinforcement Learning Under Moral Uncertainty”
Adrien Ecoffet and Joel Lehman · 2021
Closest in time.
“Measuring and Improving Consistency in Pretrained Language Models”
Yanai Elazar, Nora Kassner, Shauli Ravfogel, Abhilasha Ravichander, E. Hovy, Hinrich Sch\"utze and Yoav Goldberg · 2021
Closest in time.
“Assessment results regarding Organization Designation Authorization (ODA) Unit Member (UM) Independence”
Wendi Folkert · 2021
Closest in time.
“Augmenting Decision Making via Interactive What-If Analysis”, 2021
Sneha Gathani, Madelon Hulsebos, James Gale, P. Haas and cCaugatay Demiralp · 2021
Closest in time.
Nathan Grinsztajn, Johan Ferret, O. Pietquin, P. Preux and M. Geist · 2021
Closest in time.
“The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution Generalization”
Dan Hendrycks, Steven Basart, Norman Mu, Saurav Kadavath, Frank Wang, Evan Dorundo, Rahul Desai, Tyler Zhu, Samyak Parajuli, Mike Guo, Dawn Song, Jacob Steinhardt and Justin Gilmer · 2021
Closest in time.
“Aligning AI With Shared Human Values”
Dan Hendrycks, Collin Burns, Steven Basart, Andrew Critch, Jerry Li, Dawn Song and Jacob Steinhardt · 2021
Closest in time.
“Measuring Massive Multitask Language Understanding”
Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song and Jacob Steinhardt · 2021
Closest in time.
“What Would Jiminy Cricket Do? Towards Agents That Behave Morally”
Dan Hendrycks, Mantas Mazeika, Andy Zou, Sahil Patel, Christine Zhu, Jesus Navarro, Dawn Song, Bo Li and Jacob Steinhardt · 2021
Closest in time.
“Natural Adversarial Examples”
Dan Hendrycks, Kevin Zhao, Steven Basart, J. Steinhardt and D. Song · 2021
Closest in time.
“PixMix: Dreamlike Pictures Comprehensively Improve Safety Measures”
Dan Hendrycks, Andy Zou, Mantas Mazeika, Leonard Tang, Bo Li, Dawn Song and Jacob Steinhardt · 2021
Closest in time.
“Handcrafted Backdoors in Deep Neural Networks”
Sanghyun Hong, Nicholas Carlini and A. Kurakin · 2021
Closest in time.
“ForecastQA: A Question Answering Challenge for Event Forecasting with Temporal Text Data”
Woojeong Jin, Suji Kim, Rahul Khanna, Dong-Ho Lee, Fred Morstatter, A. Galstyan and Xiang Ren · 2021
Closest in time.
“Alignment of Language Agents”
Zachary Kenton, Tom Everitt, Laura Weidinger, Iason Gabriel, Vladimir Mikulik and Geoffrey Irving · 2021
Closest in time.
“Objective Robustness in Deep Reinforcement Learning”
Jack Koch, L. Langosco, J. Pfau, James Le and Lee Sharkey · 2021
Closest in time.
“WILDS: A Benchmark of in-the-Wild Distribution Shifts”
P.. Koh, Shiori Sagawa, H. Marklund, Sang Xie, Marvin Zhang, A. Balsubramani, Wei hua Hu, Michihiro Yasunaga, Richard. Phillips, Sara Beery, J. Leskovec, A. Kundaje, E. Pierson, Sergey Levine, Chelsea Finn and Percy Liang · 2021
Closest in time.
“Boeing 737 MAX Return to Service Report”, 2021
Patrick Ky · 2021
Closest in time.
“Perceptual Adversarial Robustness: Defense Against Unseen Threat Models”
Cassidy Laidlaw, Sahil Singla and S. Feizi · 2021
Closest in time.
“TruthfulQA: Measuring How Models Mimic Human Falsehoods”
Stephanie Lin, Jacob Hilton and Owain Evans · 2021
Closest in time.
“Localized Calibration: Metrics and Recalibration”
Rachel Luo, Aadyot Bhatnagar, Huan Wang, Caiming Xiong, Silvio Savarese, Yu Bai, Shengjia Zhao and Stefano Ermon · 2021
Closest in time.
“Test-Time Adaptation to Distribution Shift by Confidence Maximization and Input Transformation”
Chaithanya Mummadi, Robin Hutmacher, K. Rambach, Evgeny Levinkov, T. Brox and J.. Metzen · 2021
Closest in time.
“The Parliamentary Approach to Moral Uncertainty”, 2021
Toby Newberry and Toby Ord · 2021
Closest in time.
“An Empirical Cybersecurity Evaluation of GitHub Copilot’s Code Contributions”
Hammond Pearce, Baleegh Ahmad, Benjamin Tan, Brendan Dolan-Gavitt and Ramesh Karri · 2021
Closest in time.
“Robustness and Generalization via Generative Adversarial Training”, 2021
Omid Poursaeed, Tianxing Jiang, Harry Yang, Serge Belongie and Ser-Nam Lim · 2021
Closest in time.
“Grokking: Generalization Beyond Overfitting on Small Algorithmic Datasets”
Alethea Power, Yuri Burda, Harri Edwards, Igor Babuschkin and Vedant Misra · 2021
Closest in time.
“Learning Transferable Visual Models From Natural Language Supervision”
Alec Radford, Jong Kim, Chris Hallacy, A. Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger and Ilya Sutskever · 2021
Closest in time.
“Fixing Data Augmentation to Improve Adversarial Robustness”
Sylvestre-Alvise Rebuffi, Sven Gowal, D.. Calian, Florian Stimberg, Olivia Wiles and Timothy. Mann · 2021
Closest in time.
“Lethal Autonomous Weapons Exist; They Must Be Banned”, 2021
Stuart Russell, Anthony Aguirre, Emilia Javorsky and Max Tegmark · 2021
Closest in time.
“You Autocomplete Me: Poisoning Vulnerabilities in Neural Code Completion”
Roei Schuster, Congzheng Song, Eran Tromer and Vitaly Shmatikov · 2021
Closest in time.
“What are you optimizing for? Aligning Recommender Systems with Human Values”
Jonathan Stray, Ivan Vendrov, Jeremy Nixon, Steven Adler and Dylan Hadfield-Menell · 2021
Closest in time.
“Consistency Regularization for Adversarial Robustness”
Jihoon Tack, Sihyun Yu, Jongheon Jeong, Minseong Kim, Sung Hwang and Jinwoo Shin · 2021
Closest in time.
“Tesla AI Day”, 2021
Tesla · 2021
Closest in time.
“Conservative Objective Models for Effective Offline Model-Based Optimization”
Brandon Trabucco, Aviral Kumar, Xinyang Geng and Sergey Levine · 2021
Closest in time.
“Optimal Policies Tend To Seek Power”
Alexander Turner, Logan Smith, Rohin Shah, Andrew Critch and Prasad Tadepalli · 2021
Closest in time.
“Concealed Data Poisoning Attacks on NLP Models”
Eric Wallace, Tony Zhao, Shi Feng and Sameer Singh · 2021
Closest in time.
“Fighting Gradients with Gradients: Dynamic Defenses against Adversarial Attacks”
Dequan Wang, An Ju, Evan Shelhamer, David. Wagner and Trevor Darrell · 2021
Closest in time.
“Tent: Fully Test-Time Adaptation by Entropy Minimization”
Dequan Wang, Evan Shelhamer, Shaoteng Liu, B. Olshausen and Trevor Darrell · 2021
Closest in time.
“IMAGINE: Image Synthesis by Image-Guided Model Inversion”
Pei Wang, Yijun Li, Krishna Singh, Jingwan Lu and N. Vasconcelos · 2021
Closest in time.
“Towards Understanding the Generative Capability of Adversarially Robust Classifiers”
Yao Zhu, Jiacheng Ma, Jiacheng Sun, Zewei Chen, Rongxin Jiang and Zhenguo Li · 2021
Closest in time.