Fetching the paper…
Reading the bibliography…
Machine learning applications frequently come with multiple diverse objectives and constraints that can change over time.
Teoria statistica delle classi e calcolo delle probabilita
Carlo Bonferroni · 1936
Earlier work this paper cites.
Pareto optimality in multiobjective problems
Yair Censor · 1977
Earlier work this paper cites.
A simple sequentially rejective multiple test procedure
Sture Holm · 1979
Earlier work this paper cites.
Multi-Objective Optimization using Evolutionary Algorithms
Kalyanmoy Deb · 2001
Earlier work this paper cites.
Controlling computation versus quality for neural sequence models
Ankur Bapna, Naveen Arivazhagan, and Orhan Firat · 2002
Earlier work this paper cites.
Inductive confidence machines for regression
Harris Papadopoulos, Kostas Proedrou, Volodya Vovk, and Alex Gammerman · 2002
Earlier work this paper cites.
On-line confidence machines are well-calibrated
Vladimir Vovk · 2002
Earlier work this paper cites.
Introduction to optimum design
Jasbir Arora · 2004
Earlier work this paper cites.
Algorithmic learning in a random world
Vladimir Vovk, Alexander Gammerman, and Glenn Shafer · 2005
Earlier work this paper cites.
Parego: A hybrid algorithm with on-line landscape approximation for expensive multiobjective optimization problems
Joshua Knowles · 2006
Earlier work this paper cites.
Pareto-based multiobjective machine learning: An overview and case studies
Yaochu Jin and Bernhard Sendhoff · 2008
Earlier work this paper cites.
A graphical approach to sequentially rejective multiple test procedures
Frank Bretz, Willi Maurer, Werner Brannath, and Martin Posch · 2009
Earlier work this paper cites.
Learning word vectors for sentiment analysis
Andrew Maas, Raymond E Daly, Peter T Pham, Dan Huang, Andrew Y Ng, and Christopher Potts · 2011
Earlier work this paper cites.
Nonlinear multiobjective optimization , volume 12
Kaisa Miettinen · 2012
Earlier work this paper cites.
Distribution-free prediction sets
Jing Lei, James Robins, and Larry Wasserman · 2013
Earlier work this paper cites.
Surrogate-based multiobjective optimization: Parego update and test
Cristina Cristescu and Joshua Knowles · 2015
Earlier work this paper cites.
Large-scale probabilistic predictors with and without guarantees of validity
Vladimir Vovk, Ivan Petej, and Valentina Fedorova · 2015
Earlier work this paper cites.
Character-level convolutional networks for text classification
Xiang Zhang, Junbo Zhao, and Yann LeCun · 2015
Earlier work this paper cites.
Adaptive computation time for recurrent neural networks
Alex Graves · 2016
Earlier work this paper cites.
Squad: 100,000+ questions for machine comprehension of text
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang · 2016
Earlier work this paper cites.
Pareto frontier learning with expensive correlated objectives
Amar Shah and Zoubin Ghahramani · 2016
Earlier work this paper cites.
Branchynet: Fast inference via early exiting from deep neural networks
Surat Teerapittayanon, Bradley McDanel, and Hsiang-Tsung Kung · 2016
Earlier work this paper cites.
Criteria of efficiency for conformal prediction
Vladimir Vovk, Valentina Fedorova, Ilia Nouretdinov, and Alexander Gammerman · 2016
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Nonparametric predictive distributions based on conformal prediction
Vladimir Vovk, Jieli Shen, Valery Manokhin, and Min-ge Xie · 2017
Cited alongside, same era.
Efficient multiobjective optimization employing gaussian processes, spectral sampling and a genetic algorithm
Eric Bradford, Artur M Schweidtmann, and Alexei Lapkin · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Cited alongside, same era.
Dpp-net: Device-aware progressive search for pareto-optimal neural architectures
Jin-Dong Dong, An-Chieh Cheng, Da-Cheng Juan, Wei Wei, and Min Sun · 2018
Cited alongside, same era.
Multi-objective multi-fidelity hyperparameter optimization with application to fairness
Robin Schmucker, Michele Donini, Valerio Perrone, Muhammad Bilal Zafar, and Cédric Archambeau · 2020
Later among the works it cites.
Green ai
Roy Schwartz, Jesse Dodge, Noah A. Smith, and Oren Etzioni · 2020
Later among the works it cites.
Mobilebert: a compact task-agnostic bert for resource-limited devices
Zhiqing Sun, Hongkun Yu, Xiaodan Song, Renjie Liu, Yiming Yang, and Denny Zhou · 2020
Later among the works it cites.
DeeBERT: Dynamic early exiting for accelerating BERT inference
Ji Xin, Raphael Tang, Jaejun Lee, Yaoliang Yu, and Jimmy Lin · 2020
Later among the works it cites.
Learn then test: Calibrating predictive algorithms to achieve risk control
Anastasios N Angelopoulos, Stephen Bates, Emmanuel J Candès, Michael I Jordan, and Lihua Lei · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Thomas Elsken, Jan Hendrik Metzen, and Frank Hutter · 2018
Cited alongside, same era.
Distribution-free predictive inference for regression
Jing Lei, Max G’Sell, Alessandro Rinaldo, Ryan J Tibshirani, and Larry Wasserman · 2018
Cited alongside, same era.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel R Bowman · 2018
Cited alongside, same era.
Resource-efficient neural architect
Yanqi Zhou, Siavash Ebrahimi, Sercan Ö Arık, Haonan Yu, Hairong Liu, and Greg Diamos · 2018
Cited alongside, same era.
Max-value entropy search for multi-objective bayesian optimization
Syrine Belakaria, Aryan Deshwal, and Janardhan Rao Doppa · 2019
Cited alongside, same era.
Neural architecture search: A survey
Thomas Elsken, Jan Hendrik Metzen, and Frank Hutter · 2019
Cited alongside, same era.
Are sixteen heads really better than one?
Paul Michel, Omer Levy, and Graham Neubig · 2019
Cited alongside, same era.
Rina Foygel Barber, Emmanuel J Candes, Aaditya Ramdas, and Ryan J Tibshirani · 2021
Later among the works it cites.
Distribution-free, risk-controlling prediction sets
Stephen Bates, Anastasios Angelopoulos, Lihua Lei, Jitendra Malik, and Michael Jordan · 2021
Later among the works it cites.
Magic pyramid: Accelerating inference with early exiting and token pruning
Xuanli He, Iman Keivanloo, Yi Xu, Xiang He, Belinda Zeng, Santosh Rajagopalan, and Trishul Chilimbi · 2021
Later among the works it cites.
Learned token pruning for transformers
Sehoon Kim, Sheng Shen, David Thorsley, Amir Gholami, Woosuk Kwon, Joseph Hassoun, and Kurt Keutzer · 2021
Later among the works it cites.
Neurips 2020 efficientqa competition: Systems, analyses and lessons learned
Sewon Min, Jordan Boyd-Graber, Chris Alberti, Danqi Chen, Eunsol Choi, Michael Collins, Kelvin Guu, Hannaneh Hajishirzi, Kenton Lee, Jennimaria Palomaki, Colin Raffel, Adam Roberts, Tom Kwiatkowski, Patrick Lewis, Yuxiang Wu, Heinrich Küttler, Linqing Liu, Pasquale Minervini, Pontus Stenetorp, Sebastian Riedel, Sohee Yang, Minjoon Seo, Gautier Izacard, Fabio Petroni, Lucas Hosseini, Nicola De Cao, Edouard Grave, Ikuya Yamada, Sonse Shimaoka, Masatoshi Suzuki, Shumpei Miyawaki, Shun Sato, Ryo Takahashi, Jun Suzuki, Martin Fajcik, Martin Docekal, Karel Ondrej, Pavel Smrz, Hao Cheng, Yelong Shen, Xiaodong Liu, Pengcheng He, Weizhu Chen, Jianfeng Gao, Barlas Oguz, Xilun Chen, Vladimir Karpukhin, Stan Peshterliev, Dmytro Okhonko, Michael Schlichtkrull, Sonal Gupta, Yashar Mehdad, and Wen-tau Yih · 2021
Later among the works it cites.
Proceedings of the Second Workshop on Simple and Efficient Natural Language Processing , 2021
Nafise Sadat Moosavi, Iryna Gurevych, Angela Fan, Thomas Wolf, Yufang Hou, Ana Marasović, and Sujith Ravi (eds.) · 2021
Later among the works it cites.
Consistent accelerated inference via confident adaptive transformers
Tal Schuster, Adam Fisch, Tommi Jaakkola, and Regina Barzilay · 2021
Later among the works it cites.
Zero time waste: Recycling predictions in early exit neural networks
Maciej Wołczyk, Bartosz Wójcik, Klaudia Bałazy, Igor T Podolak, Jacek Tabor, Marek Śmieja, and Tomasz Trzcinski · 2021
Later among the works it cites.
Tr-bert: Dynamic token reduction for accelerating bert inference
Deming Ye, Yankai Lin, Yufei Huang, and Maosong Sun · 2021
Later among the works it cites.
Anastasios N Angelopoulos, Stephen Bates, Adam Fisch, Lihua Lei, and Tal Schuster · 2022
Closest in time.
Antonio Candelieri, Andrea Ponti, and Francesco Archetti · 2022
Closest in time.
Transkimmer: Transformer learns to layer-wise skim
Yue Guan, Zhengyi Li, Jingwen Leng, Zhouhan Lin, and Minyi Guo · 2022
Closest in time.
Multi-objective hyperparameter optimization–an overview
Florian Karl, Tobias Pielok, Julia Moosbauer, Florian Pfisterer, Stefan Coors, Martin Binder, Lennart Schneider, Janek Thomas, Jakob Richter, Michel Lang, et al · 2022
Closest in time.
Smac3: A versatile bayesian optimization package for hyperparameter optimization
Marius Lindauer, Katharina Eggensperger, Matthias Feurer, André Biedenkapp, Difan Deng, Carolin Benjamins, Tim Ruhkopf, René Sass, and Frank Hutter · 2022
Closest in time.
Adapler: Speeding up inference by adaptive length reduction
Ali Modarressi, Hosein Mohebbi, and Mohammad Taher Pilehvar · 2022
Closest in time.
Confident adaptive language modeling
Tal Schuster, Adam Fisch, Jai Gupta, Mostafa Dehghani, Dara Bahri, Vinh Q Tran, Yi Tay, and Donald Metzler · 2022
Closest in time.
Structured pruning learns compact and accurate models
Mengzhou Xia, Zexuan Zhong, and Danqi Chen · 2022
Closest in time.