Fetching the paper…
Reading the bibliography…
In this study, we delve into an emerging optimization challenge involving a black-box objective function that can only be gauged via a ranking oracle-a situation frequently encountered in real-world scenarios, especially when the function is evaluated by human judges.
A simplex method for function minimization
John A Nelder and Roger Mead · 1965
Earlier work this paper cites.
Decisions with multiple objectives: preferences and value trade-offs
Ralph L Keeney and Howard Raiffa · 1993
Earlier work this paper cites.
Direct search algorithms for optimization calculations
Michael JD Powell · 1998
Earlier work this paper cites.
Interactively optimizing information retrieval systems as a dueling bandits problem
Yisong Yue and Thorsten Joachims · 2009
Earlier work this paper cites.
Denoising diffusion implicit models
Jiaming Song, Chenlin Meng, and Stefano Ermon · 2010
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Yang Song, Jascha Sohl-Dickstein, Diederik P Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole · 2011
Earlier work this paper cites.
Robust 1-bit compressed sensing and sparse logistic regression: A convex programming approach
Yaniv Plan and Roman Vershynin · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Earlier work this paper cites.
The algorithmic foundations of differential privacy
Cynthia Dwork, Aaron Roth, et al · 2014
Earlier work this paper cites.
Optimal rates for zero-order convex optimization: The power of two function evaluations
John C Duchi, Michael I Jordan, Martin J Wainwright, and Andre Wibisono · 2015
Earlier work this paper cites.
Benchmarking deep reinforcement learning for continuous control
Yan Duan, Xi Chen, Rein Houthooft, John Schulman, and Pieter Abbeel · 2016
Earlier work this paper cites.
Cma-es for hyperparameter optimization of deep neural networks
Ilya Loshchilov and Frank Hutter · 2016
Earlier work this paper cites.
Regret analysis for continuous dueling bandit
Wataru Kumagai · 2017
Earlier work this paper cites.
Hyperband: A novel bandit-based approach to hyperparameter optimization
Lisha Li, Kevin Jamieson, Giulia DeSalvo, Afshin Rostamizadeh, and Ameet Talwalkar · 2017
Earlier work this paper cites.
Random gradient-free minimization of convex functions
Yurii Nesterov and Vladimir Spokoiny · 2017
Earlier work this paper cites.
Preference based adaptation for learning objectives
Yao-Xiang Ding and Zhi-Hua Zhou · 2018
Earlier work this paper cites.
A tutorial on bayesian optimization
Peter I Frazier · 2018
Earlier work this paper cites.
Sega: Variance reduction via gradient sketching
Filip Hanzely, Konstantin Mishchenko, and Peter Richtárik · 2018
Earlier work this paper cites.
Relation classification using coarse and fine-grained networks with sdp supervised key words selection
Yiping Sun, Yu Cui, Jinglu Hu, and Weijia Jia · 2018
Earlier work this paper cites.
Gradientless descent: High-dimensional zeroth-order optimization
Daniel Golovin, John Karro, Greg Kochanski, Chansoo Lee, Xingyou Song, and Qiuyi Zhang · 2019
Cited alongside, same era.
CMA-ES/pycma on Github
Nikolaus Hansen, Youhei Akimoto, and Petr Baudis · 2019
Cited alongside, same era.
Multi-attribute bayesian optimization with interactive preference learning
Raul Astudillo and Peter Frazier · 2020
Cited alongside, same era.
Optimal binomial reliability demonstration tests design under acceptance decision uncertainty
Suiyao Chen, Lu Lu, Qiong Zhang, and Mingyang Li · 2020
Cited alongside, same era.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Cited alongside, same era.
Learning to summarize with human feedback
Nisan Stiennon, Long Ouyang, Jeffrey Wu, Daniel Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul F Christiano · 2020
Chatgpt,https://openai.com/ blog/chatgpt/, 2022
OpenAI · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Later among the works it cites.
Fuzzy controller-based design and simulation of an automatic parking system
M. Gao, Y. Wei, Y. He, D. Zhang, Y. Tian, B. Huang, and C. Zheng · 2023
Closest in time.
Aligning text-to-image models using human feedback
Kimin Lee, Hao Liu, Moonkyung Ryu, Olivia Watkins, Yuqing Du, Craig Boutilier, Pieter Abbeel, Mohammad Ghavamzadeh, and Shixiang Shane Gu · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Locally differentially private (contextual) bandits learning
Kai Zheng, Tianle Cai, Weiran Huang, Zhenguo Li, and Liwei Wang · 2020
Cited alongside, same era.
Preference-based online learning with dueling bandits: A survey
Viktor Bengs, Róbert Busa-Fekete, Adil El Mesaoudi-Paul, and Eyke Hüllermeier · 2021
Cited alongside, same era.
Differentially private federated bayesian optimization with distributed exploration
Zhongxiang Dai, Bryan Kian Hsiang Low, and Patrick Jaillet · 2021
Cited alongside, same era.
Diffusion models beat gans on image synthesis
Prafulla Dhariwal and Alexander Nichol · 2021
Cited alongside, same era.
Gaussian differential privacy
Jinshuo Dong, Aaron Roth, and Weijie Su · 2021
Cited alongside, same era.
Human-in-the-Loop Machine Learning: Active learning and annotation for human-centered AI
Robert Munro Monarch · 2021
Cited alongside, same era.
Languages are rewards: Hindsight finetuning using human feedback, 2023
Hao Liu, Carmelo Sferrazza, and Pieter Abbeel · 2023
Closest in time.
ZeroPrompt: Streaming Acoustic Encoders are Zero-Shot Masked LMs
Xingchen Song, Di Wu, Binbin Zhang, Zhendong Peng, Bo Dang, Fuping Pan, and Zhiyong Wu · 2023
Closest in time.
Interpreting pretrained language models via concept bottlenecks
Zhen Tan, Lu Cheng, Song Wang, Yuan Bo, Jundong Li, and Huan Liu · 2023
Closest in time.
Low-rank matrix recovery with unknown correspondence
Zhiwei Tang, Tsung-Hui Chang, Xiaojing Ye, and Hongyuan Zha · 2023
Closest in time.
Embracing uncertainty: A diffusion generative model of spectrum efficiency in 5g networks
Hao Wang, Zhiwei Tang, Shutao Zhang, Chao Shen, and Tsung-Hui Chang · 2023
Closest in time.
Breast cancer prediction based on machine learning
Y. Wei, D. Zhang, M. Gao, Y. Tian, Y. He, B. Huang, and C. Zheng · 2023
Closest in time.
Hard prompts made easy: Gradient-based discrete optimization for prompt tuning and discovery
Yuxin Wen, Neel Jain, John Kirchenbauer, Micah Goldblum, Jonas Geiping, and Tom Goldstein · 2023
Closest in time.
Fostc3net: A lightweight yolov5 based on the network structure optimization
Danqing Ma, Shaojie Li, Bo Dang, Hengyi Zang, and Xinqi Dong · 2024
Closest in time.
Smartfix: Leveraging machine learning for proactive equipment maintenance in industry 4.0
Fanghao Ni, Hengyi Zang, and Yuxin Qiao · 2024
Closest in time.
Large language models for forecasting and anomaly detection: A systematic literature review
Jing Su, Chufeng Jiang, Xin Jin, Yuxin Qiao, Tingsong Xiao, Hongda Ma, Rong Wei, Zhi Jing, Jiajun Xu, and Junhong Lin · 2024
Closest in time.
Transtarec: Time-adaptive translating embedding model for next poi recommendation
Yiping Sun · 2024
Closest in time.
Fedlion: Faster adaptive federated optimization with fewer communication
Zhiwei Tang and Tsung-Hui Chang · 2024
Closest in time.
Switchtab: Switched autoencoders are effective tabular learners
Jing Wu, Suiyao Chen, Qi Zhao, Renat Sergazinov, Chen Li, Shengjie Liu, Chongchao Zhao, Tianpei Xie, Hanqing Guo, Cheng Ji, et al · 2024
Closest in time.
Precision calibration of industrial 3d scanners: An ai-enhanced approach for improved measurement accuracy
Hengyi Zang · 2024
Closest in time.