Tensorflow: a system for large-scale machine learning
M. Abadi, P. Barham, J. Chen, Z. Chen, A. Davis, J. Dean, M. Devin, S. Ghemawat, G. Irving, M. Isard, et al · 2016
Later among the works it cites.
mlr: Machine Learning in R
B. Bischl, M. Lang, L. Kotthoff, J. Schiffner, J. Richter, E. Studerus, G. Casalicchio, and Z. M. Jones · 2016
Later among the works it cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
A. Wang, A. Singh, J. Michael, F. Hill, O. Levy, and S. Bowman · 2018
Later among the works it cites.
Critical assessment of methods of protein structure prediction (CASP) - Round XIII
A. Kryshtafovych, T. Schwede, M. Topf, K. Fidelis, and J. Moult · 2019
Later among the works it cites.
Tunability: Importance of Hyperparameters of Machine Learning Algorithms
P. Probst, A.-L. Boulesteix, and B. Bischl · 2019
Later among the works it cites.
Superglue: A stickier benchmark for general-purpose language understanding systems
A. Wang, Y. Pruksachatkun, N. Nangia, A. Singh, J. Michael, F. Hill, O. Levy, and S. Bowman · 2019
Later among the works it cites.
The CAFA challenge reports improved protein function prediction and new functional annotations for hundreds of genes through experimental screens
N. Zhou, Y. Jiang, T. R. Bergquist, A. J. Lee, B. Z. Kacsoh, A. W. Crocker, K. A. Lewis, G. Georghiou, H. N. Nguyen, M. N. Hamid, et al · 2019
Later among the works it cites.
Modeling protein-protein, protein-peptide, and protein-oligosaccharide complexes: CAPRI 7th edition
M. F. Lensink, N. Nadzirin, S. Velankar, and S. J. Wodak · 2020
Closest in time.
OpenML Benchmarking Suites
B. Bischl, B. Bischl, G. Casalicchio, M. Feurer, P. Gijsbers, F. Hutter, M. Lang, R. Gomes Mantovani, J. van Rijn, and J. Vanschoren · 2021
Closest in time.
Research community dynamics behind popular AI benchmarks
F. Martínez-Plumed, P. Barredo, S. Ó. hÉigeartaigh, and J. Hernández-Orallo · 2021
Closest in time.