Fetching the paper…
Reading the bibliography…
Bridging the exponentially growing gap between the numbers of unlabeled and labeled protein sequences, several studies adopted semi-supervised learning for protein sequence modeling.
J. H. Steiger, “Tests for comparing elements of a correlation matrix.” Psychological bulletin , vol. 87, no. 2, p. 245, 1980
1980
Earlier work this paper cites.
T. E. Creighton, Proteins: structures and molecular properties . Macmillan, 1993
1993
Earlier work this paper cites.
L. Holm and C. Sander, “Mapping the protein universe,” Science , vol. 273, no. 5275, pp. 595–602, 1996
1996
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
S. F. Altschul, T. L. Madden, A. A. Schäffer, J. Zhang, Z. Zhang, W. Miller, and D. J. Lipman, “Gapped blast and psi-blast: a new generation of protein database search programs,” Nucleic acids research , vol. 25, no. 17, pp. 3389–3402, 1997
1997
Earlier work this paper cites.
A. Elofsson and E. Sonnhammer, “A comparison of sequence and structure protein domain families as a basis for structural genomics.” Bioinformatics (Oxford, England) , vol. 15, no. 6, pp. 480–500, 1999
1999
Earlier work this paper cites.
J. A. Cuff and G. J. Barton, “Evaluation and improvement of multiple sequence methods for protein secondary structure prediction,” Proteins: Structure, Function, and Bioinformatics , vol. 34, no. 4, pp. 508–519, 1999
1999
Earlier work this paper cites.
A. Bairoch and R. Apweiler, “The swiss-prot protein sequence database and its supplement trembl in 2000,” Nucleic acids research , vol. 28, no. 1, pp. 45–48, 2000
2000
Earlier work this paper cites.
H. M. Berman, P. E. Bourne, J. Westbrook, and C. Zardecki, “The protein data bank,” in Protein Structure . CRC Press, 2003, pp. 394–410
2003
Earlier work this paper cites.
S. R. Eddy, “Where did the blosum62 alignment score matrix come from?” Nature biotechnology , vol. 22, no. 8, pp. 1035–1036, 2004
2004
Earlier work this paper cites.
R. Apweiler, A. Bairoch, C. H. Wu, W. C. Barker, B. Boeckmann, S. Ferro, E. Gasteiger, H. Huang, R. Lopez, M. Magrane et al. , “Uniprot: the universal protein knowledgebase,” Nucleic acids research , vol. 32, no. suppl_1, pp. D115–D119, 2004
2004
Earlier work this paper cites.
J. Söding, A. Biegert, and A. N. Lupas, “The hhpred interactive server for protein homology detection and structure prediction,” Nucleic acids research , vol. 33, no. suppl_2, pp. W244–W248, 2005
2005
Earlier work this paper cites.
Y. Zhang and J. Skolnick, “Tm-align: a protein structure alignment algorithm based on the tm-score,” Nucleic acids research , vol. 33, no. 7, pp. 2302–2309, 2005
2005
Earlier work this paper cites.
J. M. Berg, J. L. Tymoczko, and L. Stryer, “Biochemistry. 5th,” New York: WH Freeman , vol. 38, no. 894, p. 76, 2006
2006
Earlier work this paper cites.
T. L. Bailey, N. Williams, C. Misleh, and W. W. Li, “Meme: discovering and analyzing dna and protein sequence motifs,” Nucleic acids research , vol. 34, no. suppl_2, pp. W369–W373, 2006
2006
Earlier work this paper cites.
O. Chapelle, B. Scholkopf, and A. Zien, “Semi-supervised learning (chapelle, o. et al., eds.; 2006)[book reviews],” IEEE Transactions on Neural Networks , vol. 20, no. 3, pp. 542–542, 2009
2009
Earlier work this paper cites.
H. M. Berman, J. D. Westbrook, M. J. Gabanyi, W. Tao, R. Shah, A. Kouranov, T. Schwede, K. Arnold, F. Kiefer, L. Bordoli et al. , “The protein structure initiative structural genomics knowledgebase,” Nucleic acids research , vol. 37, no. suppl_1, pp. D365–D368, 2009
2009
Earlier work this paper cites.
K.-C. Chou, Z.-C. Wu, and X. Xiao, “iloc-euk: a multi-label classifier for predicting the subcellular localization of singleplex and multiplex eukaryotic proteins,” PloS one , vol. 6, no. 3, 2011
2011
Earlier work this paper cites.
M. Remmert, A. Biegert, A. Hauser, and J. Söding, “Hhblits: lightning-fast iterative protein sequence searching by hmm-hmm alignment,” Nature methods , vol. 9, no. 2, p. 173, 2012
2012
Earlier work this paper cites.
N. K. Fox, S. E. Brenner, and J.-M. Chandonia, “Scope: Structural classification of proteins—extended, integrating scop and astral data and classification of new structures,” Nucleic acids research , vol. 42, no. D1, pp. D304–D309, 2013
2013
Cited alongside, same era.
L. S. Tavares, C. d. S. F. d. Silva, V. C. Souza, V. L. d. Silva, C. G. Diniz, and M. D. O. Santos, “Strategies and molecular tools to fight antimicrobial resistance: resistome, transcriptome, and antimicrobial peptides,” Frontiers in microbiology , vol. 4, p. 412, 2013
2013
Cited alongside, same era.
R. D. Finn, A. Bateman, J. Clements, P. Coggill, R. Y. Eberhardt, S. R. Eddy, A. Heger, K. Hetherington, L. Holm, J. Mistry et al. , “Pfam: the protein families database,” Nucleic acids research , vol. 42, no. D1, pp. D222–D230, 2014
2014
Cited alongside, same era.
2018
Later among the works it cites.
2018
Later among the works it cites.
S. Khurana, R. Rawi, K. Kunji, G.-Y. Chuang, H. Bensmail, and R. Mall, “Deepsol: a deep learning framework for sequence-based protein solubility prediction,” Bioinformatics , vol. 34, no. 15, pp. 2605–2613, 2018
2018
Later among the works it cites.
R. Poplin, P.-C. Chang, D. Alexander, S. Schwartz, T. Colthurst, A. Ku, D. Newburger, J. Dijamco, N. Nguyen, P. T. Afshar et al. , “A universal snp and small-indel variant caller using deep neural networks,” Nature biotechnology , vol. 36, no. 10, pp. 983–987, 2018
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2014
Cited alongside, same era.
2014
Cited alongside, same era.
C. C. H. Chang, J. Song, B. T. Tey, and R. N. Ramanan, “Bioinformatics approaches for improved recombinant protein production in escherichia coli: protein solubility prediction,” Briefings in bioinformatics , vol. 15, no. 6, pp. 953–962, 2014
2014
Cited alongside, same era.
E. Asgari and M. R. Mofrad, “Continuous distributed representation of biological sequences for deep proteomics and genomics,” PloS one , vol. 10, no. 11, p. e0141287, 2015
2015
Cited alongside, same era.
K. D. Tsirigos, C. Peters, N. Shu, L. Käll, and A. Elofsson, “The topcons web server for consensus prediction of membrane protein topology and signal peptides,” Nucleic acids research , vol. 43, no. W1, pp. W401–W407, 2015
2015
Cited alongside, same era.
2016
Cited alongside, same era.
K. S. Sarkisyan, D. A. Bolotin, M. V. Meer, D. R. Usmanova, A. S. Mishin, G. V. Sharonov, D. N. Ivankov, N. G. Bozhanova, M. S. Baranov, O. Soylemez et al. , “Local fitness landscape of the green fluorescent protein,” Nature , vol. 533, no. 7603, pp. 397–401, 2016
2016
Cited alongside, same era.
S. Wang, J. Peng, J. Ma, and J. Xu, “Protein secondary structure prediction using deep convolutional neural fields,” Scientific reports , vol. 6, p. 18962, 2016
2016
Cited alongside, same era.
S. Min, B. Lee, and S. Yoon, “Deep learning in bioinformatics,” Briefings in bioinformatics , vol. 18, no. 5, pp. 851–869, 2017
2017
Cited alongside, same era.
R. Rawi, R. Mall, K. Kunji, C.-H. Shen, P. D. Kwong, and G.-Y. Chuang, “Parsnip: sequence-based protein solubility prediction using gradient boosting machine,” Bioinformatics , vol. 34, no. 7, pp. 1092–1098, 2018
2018
Later among the works it cites.
L. A. Abriata, G. E. Tamò, B. Monastyrskyy, A. Kryshtafovych, and M. Dal Peraro, “Assessment of hard target modeling in casp12 reveals an emerging role of alignment-based contact prediction methods,” Proteins: Structure, Function, and Bioinformatics , vol. 86, pp. 97–112, 2018
2018
Later among the works it cites.
R. Rao, N. Bhattacharya, N. Thomas, Y. Duan, X. Chen, J. Canny, P. Abbeel, and Y. S. Song, “Evaluating protein transfer learning with tape,” in Advances in neural information processing systems , 2019, pp. 9689–9701
2019
Closest in time.
M. AlQuraishi, “End-to-end differentiable learning of protein structure,” Cell systems , vol. 8, no. 4, pp. 292–301, 2019
2019
Closest in time.
E. C. Alley, G. Khimulya, S. Biswas, M. AlQuraishi, and G. M. Church, “Unified rational protein engineering with sequence-based deep representation learning,” Nature methods , vol. 16, no. 12, pp. 1315–1322, 2019
2019
Closest in time.
T. Bepler and B. Berger, “Learning protein sequence embeddings using information from structure,” in International Conference on Learning Representations , 2019, p.
2019
Closest in time.
A. Rives, S. Goyal, J. Meier, D. Guo, M. Ott, C. L. Zitnick, J. Ma, and R. Fergus, “Biological structure and function emerge from scaling unsupervised learning to 250 million protein sequences,” bioRxiv , p. 622803, 2019
2019
Closest in time.
N. Strodthoff, P. Wagner, M. Wenzel, and W. Samek, “Udsmprot: Universal deep sequence models for protein classification,” bioRxiv , p. 704874, 2019
2019
Closest in time.
M. Heinzinger, A. Elnaggar, Y. Wang, C. Dallago, D. Nechaev, F. Matthes, and B. Rost, “Modeling aspects of the language of life through transfer-learning protein sequences,” BMC bioinformatics , vol. 20, no. 1, p. 723, 2019
2019
Closest in time.
2019
Closest in time.
M. S. Klausen, M. C. Jespersen, H. Nielsen, K. K. Jensen, V. I. Jurtz, C. K. Soenderby, M. O. A. Sommer, O. Winther, M. Nielsen, B. Petersen et al. , “Netsurfp-2.0: Improved prediction of protein structural features by integrated deep learning,” Proteins: Structure, Function, and Bioinformatics , vol. 87, no. 6, pp. 520–527, 2019
2019
Closest in time.
A. Kryshtafovych, T. Schwede, M. Topf, K. Fidelis, and J. Moult, “Critical assessment of methods of protein structure prediction (casp)—round xiii,” Proteins: Structure, Function, and Bioinformatics , vol. 87, no. 12, pp. 1011–1020, 2019
2019
Closest in time.
A. X. Lu, H. Zhang, M. Ghassemi, and A. Moses, “Self-supervised contrastive learning of protein representations by mutual information maximization,” bioRxiv , 2020
2020
Closest in time.
2020
Closest in time.
S. Sukhbaatar, E. Grave, P. Bojanowski, and A. Joulin, “Adaptive attention span in transformers,” in ACL , 2019
2021
Closest in time.