Fetching the paper…
Reading the bibliography…
This paper presents solutions to the Machine Learning Model Attribution challenge (MLMAC) collectively organized by MITRE, Microsoft, Schmidt-Futures, Robust-Intelligence, Lincoln-Network, and Huggingface community.
1906
Earlier work this paper cites.
1907
Earlier work this paper cites.
1909
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics . Philadelphia, Pennsylvania, USA: Association for Computational Linguistics, Jul. 2002, pp. 311–318. [Online]. Available: https://aclanthology.org/P02-1040
2002
Earlier work this paper cites.
X. Li and D. Roth, “Learning question classifiers,” in COLING 2002: The 19th International Conference on Computational Linguistics , 2002. [Online]. Available: https://www.aclweb.org/anthology/C02-1150
2002
Earlier work this paper cites.
D. Dubin, “The most influential paper gerard salton never wrote,” Libr. Trends , vol. 52, pp. 748–764, 2004
2004
Earlier work this paper cites.
R. Rifkin and A. Klautau, “In defense of one-vs-all classification,” J. Mach. Learn. Res. , vol. 5, p. 101–141, dec 2004
2004
Earlier work this paper cites.
B. Pang and L. Lee, “Seeing stars: Exploiting class relationships for sentiment categorization with respect to rating scales,” in Proceedings of the ACL , 2005
2005
Earlier work this paper cites.
M. Snover, B. Dorr, R. Schwartz, L. Micciulla, and J. Makhoul, “A study of translation edit rate with targeted human annotation,” in Proceedings of the 7th Conference of the Association for Machine Translation in the Americas: Technical Papers . Cambridge, Massachusetts, USA: Association for Machine Translation in the Americas, Aug. 8-12 2006, pp. 223–231. [Online]. Available: https://aclanthology.org/2006.amta-papers.25
2006
Earlier work this paper cites.
A. L. Maas, R. E. Daly, P. T. Pham, D. Huang, A. Y. Ng, and C. Potts, “Learning word vectors for sentiment analysis,” in Proceedings of the 49th Annual Meeting of the Association for Computational Linguistics: Human Language Technologies . Portland, Oregon, USA: Association for Computational Linguistics, June 2011, pp. 142–150. [Online]. Available: http://www.aclweb.org/anthology/P11-1015
2011
Earlier work this paper cites.
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. u. Kaiser, and I. Polosukhin, “Attention is all you need,” in Advances in Neural Information Processing Systems , I. Guyon, U. V. Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, and R. Garnett, Eds., vol. 30. Curran Associates, Inc., 2017, https://proceedings.neurips.cc/paper/2017/file/3f5ee243547dee91fbd053c1c4a845aa-Paper.pdf
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
A. Radford and K. Narasimhan, “Improving language understanding by generative pre-training,” 2018
2018
Cited alongside, same era.
A. Wang, A. Singh, J. Michael, F. Hill, O. Levy, and S. R. Bowman, “GLUE: A multi-task benchmark and analysis platform for natural language understanding,” 2019, in the Proceedings of ICLR
2019
Cited alongside, same era.
2019
Cited alongside, same era.
V. Sanh, L. Debut, J. Chaumond, and T. Wolf, “Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter,” in NeurIPS E M C 2 EMC^{2} Workshop , 2019
K. Clark, M.-T. Luong, Q. V. Le, and C. D. Manning, “ELECTRA: Pre-training text encoders as discriminators rather than generators,” in ICLR , 2020. [Online]. Available: https://openreview.net/pdf?id=r1xMH1BtvB
2020
Later among the works it cites.
S. Black, G. Leo, P. Wang, C. Leahy, and S. Biderman, “GPT-Neo: Large Scale Autoregressive Language Modeling with Mesh-Tensorflow,” Mar. 2021, If you use this software, please cite it using these metadata. [Online]. Available: https://doi.org/10.5281/zenodo.5297715
2021
Later among the works it cites.
B. Wang, “Mesh-Transformer-JAX: Model-Parallel Implementation of Transformer Language Model with JAX,” https://github.com/kingoflolz/mesh-transformer-jax , May 2021
2021
Later among the works it cites.
“MLMAC — mlmac.io,” https://mlmac.io/#overview , [Accessed 04-Nov-2022]
2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever et al. , “Language models are unsupervised multitask learners,” OpenAI blog , vol. 1, no. 8, p. 9, 2019
2019
Cited alongside, same era.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, and I. Sutskever, “Language models are unsupervised multitask learners,” 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
P. J. Ortiz Su’arez, L. Romary, and B. Sagot, “A monolingual approach to contextualized word embeddings for mid-resource languages,” in Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . Online: Association for Computational Linguistics, Jul. 2020, pp. 1703–1714. [Online]. Available: https://www.aclweb.org/anthology/2020.acl-main.156
2020
Cited alongside, same era.
Y. Bisk, R. Zellers, R. L. Bras, J. Gao, and Y. Choi, “Piqa: Reasoning about physical commonsense in natural language,” in Thirty-Fourth AAAI Conference on Artificial Intelligence , 2020
2020
Cited alongside, same era.
W. Wang, F. Wei, L. Dong, H. Bao, N. Yang, and M. Zhou, “Minilm: Deep self-attention distillation for task-agnostic compression of pre-trained transformers,” 2020
2020
Cited alongside, same era.
Y. Zhang, S. Sun, M. Galley, Y.-C. Chen, C. Brockett, X. Gao, J. Gao, J. Liu, and B. Dolan, “Dialogpt: Large-scale generative pre-training for conversational response generation,” in ACL, system demonstration , 2020
2020
Cited alongside, same era.
Wikipedia, “Prisoner’s dilemma — Wikipedia, the free encyclopedia,” http://en.wikipedia.org/w/index.php?title=Prisoner’s%20dilemma&oldid=1112692700 , 2022, [Online; accessed 29-September-2022]
2022
Closest in time.
2022
Closest in time.
“GitHub - FarhanDhanani/MLMAC — github.com,” https://github.com/FarhanDhanani/MLMAC , [Accessed 06-Nov-2022]
2022
Closest in time.
“model-attribution-challenge/bloom-2b5 · Hugging Face — huggingface.co,” https://huggingface.co/model-attribution-challenge/bloom-2b5 , [Accessed 06-Nov-2022]
2022
Closest in time.
“model-attribution-challenge/bloom-350m · Hugging Face — huggingface.co,” https://huggingface.co/model-attribution-challenge/bloom-350m , [Accessed 06-Nov-2022]
2022
Closest in time.
E. Nijkamp, B. Pang, H. Hayashi, L. Tu, H. Wang, Y. Zhou, S. Savarese, and C. Xiong, “A conversational paradigm for program synthesis,” arXiv preprint , 2022
2022
Closest in time.
S. Zhang, S. Roller, N. Goyal, M. Artetxe, M. Chen, S. Chen, C. Dewan, M. Diab, X. Li, X. V. Lin, T. Mihaylov, M. Ott, S. Shleifer, K. Shuster, D. Simig, P. S. Koura, A. Sridhar, T. Wang, and L. Zettlemoyer, “Opt: Open pre-trained transformer language models,” 2022
2022
Closest in time.