Fetching the paper…
Reading the bibliography…
Language models serve as a cornerstone in natural language processing (NLP), utilizing mathematical methods to generalize language laws and knowledge for prediction and generation.
Pinker, S.: The language instinct: How the mind creates. Language. New York: Harper Collins (1994)
1994
Earlier work this paper cites.
Jelinek, F.: Statistical methods for speech recognition. MIT Press (1998)
1998
Earlier work this paper cites.
LeCun, Y., Bottou, L., Bengio, Y., Haffner, P.: Gradient-based learning applied to document recognition. Proceedings of the IEEE 86
1998
Earlier work this paper cites.
Turing, A.M., Geirsson, H., Losonsky, M.: Computing machinery and intelligence. Artificial Intelligence: Critical Concepts 2
2000
Earlier work this paper cites.
Rosenfeld, R.: Two decades of statistical language modeling: Where do we go from here? Proceedings of the IEEE 88
2000
Earlier work this paper cites.
Bengio, Y., Ducharme, R., Vincent, P.: A neural probabilistic language model. Advances in neural information processing systems 13
2000
Earlier work this paper cites.
Stolcke, A., et al
2002
Earlier work this paper cites.
Mikolov, T., Karafiát, M., Burget, L., Cernockỳ, J., Khudanpur, S.: Recurrent neural network based language model. In: Interspeech, vol. 2, pp. 1045–1048 (2010). Makuhari
2010
Earlier work this paper cites.
Kombrink, S., Mikolov, T., Karafiát, M., Burget, L.: Recurrent neural network based language modeling in meeting recognition. In: Interspeech, vol. 11, pp. 2877–2880 (2011)
2011
Earlier work this paper cites.
Mikolov, T., Sutskever, I., Chen, K., Corrado, G.S., Dean, J.: Distributed representations of words and phrases and their compositionality. Advances in neural information processing systems 26
2013
Earlier work this paper cites.
Kiros, R., Zhu, Y., Salakhutdinov, R.R., Zemel, R., Urtasun, R., Torralba, A., Fidler, S.: Skip-thought vectors. Advances in neural information processing systems 28
2015
Earlier work this paper cites.
Fredrikson, M., Jha, S., Ristenpart, T.: Model inversion attacks that exploit confidence information and basic countermeasures. In: Proceedings of the 22nd ACM SIGSAC Conference on Computer and Communications Security, pp. 1322–1333 (2015)
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
Kaiming, H., Xiangyu, Z., Shaoqing, R., Jian, S., et al
2016
Earlier work this paper cites.
Ba, J.L., Kiros, J.R., Hinton, G.E.: Layer normalization. arXiv preprint arXiv:1607.06450 (2016)
2016
Earlier work this paper cites.
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, Ł., Polosukhin, I.: Attention is all you need. Advances in neural information processing systems 30
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Radford, A., Narasimhan, K., Salimans, T., Sutskever, I., et al.: Improving language understanding by generative pre-training (2018)
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., Sutskever, I., et al
2019
Earlier work this paper cites.
Zhang, W., Ntoutsi, E.: Faht: an adaptive fairness-aware decision tree classifier. In: International Joint Conference on Artificial Intelligence (IJCAI), pp. 1480–1486 (2019)
2019
Earlier work this paper cites.
Yang, Z., Dai, Z., Yang, Y., Carbonell, J., Salakhutdinov, R.R., Le, Q.V.: Xlnet: Generalized autoregressive pretraining for language understanding. Advances in neural information processing systems 32
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J.D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al
2020
Earlier work this paper cites.
Gupta, A., Dengre, V., Kheruwala, H.A., Shah, M.: Comprehensive review of text-mining applications in finance. Financial Innovation 6
2020
Earlier work this paper cites.
Mozafari, M., Farahbakhsh, R., Crespi, N.: Hate speech detection and racial bias mitigation in social media based on bert model. PloS one 15
2020
Earlier work this paper cites.
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., Liu, P.J.: Exploring the limits of transfer learning with a unified text-to-text transformer. Journal of machine learning research 21
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Haley, B., Roudnicky, F.: Functional genomics for cancer drug target discovery. Cancer Cell 38
2020
Earlier work this paper cites.
Paananen, J., Fortino, V.: An omics perspective on drug target discovery platforms. Briefings in bioinformatics 21
2020
Earlier work this paper cites.
Zhang, Z., Zohren, S., Roberts, S.: Deep learning for portfolio optimization. The Journal of Financial Data Science (2020)
2020
Earlier work this paper cites.
Mashrur, A., Luo, W., Zaidi, N.A., Robles-Kelly, A.: Machine learning for financial risk management: a survey. Ieee Access 8
2020
Earlier work this paper cites.
Han, X., Zhang, Z., Ding, N., Gu, Y., Liu, X., Huo, Y., Qiu, J., Yao, Y., Zhang, A., Zhang, L., et al
2021
Earlier work this paper cites.
Dodge, J., Sap, M., Marasović, A., Agnew, W., Ilharco, G., Groeneveld, D., Mitchell, M., Gardner, M.: Documenting large webtext corpora: A case study on the colossal clean crawled corpus. In: Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing (EMNLP), pp. 1286–1305 (2021)
2021
Earlier work this paper cites.
Kim, B., Kim, H., Lee, S.-W., Lee, G., Kwak, D., Hyeon, J.D., Park, S., Kim, S., Kim, S., Seo, D., et al
2021
Earlier work this paper cites.
Ahmad, W., Chakraborty, S., Ray, B., Chang, K.: Unified pre-training for program understanding and generation. In: Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (2021)
2021
Earlier work this paper cites.
Sheng, E., Chang, K.-W., Natarajan, P., Peng, N.: Societal biases in language generation: Progress and challenges. In: Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers), pp. 4275–4293 (2021)
2021
Earlier work this paper cites.
Wei, J., Bosma, M., Zhao, V., Guu, K., Yu, A.W., Lester, B., Du, N., Dai, A.M., Le, Q.V.: Finetuned language models are zero-shot learners. In: International Conference on Learning Representations (2021)
2021
Earlier work this paper cites.
Pagliaro, C., Mehta, D., Shiao, H.-T., Wang, S., Xiong, L.: Investor behavior modeling by analyzing financial advisor notes: a machine learning perspective. In: Proceedings of the Second ACM International Conference on AI in Finance, pp. 1–8 (2021)
2021
Earlier work this paper cites.
Zhang, W., Weiss, J.: Fair decision-making under uncertainty. In: 2021 IEEE International Conference on Data Mining (ICDM) (2021). IEEE
2021
Earlier work this paper cites.
Zhang, W., Bifet, A., Zhang, X., Weiss, J.C., Nejdl, W.: Farf: A fair and adaptive random forests classifier. In: Pacific-Asia Conference on Knowledge Discovery and Data Mining, pp. 245–256 (2021). Springer
2021
Earlier work this paper cites.
Jin, D., Pan, E., Oufattole, N., Weng, W.-H., Fang, H., Szolovits, P.: What disease does this patient have? a large-scale open domain question answering dataset from medical exams. Applied Sciences 11
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Zhuang, L., Wayne, L., Ya, S., Jun, Z.: A robustly optimized bert pre-training approach with post-training. In: Proceedings of the 20th Chinese National Conference on Computational Linguistics, pp. 1218–1227 (2021)
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Lieber, O., Sharir, O., Lenz, B., Shoham, Y.: Jurassic-1: Technical details and evaluation. White Paper. AI21 Labs 1
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Carlini, N., Tramer, F., Wallace, E., Jagielski, M., Herbert-Voss, A., Lee, K., Roberts, A., Brown, T., Song, D., Erlingsson, U., et al
2021
Earlier work this paper cites.
Bender, E.M., Gebru, T., McMillan-Major, A., Shmitchell, S.: On the dangers of stochastic parrots: Can language models be too big? In: Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency, pp. 610–623 (2021)
2021
Earlier work this paper cites.
2021
Cited alongside, same era.
Fu, Y., Peng, H., Khot, T.: How does gpt obtain its ability? tracing emergent abilities of language models to their sources. Yao Fu’s Notion (2022)
2022
Cited alongside, same era.
Wei, J., Tay, Y., Bommasani, R., Raffel, C., Zoph, B., Borgeaud, S., Yogatama, D., Bosma, M., Zhou, D., Metzler, D., et al.: Emergent abilities of large language models. Transactions on Machine Learning Research (2022)
2022
Cited alongside, same era.
Nijkamp, E., Pang, B., Hayashi, H., Tu, L., Wang, H., Zhou, Y., Savarese, S., Xiong, C.: Codegen: An open large language model for code with multi-turn program synthesis. In: The Eleventh International Conference on Learning Representations (2022)
2022
Cited alongside, same era.
2023
Later among the works it cites.
Le Scao, T., Fan, A., Akiki, C., Pavlick, E., Ilić, S., Hesslow, D., Castagné, R., Luccioni, A.S., Yvon, F., Gallé, M., et al.: Bloom: A 176b-parameter open-access multilingual language model (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Sadybekov, A.V., Katritch, V.: Computational approaches streamlining drug discovery. Nature 616
2023
Later among the works it cites.
Savage, N.: Drug discovery companies are customizing chatgpt: here’s how. Nat Biotechnol 41
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sanh, V., Webson, A., Raffel, C., Bach, S.H., Sutawika, L., Alyafeai, Z., Chaffin, A., Stiegler, A., Le Scao, T., Raja, A., et al
2022
Cited alongside, same era.
Merity, S., Xiong, C., Bradbury, J., Socher, R.: Pointer sentinel mixture models (2022)
2022
Cited alongside, same era.
Du, N., Huang, Y., Dai, A.M., Tong, S., Lepikhin, D., Xu, Y., Krikun, M., Zhou, Y., Yu, A.W., Firat, O., et al
2022
Cited alongside, same era.
Wang, T., Roberts, A., Hesslow, D., Le Scao, T., Chung, H.W., Beltagy, I., Launay, J., Raffel, C.: What language model architecture and pretraining objective works best for zero-shot generalization? In: International Conference on Machine Learning, pp. 22964–22984 (2022). PMLR
2022
Cited alongside, same era.
Zhang, W., Weiss, J.C.: Longitudinal fairness with censorship. In: Proceedings of the AAAI Conference on Artificial Intelligence, vol. 36, pp. 12235–12243 (2022)
2022
Cited alongside, same era.
Fedus, W., Zoph, B., Shazeer, N.: Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity. Journal of Machine Learning Research 23
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Li, Y., Choi, D., Chung, J., Kushman, N., Schrittwieser, J., Leblond, R., Eccles, T., Keeling, J., Gimeno, F., Dal Lago, A., et al
2022
Cited alongside, same era.
2023
Later among the works it cites.
2023
Later among the works it cites.
Thirunavukarasu, A.J., Ting, D.S.J., Elangovan, K., Gutierrez, L., Tan, T.F., Ting, D.S.W.: Large language models in medicine. Nature medicine 29
2023
Later among the works it cites.
2023
Later among the works it cites.
Arora, A., Arora, A.: The promise of large language models in health care. The Lancet 401
2023
Later among the works it cites.
Iu, K.Y., Wong, V.M.-Y.: Chatgpt by openai: The end of litigation lawyers? Available at SSRN 4339839 (2023)
2023
Later among the works it cites.
Markel, J.M., Opferman, S.G., Landay, J.A., Piech, C.: Gpteach: Interactive ta training with gpt-based students. In: Proceedings of the Tenth Acm Conference on Learning@ Scale, pp. 226–236 (2023)
2023
Later among the works it cites.
Tu, S., Zhang, Z., Yu, J., Li, C., Zhang, S., Yao, Z., Hou, L., Li, J.: Littlemu: Deploying an online virtual teaching assistant via heterogeneous sources integration and chain of teach prompts. In: Proceedings of the 32nd ACM International Conference on Information and Knowledge Management, pp. 4843–4849 (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Zhang, B.: Preparing educators and students for chatgpt and ai technology in higher education. ResearchGate (2023)
2023
Later among the works it cites.
Dwivedi, Y.K., Kshetri, N., Hughes, L., Slade, E.L., Jeyaraj, A., Kar, A.K., Baabdullah, A.M., Koohang, A., Raghavan, V., Ahuja, M., et al
2023
Later among the works it cites.
Chen, Y., Jensen, S., Albert, L.J., Gupta, S., Lee, T.: Artificial intelligence (ai) student assistants in the classroom: Designing chatbots to support student success. Information Systems Frontiers 25
2023
Later among the works it cites.
Leboukh, F., Aduku, E.B., Ali, O.: Balancing chatgpt and data protection in germany: challenges and opportunities for policy makers. Journal of Politics and Ethics in New Technologies and AI 2
2023
Later among the works it cites.
2023
Later among the works it cites.
Amos, Z.: What is fraudgpt? (2023)
2023
Later among the works it cites.
Delley, D.: Wormgpt–the generative ai tool cybercriminals are using to launch business email compromise attacks. SlashNext. Retrieved August 24
2023
Later among the works it cites.
Saxena, N.A., Zhang, W., Shahabi, C.: Missed opportunities in fair ai. In: Proceedings of the 2023 SIAM International Conference on Data Mining (SDM), pp. 961–964 (2023). SIAM
2023
Later among the works it cites.
Wang, Z., Wallace, C., Bifet, A., Yao, X., Zhang, W.: Fg 2 an: Fairness-aware graph generative adversarial networks. In: Joint European Conference on Machine Learning and Knowledge Discovery in Databases, pp. 259–275 (2023). Springer Nature Switzerland
2023
Later among the works it cites.
Wang, Z., Saxena, N., Yu, T., Karki, S., Zetty, T., Haque, I., Zhou, S., Kc, D., Stockwell, I., Bifet, A., et al
2023
Later among the works it cites.
Zhang, W., Weiss, J.C.: Fairness with censorship and group constraints. Knowledge and Information Systems, 1–24 (2023)
2023
Later among the works it cites.
Mei, K., Fereidooni, S., Caliskan, A.: Bias against 93 stigmatized groups in masked language models and downstream sentiment classification tasks. In: Proceedings of the 2023 ACM Conference on Fairness, Accountability, and Transparency, pp. 1699–1710 (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Pal, A., Umapathi, L.K., Sankarasubbu, M.: Med-halt: Medical domain hallucination test for large language models. In: Proceedings of the 27th Conference on Computational Natural Language Learning (CoNLL), pp. 314–334 (2023)
2023
Later among the works it cites.
Small, Z.: Sarah silverman sues openai and meta over copyright infringement. The New York Times (2023)
2023
Later among the works it cites.
Stempel, J.: NY Times sues openai, Microsoft for infringing copyrighted works … Thomson Reuters Corporation (2023). https://www.reuters.com/legal/transactional/ny-times-sues-openai-microsoft-infringing-copyrighted-work-2023-12-27/
2023
Later among the works it cites.
Li, Z., Wang, C., Wang, S., Gao, C.: Protecting intellectual property of large language model-based code generation apis via watermarks. In: Proceedings of the 2023 ACM SIGSAC Conference on Computer and Communications Security, pp. 2336–2350 (2023)
2023
Later among the works it cites.
Shanahan, M.: Talking about large language models. Communications of the ACM 67
2024
Closest in time.
Yin, Z., Wang, Z., Zhang, W.: Improving fairness in machine learning software via counterfactual fairness thinking. In: Proceedings of the 2024 IEEE/ACM 46th International Conference on Software Engineering: Companion Proceedings, pp. 420–421 (2024)
2024
Closest in time.
Saxena, N.A., Zhang, W., Shahabi, C.: Unveiling and mitigating bias in ride-hailing pricing for equitable policy making. AI and Ethics, 1–12 (2024)
2024
Closest in time.
Zhang, W.: Fairness with censorship: Bridging the gap between fairness research and real-world deployment. In: Proceedings of the AAAI Conference on Artificial Intelligence, vol. 38, pp. 22685–22685 (2024)
2024
Closest in time.
Chung, H.W., Hou, L., Longpre, S., Zoph, B., Tay, Y., Fedus, W., Li, Y., Wang, X., Dehghani, M., Brahma, S., et al
2024
Closest in time.
2024
Closest in time.
Yao, Y., Duan, J., Xu, K., Cai, Y., Sun, Z., Zhang, Y.: A survey on large language model (llm) security and privacy: The good, the bad, and the ugly. High-Confidence Computing, 100211 (2024)
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
He, J., Zhou, X., Xu, B., Zhang, T., Kim, K., Yang, Z., Thung, F., Irsan, I.C., Lo, D.: Representation learning for stack overflow posts: How far are we? ACM Transactions on Software Engineering and Methodology 33
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Wang, Z., Dzuong, J., Yuan, X., Chen, Z., Wu, Y., Yao, X., Zhang, W.: Individual fairness with group awareness under uncertainty. In: Joint European Conference on Machine Learning and Knowledge Discovery in Databases (2024). Springer Nature Switzerland
2024
Closest in time.
Wang, Z., Qiu, M., Chen, M., Salem, M.B., Yao, X., Zhang, W.: Toward fair graph neural networks via real counterfactual samples. Knowledge and Information Systems, 1–25 (2024)
2024
Closest in time.
Chu, Z., Wang, Z., Zhang, W.: Fairness in large language models: A taxonomic survey. ACM SIGKDD Explorations Newsletter, 2024, 34–48 (2024)
2024
Closest in time.
Doan, T.V., Wang, Z., Nguyen, M.N., Zhang, W.: Fairness in large language models in three hours. In: Proceedings of the 33rd ACM International Conference on Information & Knowledge Management (Boise, USA) (2024)
2024
Closest in time.
2024
Closest in time.
Zhang, W.: Ai fairness in practice: Paradigm, challenges, and prospects. Ai Magazine (2024)
2024
Closest in time.
Gallegos, I.O., Rossi, R.A., Barrow, J., Tanjim, M.M., Kim, S., Dernoncourt, F., Yu, T., Zhang, R., Ahmed, N.K.: Bias and fairness in large language models: A survey. Computational Linguistics, 1–79 (2024)
2024
Closest in time.
Wang, Z., Chu, Z., Blanco, R., Chen, Z., Chen, S.-C., Zhang, W.: Advancing graph counterfactual fairness through fair representation learning. In: Joint European Conference on Machine Learning and Knowledge Discovery in Databases (2024). Springer Nature Switzerland
2024
Closest in time.
2024
Closest in time.
Yazdani, S., Saxena, N., Wang, Z., Wu, Y., Zhang, W.: A comprehensive survey of image and video generative ai: Recent advances, variants, and applications (2024)
2024
Closest in time.