Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have advanced various Natural Language Processing (NLP) tasks, such as text generation and translation, among others.
Fleiss, J.L.: Measuring nominal scale agreement among many raters. Psychological bulletin 76
1971
Earlier work this paper cites.
Kim, T.K.: T test as a parametric statistic. Korean journal of anesthesiology 68
2015
Earlier work this paper cites.
2017
Earlier work this paper cites.
Ross, A., Willson, V.L.: One-sample T-test. In: Basic and Advanced Statistical Tests, pp. 9–12. Brill, ??? (2017)
2017
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., Sutskever, I., et al
2019
Earlier work this paper cites.
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J.D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Wolf, T., Debut, L., Sanh, V., Chaumond, J., Delangue, C., Moi, A., Cistac, P., Rault, T., Louf, R., Funtowicz, M., Davison, J., Shleifer, S., Platen, P., Ma, C., Jernite, Y., Plu, J., Xu, C., Le Scao, T., Gugger, S., Drame, M., Lhoest, Q., Rush, A.: Transformers: State-of-the-art natural language processing. In: Liu, Q., Schlangen, D. (eds.) Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations, pp. 38–45. Association for Computational Linguistics, Online (2020). https://doi.org/10.18653/v1/2020.emnlp-demos.6 . https://aclanthology.org/2020.emnlp-demos.6
2020
Earlier work this paper cites.
Bender, E.M., Gebru, T., McMillan-Major, A., Shmitchell, S.: On the dangers of stochastic parrots: Can language models be too big? In: Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency, pp. 610–623 (2021)
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
Ouyang, L., Wu, J., Jiang, X., Almeida, D., Wainwright, C., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., et al
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Dodge, J., Prewitt, T., Combes, R., Odmark, E., Schwartz, R., Strubell, E., Luccioni, A.S., Smith, N.A., DeCario, N., Buchanan, W.: Measuring the carbon intensity of AI in cloud instances. In: Proceedings of the 2022 ACM Conference on Fairness, Accountability, and Transparency, pp. 1877–1894 (2022)
2022
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Guardrails: Guardrails AI | Your Enterprise AI needs Guardrails — guardrailsai.com (2024). https://www.guardrailsai.com/docs/ Accessed 2024-02-01
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Wang, B., Chen, W., Pei, H., Xie, C., Kang, M., Zhang, C., Xu, C., Xiong, Z., Dutta, R., Schaeffer, R., et al.: Decodingtrust: A comprehensive assessment of trustworthiness in gpt models. Advances in Neural Information Processing Systems 36
2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Smith, E.M., Hall, M., Kambadur, M., Presani, E., Williams, A.: “I’m sorry to hear that”: Finding New Biases in Language Models with a Holistic Descriptor Dataset. In: Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing, pp. 9180–9211. Association for Computational Linguistics, Abu Dhabi, United Arab Emirates (2022). https://doi.org/10.18653/v1/2022.emnlp-main.625 . https://aclanthology.org/2022.emnlp-main.625
2023
Cited alongside, same era.
Hartvigsen, T., Gabriel, S., Palangi, H., Sap, M., Ray, D., Kamar, E.: ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection. In: Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pp. 3309–3326. Association for Computational Linguistics, Dublin, Ireland (2022). https://doi.org/10.18653/v1/2022.acl-long.234 . https://aclanthology.org/2022.acl-long.234
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Closest in time.
Gallegos, I.O., Rossi, R.A., Barrow, J., Tanjim, M.M., Kim, S., Dernoncourt, F., Yu, T., Zhang, R., Ahmed, N.K.: Bias and fairness in large language models: A survey. Computational Linguistics, 1–79 (2024)
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Raza, S., Bamgbose, O., Ghuge, S., Pandya, D.: Safe and sound: Evaluating language models for bias mitigation and understanding. In: Neurips Safe Generative AI Workshop 2024 (2024)
2024
Closest in time.
2024
Closest in time.
Nadeem, M., Bethke, A., Reddy, S.: StereoSet: Measuring stereotypical bias in pretrained language models. In: Zong, C., Xia, F., Li, W., Navigli, R. (eds.) Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers), pp. 5356–5371. Association for Computational Linguistics, Online (2021). https://doi.org/10.18653/v1/2021.acl-long.416 . https://aclanthology.org/2021.acl-long.416
2024
Closest in time.
Chung, H.W., Hou, L., Longpre, S., Zoph, B., Tay, Y., Fedus, W., Li, Y., Wang, X., Dehghani, M., Brahma, S., et al
2024
Closest in time.
API, P.: Perspective API (2024). https://www.perspectiveapi.com/
2024
Closest in time.
OpenAI: Moderation - OpenAI API (2024). https://platform.openai.com/docs/guides/moderation
2024
Closest in time.
AI, C.: Confident AI Documentation. [Online; accessed 10-May-2024] (2024). https://docs.confident-ai.com/docs/getting-started
2024
Closest in time.
2024
Closest in time.