Fetching the paper…
Reading the bibliography…
As AI systems become increasingly prevalent and impactful, the need for effective AI governance and accountability measures is paramount.
M. Goddard, “The eu general data protection regulation (gdpr): European regulation that has a global impact,” International Journal of Market Research , vol. 59, no. 6, pp. 703–705, 2017
2017
Earlier work this paper cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “Bert: Pre-training of deep bidirectional transformers for language understanding,” 2019
2019
Earlier work this paper cites.
D. Schiff, J. Biddle, J. Borenstein, and K. Laas, “What’s next for ai ethics, policy, and governance? a global overview,” in Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society , ser. AIES ’20. New York, NY, USA: Association for Computing Machinery, 2020, p. 153–158. [Online]. Available: https://doi.org/10.1145/3375627.3375804
2020
Earlier work this paper cites.
I. D. Raji, A. Smart, R. N. White, M. Mitchell, T. Gebru, B. Hutchinson, J. Smith-Loud, D. Theron, and P. Barnes, “Closing the ai accountability gap: Defining an end-to-end framework for internal algorithmic auditing,” 2020
2020
Earlier work this paper cites.
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei, “Language models are few-shot learners,” 2020
2020
Earlier work this paper cites.
E. Arazo, D. Ortego, P. Albert, N. E. O’Connor, and K. McGuinness, “Pseudo-labeling and confirmation bias in deep semi-supervised learning,” in 2020 International joint conference on neural networks (IJCNN) . IEEE, 2020, pp. 1–8
2020
Earlier work this paper cites.
M. Mäntymäki, M. Minkkinen, T. Birkstedt, and M. Viljanen, “Defining organizational ai governance,” AI and Ethics , vol. 2, no. 4, pp. 603–609, 2022
2022
Earlier work this paper cites.
Y. Bai, S. Kadavath, S. Kundu, A. Askell, J. Kernion, A. Jones, A. Chen, A. Goldie, A. Mirhoseini, C. McKinnon, C. Chen, C. Olsson, C. Olah, D. Hernandez, D. Drain, D. Ganguli, D. Li, E. Tran-Johnson, E. Perez, J. Kerr, J. Mueller, J. Ladish, J. Landau, K. Ndousse, K. Lukosuite, L. Lovitt, M. Sellitto, N. Elhage, N. Schiefer, N. Mercado, N. DasSarma, R. Lasenby, R. Larson, S. Ringer, S. Johnston, S. Kravec, S. E. Showk, S. Fort, T. Lanham, T. Telleen-Lawton, T. Conerly, T. Henighan, T. Hume, S. R. Bowman, Z. Hatfield-Dodds, B. Mann, D. Amodei, N. Joseph, S. McCandlish, T. Brown, and J. Kaplan, “Constitutional AI: Harmlessness from AI Feedback,” arXiv , Dec. 2022
2022
Earlier work this paper cites.
A. Parrish, A. Chen, N. Nangia, V. Padmakumar, J. Phang, J. Thompson, P. M. Htut, and S. Bowman, “BBQ: A hand-built bias benchmark for question answering,” in Findings of the Association for Computational Linguistics: ACL 2022 , S. Muresan, P. Nakov, and A. Villavicencio, Eds. Dublin, Ireland: Association for Computational Linguistics, May 2022, pp. 2086–2105. [Online]. Available: https://aclanthology.org/2022.findings-acl.165
2022
Earlier work this paper cites.
Aug. 2022, [Online; accessed 28. Apr. 2024]. [Online]. Available: h t t p s : / / w w w . n i s t . g o v / s y s t e m / f i l e s / d o c u m e n t s / 2022 / 08 / 18 / A I R M F 2 n d d r a f t . p d f https://www.nist.gov/system/files/documents/2022/08/18/AI_{R}MF_{2}nd_{d}raft.pdf
2022
Earlier work this paper cites.
A. Parrish, A. Chen, N. Nangia, V. Padmakumar, J. Phang, J. Thompson, P. M. Htut, and S. Bowman, “BBQ: A hand-built bias benchmark for question answering,” ACL Anthology , pp. 2086–2105, May 2022
2022
Earlier work this paper cites.
T. Birkstedt, M. Minkkinen, A. Tandon, and M. Mäntymäki, “Ai governance: themes, knowledge gaps and future agendas,” Internet Research , vol. 33, no. 7, pp. 133–167, 2023
2023
Earlier work this paper cites.
E. Papagiannidis, I. M. Enholm, C. Dremel, P. Mikalef, and J. Krogstie, “Toward ai governance: Identifying best practices and potential barriers and outcomes,” Information Systems Frontiers , vol. 25, no. 1, pp. 123–141, 2023
2023
Earlier work this paper cites.
R. Bommasani, K. Klyman, S. Longpre, S. Kapoor, N. Maslej, B. Xiong, D. Zhang, and P. Liang, “The foundation model transparency index,” 2023
2023
Earlier work this paper cites.
J. Li, X. Cheng, X. Zhao, J.-Y. Nie, and J.-R. Wen, “HaluEval: A Large-Scale Hallucination Evaluation Benchmark for Large Language Models,” OpenReview , Dec. 2023. [Online]. Available: https://openreview.net/forum?id=bxsrykzSnq
2023
Cited alongside, same era.
Y. Huang and D. Xiong, “CBBQ: A Chinese Bias Benchmark Dataset Curated with Human-AI Collaboration for Large Language Models,” arXiv , Jun. 2023
2023
Cited alongside, same era.
V. K. Felkner, H.-C. H. Chang, E. Jang, and J. May, “WinoQueer: A Community-in-the-Loop Benchmark for Anti-LGBTQ+ Bias in Large Language Models,” arXiv , Jun. 2023
2023
Cited alongside, same era.
D. Yao, “Zoom Invests in Anthropic, Partners on Generative AI,” AI Business , May 2023. [Online]. Available: h t t p s : / / a i b u s i n e s s . c o m / n l p / z o o m − i n v e s t s − i n − a n t h r o p i c − p a r t n e r s − o n − g e n e r a t i v e − a i # c l o s e − m o d a l https://aibusiness.com/nlp/zoom-invests-in-anthropic-partners-on-generative-ai\#close-modal
V. K. Uppalapati and D. S. Nag, “A comparative analysis of ai models in complex medical decision-making scenarios: Evaluating chatgpt, claude ai, bard, and perplexity,” Cureus , vol. 16, no. 1, 2024
2024
Closest in time.
“Claude’s Constitution,” Apr. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://www.anthropic.com/news/claudes-constitution
2024
Closest in time.
United Nations, “Universal Declaration of Human Rights | | United Nations,” Apr. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://www.un.org/en/about-us/universal-declaration-of-human-rights
2024
Closest in time.
Sep. 2022, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://storage.googleapis.com/deepmind-media/DeepMind.com/Authors-Notes/sparrow/sparrow-final.pdf
2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
“Partnering with Scale to Bring Generative AI to Enterprises,” Apr. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://www.anthropic.com/news/partnering-with-scale
2024
Cited alongside, same era.
“Zoom Partnership and Investment in Anthropic,” Apr. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://www.anthropic.com/news/zoom-partnership-and-investment
2024
Cited alongside, same era.
“Anthropic partners with BCG,” Apr. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://www.anthropic.com/news/anthropic-bcg
2024
Cited alongside, same era.
“Accenture, AWS, Anthropic Collaboration,” Apr. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://www.anthropic.com/news/accenture-aws-anthropic
2024
Cited alongside, same era.
“SKT Partnership Announcement,” Apr. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://www.anthropic.com/news/skt-partnership-announcement
2024
Cited alongside, same era.
“Kief Studio & Anthropic Partnership: Transform Your Business with AI C,” Apr. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://kief.studio/anthropic
2024
Cited alongside, same era.
“AI Risk Management Framework | | NIST,” Jan. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://www.nist.gov/itl/ai-risk-management-framework
2024
Cited alongside, same era.
“Eu ai act: first regulation on artificial intelligence,” European Parliament, Apr. 2024. [Online]. Available: https://www.europarl.europa.eu/topics/en/article/20230601STO93804/eu-ai-act-first-regulation-on-artificial-intelligence
2024
Cited alongside, same era.
“Collective Constitutional AI: Aligning a Language Model with Public Input,” Apr. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://www.anthropic.com/news/collective-constitutional-ai-aligning-a-language-model-with-public-input
2024
Closest in time.
“pol.is report,” Apr. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://pol.is/report/r3rwrinr5udrzwkvxtdkj
2024
Closest in time.
Dec. 2023, [Online; accessed 28. Apr. 2024]. [Online]. Available: h t t p s : / / w w w − c d n . a n t h r o p i c . c o m / 65408 e e 2 b 9 c 99 a b e 53 e 432 f 300 e 7 f 43 e f 69 f b 6 e 4 / C C A I p u b l i c c o m p a r i s o n 2 023 . p d f https://www-cdn.anthropic.com/65408ee2b9c99abe53e432f300e7f43ef69fb6e4/CCAI_{p}ublic_{c}omparison_{2}023.pdf
2024
Closest in time.
G. Del Valle, “DHS testing out AI pilot programs for FEMA, ICE, and USCIS,” Verge , Mar. 2024. [Online]. Available: https://www.theverge.com/2024/3/18/24104843/dhs-ai-pilot-programs-chatgpt-openai-anthropic-meta
2024
Closest in time.
Mar. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://www.ntia.gov/sites/default/files/ntia-ai-report-final.pdf
2024
Closest in time.
“Anthropic’s Responsible Scaling Policy,” Apr. 2024, [Online; accessed 28. Apr. 2024]. [Online]. Available: https://www.anthropic.com/news/anthropics-responsible-scaling-policy
2024
Closest in time.
Z. Zhu, Z. Sun, and Y. Yang, “HaluEval-Wild: Evaluating Hallucinations of Language Models in the Wild,” arXiv , Mar. 2024
2024
Closest in time.
T. Yuan, Z. He, L. Dong, Y. Wang, R. Zhao, T. Xia, L. Xu, B. Zhou, F. Li, Z. Zhang, R. Wang, and G. Liu, “R-Judge: Benchmarking Safety Risk Awareness for LLM Agents,” arXiv , Jan. 2024
2024
Closest in time.
J. Lee, M. Kim, S. Kim, J. Kim, S. Won, H. Lee, and E. Choi, “KorNAT: LLM Alignment Benchmark for Korean Social Values and Common Knowledge,” arXiv , Feb. 2024
2024
Closest in time.