Fetching the paper…
Reading the bibliography…
The emergence of tools based on Large Language Models (LLMs), such as OpenAI's ChatGPT, Microsoft's Bing Chat, and Google's Bard, has garnered immense public attention.
Vulnerabilities of the online public square to manipulation
Truong, B. T., Lou, X., Flammini, A. & Menczer, F · 1907
Earlier work this paper cites.
A mathematical theory of communication
Shannon, C. E · 1948
Earlier work this paper cites.
Improving language understanding by generative pre-training (2018)
Radford, A., Narasimhan, K., Salimans, T., Sutskever, I. et al · 2018
Earlier work this paper cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Wang, A. et al · 2018
Earlier work this paper cites.
Language models are unsupervised multitask learners
Radford, A. et al · 2019
Earlier work this paper cites.
Physicians’ perceptions of chatbots in health care: Cross-sectional web-based survey
Palanica, A., Flaschner, P., Thommandram, A., Li, M. & Fossat, Y · 2019
Earlier work this paper cites.
The history of digital spam
Ferrara, E · 2019
Earlier work this paper cites.
SuperGLUE: A Stickier Benchmark for General-Purpose Language Understanding Systems
Wang, A. et al · 2019
Earlier work this paper cites.
Generalization in Generation: A closer look at Exposure Bias
Schmidt, F · 2019
Earlier work this paper cites.
MoverScore: Text generation evaluating with contextualized embeddings and earth mover distance
Zhao, W. et al · 2019
Earlier work this paper cites.
Language Models are Few-Shot Learners
Brown, T. et al · 2020
Earlier work this paper cites.
Judging truth
Brashier, N. M. & Marsh, E. J · 2020
Earlier work this paper cites.
Exposure to social engagement metrics increases vulnerability to misinformation
Avram, M., Micallef, N., Patil, S. & Menczer, F · 2020
Earlier work this paper cites.
Realm: retrieval-augmented language model pre-training
Guu, K., Lee, K., Tung, Z., Pasupat, P. & Chang, M.-W · 2020
Earlier work this paper cites.
Controlled hallucinations: learning to generate faithfully from noisy data
Filippova, K · 2020
Earlier work this paper cites.
Bertscore: Evaluating text generation with BERT
Zhang, T., Kishore, V., Wu, F., Weinberger, K. Q. & Artzi, Y · 2020
Earlier work this paper cites.
Did chatbots miss their “Apollo moment”? Potential, gaps, and lessons from using collaboration assistants during COVID-19
Srivastava, B · 2021
Earlier work this paper cites.
The Creation and Detection of Deepfakes: A Survey
Mirsky, Y. & Lee, W · 2021
Earlier work this paper cites.
Towards few-shot fact-checking via perplexity
Lee, N., Bang, Y., Madotto, A. & Fung, P · 2021
Earlier work this paper cites.
Neural path hunter: reducing hallucination in dialogue systems via path grounding
Dziri, N., Madotto, A., Zaïane, O. & Bose, A. J · 2021
Earlier work this paper cites.
Editing factual knowledge in language models
De Cao, N., Aziz, W. & Titov, I · 2021
Earlier work this paper cites.
Mitchell, E., Lin, C., Bosselut, A., Finn, C. & Manning, C. D · 2021
Earlier work this paper cites.
Infosurgeon: cross-media fine-grained information consistency checking for fake news detection
Fung, Y. et al · 2021
Earlier work this paper cites.
Adversarial deepfakes: Evaluating vulnerability of deepfake detectors to adversarial examples
Hussain, S., Neekhara, P., Jere, M., Koushanfar, F. & McAuley, J · 2021
Earlier work this paper cites.
Desyr: definition and syntactic representation based claim detection on the web
Sundriyal, M., Singh, P., Akhtar, M. S., Sengupta, S. & Chakraborty, T · 2021
Earlier work this paper cites.
Uncovering coordinated networks on social media: Methods and case studies
Pacheco, D. et al · 2021
Earlier work this paper cites.
Lamda: Language models for dialog applications
Thoppilan, R. et al · 2022
Earlier work this paper cites.
Emergent abilities of large language models
Wei, J. et al · 2022
Earlier work this paper cites.
Deep Learning Is Hitting a Wall (2022)
Marcus, G · 2022
Earlier work this paper cites.
TruthfulQA: Measuring how models mimic human falsehoods
Lin, S., Hilton, J. & Evans, O · 2022
Earlier work this paper cites.
Online misinformation is linked to early COVID-19 vaccination hesitancy and refusal
Pierri, F. et al · 2022
Earlier work this paper cites.
A survey of knowledge-enhanced text generation
Yu, W. et al · 2022
Earlier work this paper cites.
Memory-based model editing at scale
Mitchell, E., Lin, C., Bosselut, A., Manning, C. D. & Finn, C · 2022
Earlier work this paper cites.
Locating and editing factual associations in gpt
Meng, K., Bau, D., Andonian, A. & Belinkov, Y · 2022
Earlier work this paper cites.
Mass-editing memory in a transformer
Meng, K., Sharma, A. S., Andonian, A., Belinkov, Y. & Bau, D · 2022
Earlier work this paper cites.
NewsClaims: A new benchmark for claim detection from news with attribute knowledge
Gangi Reddy, R. et al · 2022
Earlier work this paper cites.
Empowering the fact-checkers! automatic identification of claim spans on Twitter
Sundriyal, M., Kulkarni, A., Pulastya, V., Akhtar, M. S. & Chakraborty, T · 2022
Earlier work this paper cites.
CONCRETE: Improving cross-lingual fact-checking with cross-lingual retrieval
Huang, K.-H., Zhai, C. & Ji, H · 2022
Earlier work this paper cites.
How would stance detection techniques evolve after the launch of chatgpt?
Zhang, B., Ding, D. & Jing, L · 2022
Earlier work this paper cites.
Towards Reasoning in Large Language Models: A Survey
Huang, J. & Chang, K. C.-C · 2023
Earlier work this paper cites.
OpenAI · 2023
Cited alongside, same era.
Llama: Open and efficient foundation language models
Touvron, H. et al · 2023
Cited alongside, same era.
https://openai.com/blog/chatgpt
Introducing ChatGPT — openai.com · 2023
Cited alongside, same era.
ChatGPT sets record for fastest-growing user base - analyst note
Hu, K · 2023
Cited alongside, same era.
Stanford Alpaca: An instruction-following LLaMA model
Taori, R. et al · 2023
Cited alongside, same era.
https://lmsys.org/blog/2023-03-30-vicuna/
Vicuna: An open-source chatbot impressing GPT-4 with 90%* ChatGPT quality · 2023
https://falconllm.tii.ae/falcon-180b.html
Falcon LLM — falconllm.tii.ae · 2023
Closest in time.
Halo effect — wikipedia.org
Volunteering, C · 2023
Closest in time.
Lawyer used ChatGPT in court—and cited fake cases. A judge is considering sanctions
OpenAI · 2023
Closest in time.
Judging the creative prowess of AI
Chakraborty, T. & Masud, S · 2023
Closest in time.
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Srivastava, A. et al · 2023
Closest in time.
FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation
Min, S. et al · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
https://www.anthropic.com/index/introducing-claude
Introducing Claude · 2023
Cited alongside, same era.
https://falconllm.tii.ae/falcon.html
Falcon · 2023
Cited alongside, same era.
https://docs.ai21.com/docs/jurassic-2-models
Jurassic-2 models · 2023
Cited alongside, same era.
Jais and Jais-chat: Arabic-centric foundation and instruction-tuned open generative large language models
Sengupta, N. et al · 2023
Cited alongside, same era.
A survey of large language models
Zhao, W. X. et al · 2023
Cited alongside, same era.
Open-domain hierarchical event schema induction by incremental prompting and verification
Li, S. et al · 2023
Cited alongside, same era.
Time Travel in LLMs: Tracing Data Contamination in Large Language Models (2023)
Golchin, S. & Surdeanu, M · 2023
Closest in time.
GPTScore: Evaluate as you desire (2023)
Fu, J., Ng, S.-K., Jiang, Z. & Liu, P · 2023
Closest in time.
G-Eval: NLG evaluation using GPT-4 with better human alignment (2023)
Liu, Y. et al · 2023
Closest in time.
SelfCheckGPT: Zero-resource black-box hallucination detection for generative large language models (2023)
Manakul, P., Liusie, A. & Gales, M. J. F · 2023
Closest in time.
Large language models are not fair evaluators (2023)
Wang, P. et al · 2023
Closest in time.
11% of data employees paste into ChatGPT is confidential - Cyberhaven — cyberhaven.com
Coles, C · 2023
Closest in time.
Do-not-answer: A dataset for evaluating safeguards in llms
Wang, Y., Li, H., Han, X., Nakov, P. & Baldwin, T · 2023
Closest in time.
Smartbook: AI-assisted situation report generation (2023)
Reddy, R. G. et al · 2023
Closest in time.
Critic: large language models can self-correct with tool-interactive critiquing
Gou, Z. et al · 2023
Closest in time.
Lm vs lm: detecting factual errors via cross examination
Cohen, R., Hamri, M., Geva, M. & Globerson, A · 2023
Closest in time.
Improving factuality and reasoning in language models through multiagent debate
Du, Y., Li, S., Torralba, A., Tenenbaum, J. B. & Mordatch, I · 2023
Closest in time.
Self information update for large language models through mitigating exposure bias
Yu, P. & Ji, H · 2023
Closest in time.
Security & privacy
OpenAI · 2023
Closest in time.
Privacy and data protection in chatgpt and other ai chatbots: Strategies for securing user information
Sebastian, G · 2023
Closest in time.
M4: Multi-generator, Multi-domain, and Multi-lingual Black-Box Machine-Generated Text Detection
Wang, Y. et al · 2023
Closest in time.
Faking fake news for real fake news detection: propaganda-loaded training data generation
Huang, K.-H., McKeown, K., Nakov, P., Choi, Y. & Ji, H · 2023
Closest in time.
Human detection of political speech deepfakes across transcripts, audio, and video (2023)
Groh, M. et al · 2023
Closest in time.
Can ai-generated text be reliably detected?
Sadasivan, V. S., Kumar, A., Balasubramanian, S., Wang, W. & Feizi, S · 2023
Closest in time.
https://contentauthenticity.org/
Content Authenticity Initiative — contentauthenticity.org · 2023
Closest in time.
https://www.garanteprivacy.it/home/docweb/-/docweb-display/docweb/9881490
ChatGPT: OpenAI reopens the platform in Italy guaranteeing more transparency and more rights to European users and non-users — tbs-sct.canada.ca · 2023
Closest in time.
https://www.ftc.gov/business-guidance/blog/2023/03/chatbots-deepfakes-voice-clones-ai-deception-sale
Chatbots, deepfakes, and voice clones: AI deception for sale — ftc.gov · 2023
Closest in time.
https://www.tbs-sct.canada.ca/pol/doc-eng.aspx?id=32592
Directive on Automated Decision-Making — tbs-sct.canada.ca · 2023
Closest in time.
Zero-shot faithful factual error correction
Huang, K.-H., Chan, H. P. & Ji, H · 2023
Closest in time.
ChatGPT: Jack of all trades, master of none
Kocoń, J. et al · 2023
Closest in time.
Remembering Conversations: Building Chatbots with Short and Long-Term Memory on AWS — itnext.io
Shankar, A · 2023
Closest in time.
Ferrara, E · 2023
Closest in time.
https://nap.nationalacademies.org/download/24623
Login | The National Academies Press — nap.nationalacademies.org · 2023
Closest in time.
ChatGPT banned in Italy over privacy concerns
McCallum, S · 2023
Closest in time.
https://www.universityworldnews.com/post.php?story=20230704155107330
New UK university principles promote AI literacy and integrity — universityworldnews.com · 2023
Closest in time.
Generative artificial intelligence in education departmental statement
Department for Education · 2023
Closest in time.
Right on Track: NVIDIA Open-Source Software Helps Developers Add Guardrails to AI Chatbots — blogs.nvidia.com
Cohen, J · 2023
Closest in time.
Accuracy of chatbots in citing journal articles
Chen, A. & Chen, D. O · 2023
Closest in time.
Introducing Microsoft 365 Copilot – your copilot for work - The Official Microsoft Blog — blogs.microsoft.com
Spataro, J · 2023
Closest in time.