Fetching the paper…
Reading the bibliography…
Climate decision making is constrained by the complexity and inaccessibility of key information within lengthy, technical, and multi-lingual documents.
The probabilistic relevance framework: Bm25 and beyond
Stephen Robertson and Hugo Zaragoza · 2009
Earlier work this paper cites.
Integrating cognitive load theory and concepts of human–computer interaction
Nina Hollender, Cristian Hofmann, Michael Deneke, and Bernhard Schmitz · 2010
Earlier work this paper cites.
What makes users trust a chatbot for customer service? an exploratory interview study
Asbjørn Følstad, Cecilie B. Nordheim, and Cato A. Bjørkli · 2018
Earlier work this paper cites.
Data quality and artificial intelligence – mitigating bias and error to protect fundamental rights
European Union Agency for Fundamental Rights · 2019
Earlier work this paper cites.
Best practices for the human evaluation of automatically generated text
Chris van der Lee, Albert Gatt, Emiel van Miltenburg, Sander Wubben, and Emiel Krahmer · 2019
Earlier work this paper cites.
Conversation considered harmful?
Stuart Reeves · 2019
Earlier work this paper cites.
Analyzing sustainability reports using natural language processing, 2020
Alexandra Luccioni, Emily Baylor, and Nicolas Duchene · 2020
Earlier work this paper cites.
Implementation of defense in depth strategy to secure industrial control system in critical infrastructures
Abdelghani Tschroub · 2020
Earlier work this paper cites.
With little power comes great responsibility, 2020
Dallas Card, Peter Henderson, Urvashi Khandelwal, Robin Jia, Kyle Mahowald, and Dan Jurafsky · 2020
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive nlp tasks, 2021
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen tau Yih, Tim Rocktäschel, Sebastian Riedel, and Douwe Kiela · 2021
Earlier work this paper cites.
Retrieval augmentation reduces hallucination in conversation, 2021
Kurt Shuster, Spencer Poff, Moya Chen, Douwe Kiela, and Jason Weston · 2021
Earlier work this paper cites.
“everyone wants to do the model work, not the data work”: Data cascades in high-stakes ai
Nithya Sambasivan, Shivani Kapania, Hannah Highfill, Diana Akrong, Praveen Paritosh, and Lora M Aroyo · 2021
Earlier work this paper cites.
Human evaluation of automatically generated text: Current trends and best practice guidelines
Chris van der Lee, Albert Gatt, Emiel van Miltenburg, and Emiel Krahmer · 2021
Earlier work this paper cites.
Anticipating safety issues in e2e conversational ai: Framework and tooling, 2021
Emily Dinan, Gavin Abercrombie, A. Stevie Bergman, Shannon Spruit, Dirk Hovy, Y-Lan Boureau, and Verena Rieser · 2021
Earlier work this paper cites.
Training language models to follow instructions with human feedback, 2022
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul Christiano, Jan Leike, and Ryan Lowe · 2022
Earlier work this paper cites.
Cheap talk and cherry-picking: What climatebert has to say on corporate climate risk disclosures
Julia Anna Bingler, Mathias Kraus, Markus Leippold, and Nicolas Webersinke · 2022
Earlier work this paper cites.
Red teaming language models with language models, 2022
Ethan Perez, Saffron Huang, Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving · 2022
Earlier work this paper cites.
On the safety of conversational models: Taxonomy, dataset, and benchmark
Hao Sun, Guangxuan Xu, Jiawen Deng, Jiale Cheng, Chujie Zheng, Hao Zhou, Nanyun Peng, Xiaoyan Zhu, and Minlie Huang · 2022
Earlier work this paper cites.
SafetyKit: First aid for measuring safety in open-domain conversational systems
Emily Dinan, Gavin Abercrombie, A. Bergman, Shannon Spruit, Dirk Hovy, Y-Lan Boureau, and Verena Rieser · 2022
Earlier work this paper cites.
True: Re-evaluating factual consistency evaluation, 2022
Or Honovich, Roee Aharoni, Jonathan Herzig, Hagai Taitelbaum, Doron Kukliansy, Vered Cohen, Thomas Scialom, Idan Szpektor, Avinatan Hassidim, and Yossi Matias · 2022
Earlier work this paper cites.
Streamlining evaluation with ir-measures
Sean MacAvaney, Craig Macdonald, and Iadh Ounis · 2022
Cited alongside, same era.
From human writing to artificial intelligence generated text: examining the prospects and potential threats of ChatGPT in academic writing
Ismail Dergaa, Karim Chamari, Piotr Zmijewski, and Helmi Ben Saad · 2023
Cited alongside, same era.
A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions, 2023
Lei Huang, Weijiang Yu, Weitao Ma, Weihong Zhong, Zhangyin Feng, Haotian Wang, Qianglong Chen, Weihua Peng, Xiaocheng Feng, Bing Qin, and Ting Liu · 2023
Cited alongside, same era.
Lima: Less is more for alignment, 2023
Chunting Zhou, Pengfei Liu, Puxin Xu, Srini Iyer, Jiao Sun, Yuning Mao, Xuezhe Ma, Avia Efrat, Ping Yu, Lili Yu, Susan Zhang, Gargi Ghosh, Mike Lewis, Luke Zettlemoyer, and Omer Levy · 2023
Cited alongside, same era.
Progress on climate action: a multilingual machine learning analysis of the global stocktake
Anne J Sietsma, Rick W Groenendijk, and Robbert Biesbroek · 2023
Cited alongside, same era.
Gemini: A family of highly capable multimodal models, 2024
Gemini Team, Anil, R., Borgeaud, S., Alayrac, J.-B., Yu, J., Soricut, R., Schalkwyk, J., Dai, A. M., Hauth, A., Millican, K., Silver, D., Johnson, M., Antonoglou, I., Schrittwieser, J., Glaese, A., Chen, J., Pitler, E., Lillicrap, T., Lazaridou, A., Firat, O., …, Vinyals, O · 2024
Closest in time.
Hallucination-free? assessing the reliability of leading ai legal research tools, 2024
Varun Magesh, Faiz Surani, Matthew Dahl, Mirac Suzgun, Christopher D. Manning, and Daniel E. Ho · 2024
Closest in time.
Identifying climate targets in national laws and policies using machine learning, 2024
Matyas Juhasz, Tina Marchand, Roshan Melwani, Kalyan Dutia, Sarah Goodenough, Harrison Pim, and Henry Franks · 2024
Closest in time.
Climategpt: Towards ai synthesizing interdisciplinary research on climate change, 2024
David Thulke, Yingbo Gao, Petrus Pelser, Rein Brune, Rricha Jalota, Floris Fok, Michael Ramos, Ian van Wyk, Abdallah Nasir, Hayden Goldstein, Taylor Tragemann, Katie Nguyen, Ariana Fowler, Andrew Stanco, Jon Gabriel, Jordan Taylor, Dean Moro, Evgenii Tsymbalov, Juliette de Waal, Evgeny Matusov, Mudar Yaghi, Mohammad Shihadah, Hermann Ney, Christian Dugast, Jonathan Dotan, and Daniel Erasmus · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
chatclimate: Grounding conversational ai in climate science, 2023
Saeid Ashraf Vaghefi, Qian Wang, Veruska Muccione, Jingwei Ni, Mathias Kraus, Julia Bingler, Tobias Schimanski, Chiara Colesanti-Senni, Nicolas Webersinke, Christrian Huggel, and Markus Leippold · 2023
Cited alongside, same era.
Measuring data, 2023
Margaret Mitchell, Alexandra Sasha Luccioni, Nathan Lambert, Marissa Gerchick, Angelina McMillan-Major, Ezinwanne Ozoani, Nazneen Rajani, Tristan Thrush, Yacine Jernite, and Douwe Kiela · 2023
Cited alongside, same era.
Data-centric artificial intelligence: A survey, 2023
Daochen Zha, Zaid Pervaiz Bhat, Kwei-Herng Lai, Fan Yang, Zhimeng Jiang, Shaochen Zhong, and Xia Hu · 2023
Cited alongside, same era.
Towards answering climate questionnaires from unstructured climate reports, 2023
Daniel Spokoyny, Tanmay Laud, Tom Corringham, and Taylor Berg-Kirkpatrick · 2023
Cited alongside, same era.
Jailbroken: How does llm safety training fail?
Alexander Wei, Nika Haghtalab, and Jacob Steinhardt · 2023
Cited alongside, same era.
Guardrails ai, 2023
S. Rajpal · 2023
Cited alongside, same era.
Judging llm-as-a-judge with mt-bench and chatbot arena, 2023
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric P. Xing, Hao Zhang, Joseph E. Gonzalez, and Ion Stoica · 2023
Cited alongside, same era.
Angel Hsu, Mason Laney, Ji Zhang, Diego Manya, and Linda Farczadi · 2024
Closest in time.
(a)i am not a lawyer, but…: Engaging legal experts towards responsible llm policies for legal advice, 2024
Inyoung Cheong, King Xia, K. J. Kevin Feng, Quan Ze Chen, and Amy X. Zhang · 2024
Closest in time.
Towards faithful and robust llm specialists for evidence-based question-answering, 2024
Tobias Schimanski, Jingwei Ni, Mathias Kraus, Elliott Ash, and Markus Leippold · 2024
Closest in time.
Jailbreaking leading safety-aligned llms with simple adaptive attacks, 2024
Maksym Andriushchenko, Francesco Croce, and Nicolas Flammarion · 2024
Closest in time.
Climate policy radar database
Climate Policy Radar · 2024
Closest in time.
rag-climate-expert-eval (revision b282d48), 2024
Climate Policy Radar · 2024
Closest in time.
Dapr: A benchmark on document-aware passage retrieval, 2024
Kexin Wang, Nils Reimers, and Iryna Gurevych · 2024
Closest in time.
NextLevelBERT: Masked language modeling with higher-level representations for long documents
Tamara Czinczoll, Christoph Hönes, Maximilian Schall, and Gerard De Melo · 2024
Closest in time.
Colpali: Efficient document retrieval with vision language models, 2024
Manuel Faysse, Hugues Sibille, Tony Wu, Bilel Omrani, Gautier Viaud, Céline Hudelot, and Pierre Colombo · 2024
Closest in time.
Statements: Universal information extraction from tables with large language models for ESG KPIs
Lokesh Mishra, Sohayl Dhibi, Yusik Kim, Cesar Berrospi Ramis, Shubham Gupta, Michele Dolfi, and Peter Staar · 2024
Closest in time.
Chatgpt (gpt-4o version), 2024
OpenAI · 2024
Closest in time.
Lynx: An open source hallucination evaluation model, 2024
Selvan Sunitha Ravi, Bartosz Mielczarek, Anand Kannappan, Douwe Kiela, and Rebecca Qian · 2024
Closest in time.
Llm evaluators recognize and favor their own generations, 2024
Arjun Panickssery, Samuel R. Bowman, and Shi Feng · 2024
Closest in time.
Zhi Rui Tam, Cheng-Kuang Wu, Yi-Lin Tsai, Chieh-Yen Lin, Hung yi Lee, and Yun-Nung Chen · 2024
Closest in time.
Dual use concerns of generative ai and large language models
Alexei Grinbaum and Laurynas Adomaitis · 2024
Closest in time.
Perceived conversational ability of task-based chatbots – which conversational elements influence the success of text-based dialogues?
Alexandra Rese and Pauline Tränkner · 2024
Closest in time.