Fetching the paper…
Reading the bibliography…
We introduce Fraud-R1, a benchmark designed to evaluate LLMs' ability to defend against internet fraud and phishing in dynamic, real-world scenarios.
An intelligent classification model for phishing email detection
Adwan Yasin and Abdelmunem Abuhasan. 2016 · 2016
Earlier work this paper cites.
Fake job recruitment detection using machine learning approach
Shawni Dutta and Samir Kumar Bandyopadhyay. 2020 · 2020
Earlier work this paper cites.
Fraud dataset benchmark and applications
Prince Grover, Julia Xu, Justin Tittelfitz, Anqi Cheng, Zheng Li, Jakub Zablocki, Jianbo Liu, and Hao Zhou. 2022 · 2022
Earlier work this paper cites.
Introducing chatgpt
OpenAI. 2022 · 2022
Earlier work this paper cites.
Trustworthy llms: a survey and guideline for evaluating large language models’ alignment
Yang Liu, Yuanshun Yao, Jean-Francois Ton, Xiaoying Zhang, Ruocheng Guo, Hao Cheng, Yegor Klochkov, Muhammad Faaiz Taufiq, and Hang Li. 2023 · 2023
Earlier work this paper cites.
The evolution of the nigerian prince scam
Ojeifoh Okosun and Uchenna Ilo. 2023 · 2023
Earlier work this paper cites.
Adversarial data poisoning for fake news detection: How to make a model misclassify a target news without modifying it
Federico Siciliano, Luca Maiano, Lorenzo Papa, Federica Baccini, Irene Amerini, and Fabrizio Silvestri. 2023 · 2023
Earlier work this paper cites.
Fake news detectors are biased against texts generated by large language models
Jinyan Su, Terry Yue Zhuo, Jonibek Mansurov, Di Wang, and Preslav Nakov. 2023 · 2023
Earlier work this paper cites.
An llm can fool itself: A prompt-based adversarial attack
Xilie Xu, Keyi Kong, Ning Liu, Lizhen Cui, Di Wang, Jingfeng Zhang, and Mohan Kankanhalli. 2023 · 2023
Earlier work this paper cites.
A study on the discourse strategy of telecommunication fraud based on proximization theory
Hong Ye and Kexin Chen. 2023 · 2023
Earlier work this paper cites.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing, et al. 2023 · 2023
Earlier work this paper cites.
Novel interpretable and robust web-based ai platform for phishing email detection
Abdulla Al-Subaiey, Mohammed Al-Thani, Naser Abdullah Alam, Kaniz Fatema Antora, Amith Khandakar, and SM Ashfaq Uz Zaman. 2024 · 2024
Earlier work this paper cites.
Detoxbench: Benchmarking large language models for multitask fraud & abuse detection
Joymallya Chakraborty, Wei Xia, Anirban Majumder, Dan Ma, Walid Chaabene, and Naveed Janvekar. 2024 · 2024
Earlier work this paper cites.
A complete survey on llm-based ai chatbots
Sumit Kumar Dam, Choong Seon Hong, Yu Qiao, and Chaoning Zhang. 2024 · 2024
Cited alongside, same era.
Chatglm: A family of large language models from glm-130b to glm-4 all tools
Team GLM, :, Aohan Zeng, Bin Xu, Bowen Wang, Chenhui Zhang, and Da Yin. 2024 · 2024
Cited alongside, same era.
Jiawei Gu, Xuhui Jiang, Zhichao Shi, Hexiang Tan, Xuehao Zhai, Chengjin Xu, Wei Li, Yinghan Shen, Shengjie Ma, Honghao Liu, et al. 2024 · 2024
Cited alongside, same era.
Large language models meet collaborative filtering: An efficient all-round llm-based recommender system
Sein Kim, Hongseok Kang, Seungyoon Choi, Donghyun Kim, Minchul Yang, and Chanyoung Park. 2024 · 2024
Cited alongside, same era.
Personalllm: Tailoring llms to individual preferences
Thomas P Zollo, Andrew Wei Tung Siah, Naimeng Ye, Ang Li, and Hongseok Namkoong · 2024
Later among the works it cites.
Claude 3.5 haiku
Anthropic. 2025a · 2025
Closest in time.
Claude sonnet
Anthropic. 2025b · 2025
Closest in time.
Real or fake? fake job posting prediction dataset
Shivam Bansal. 2019 · 2025
Closest in time.
Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning
DeepSeek-AI. 2025a · 2025
Closest in time.
Deepseek-v3 technical report
DeepSeek-AI. 2025b · 2025
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jean Lee, Nicholas Stevens, Soyeon Caren Han, and Minseok Song. 2024 · 2024
Cited alongside, same era.
Large language models for generative recommendation: A survey and visionary discussions
Lei Li, Yongfeng Zhang, Dugang Liu, and Li Chen. 2024a · 2024
Cited alongside, same era.
Combining fine-tuning and llm-based agents for intuitive smart contract auditing with justifications
Wei Ma, Daoyuan Wu, Yuqiang Sun, Tianwen Wang, Shangqing Liu, Jian Zhang, Yue Xue, and Yang Liu. 2024 · 2024
Cited alongside, same era.
OpenAI, :, Aaron Hurst, Adam Lerer, Adam P. Goucher, Adam Perelman, Aditya Ramesh, Aidan Clark, and AJ Ostrow. 2024 · 2024
Cited alongside, same era.
Investigating llm applications in e-commerce
Chester Palen-Michel, Ruixiang Wang, Yipeng Zhang, David Yu, Canran Xu, and Zhe Wu. 2024 · 2024
Cited alongside, same era.
Securing large language models: Addressing bias, misinformation, and prompt attacks
Benji Peng, Keyu Chen, Ming Li, Pohsun Feng, Ziqian Bi, Junyu Liu, and Qian Niu. 2024 · 2024
Cited alongside, same era.
Two tales of persona in LLMs: A survey of role-playing and personalization
Yu-Min Tseng, Yu-Chao Huang, Teng-Yun Hsiao, Wei-Lin Chen, Chao-Wei Huang, Yu Meng, and Yun-Nung Chen. 2024 · 2024
Cited alongside, same era.
Yangyang Yu, Zhiyuan Yao, Haohang Li, Zhiyang Deng, Yupeng Cao, Zhi Chen, Jordan W Suchow, Rong Liu, Zhenyu Cui, Zhaozhuo Xu, et al. 2024 · 2024
Cited alongside, same era.
Shaopeng Fu, Liang Ding, and Di Wang. 2025 · 2025
Closest in time.
Gemini gemma developer updates (may 2024)
Google. 2024a · 2025
Closest in time.
Google gemini: Next-generation model (february 2024)
Google. 2024b · 2025
Closest in time.
Meta llama 3.1
Meta. 2025 · 2025
Closest in time.
Openai o3-mini
OpenAI. 2025 · 2025
Closest in time.
Can large language model analyze financial statements well?
Xinlin Wang and Mats Brorsson. 2025 · 2025
Closest in time.