Fetching the paper…
Reading the bibliography…
Social media abounds with multimodal sarcasm, and identifying sarcasm targets is particularly challenging due to the implicit incongruity not directly evident in the text and image modalities.
Language models are few-shot learners
Tom B Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Yolov4: Optimal speed and accuracy of object detection
Alexey Bochkovskiy, Chien-Yao Wang, and Hong-Yuan Mark Liao. 2020 · 2004
Earlier work this paper cites.
Detection of harassment on web 2.0
Dawei Yin, Zhenzhen Xue, Liangjie Hong, Brian D Davison, April Kontostathis, Lynne Edwards, et al. 2009 · 2009
Earlier work this paper cites.
Semi-supervised recognition of sarcasm in twitter and amazon
Dmitry Davidov, Oren Tsur, and Ari Rappoport. 2010 · 2010
Earlier work this paper cites.
Sarcasm as contrast between a positive sentiment and negative situation
Ellen Riloff, Ashequl Qadir, Prafulla Surve, Lalindra De Silva, Nathan Gilbert, and Ruihong Huang. 2013 · 2013
Earlier work this paper cites.
Convolutional neural networks for sentence classification
Yoon Kim. 2014 · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick. 2014 · 2014
Earlier work this paper cites.
Fast r-cnn
Ross Girshick. 2015 · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Detecting sarcasm in multimodal social platforms
Rossano Schifanella, Paloma De Juan, Joel Tetreault, and Liangliang Cao. 2016 · 2016
Earlier work this paper cites.
Mask r-cnn
Kaiming He, Georgia Gkioxari, Piotr Dollár, and Ross Girshick. 2017 · 2017
Earlier work this paper cites.
Automatic sarcasm detection: A survey
Aditya Joshi, Pushpak Bhattacharyya, and Mark J Carman. 2017 · 2017
Earlier work this paper cites.
Sarcasm target identification: Dataset and an introductory approach
Aditya Joshi, Pranav Goel, Pushpak Bhattacharyya, and Mark Carman. 2018 · 2018
Earlier work this paper cites.
Multi-modal sarcasm detection in twitter with hierarchical fusion model
Yitao Cai, Huiyu Cai, and Xiaojun Wan. 2019 · 2019
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Overview of the 2019 alta shared task: Sarcasm target identification
Diego Molla and Aditya Joshi. 2019 · 2019
Earlier work this paper cites.
Detecting target of sarcasm using ensemble methods
Pradeesh Parameswaran, Andrew Trotman, Veronica Liesaputra, and David Eyers. 2019 · 2019
Earlier work this paper cites.
A deep-learning framework to detect sarcasm targets
Jasabanta Patro, Srijan Bansal, and Animesh Mukherjee. 2019 · 2019
Earlier work this paper cites.
Generalized intersection over union: A metric and a loss for bounding box regression
Hamid Rezatofighi, Nathan Tsoi, JunYoung Gwak, Amir Sadeghian, Ian Reid, and Silvio Savarese. 2019 · 2019
Earlier work this paper cites.
Sarcasm detection with self-matching networks and low-rank bilinear pooling
Tao Xiong, Peiran Zhang, Hongbo Zhu, and Yihui Yang. 2019 · 2019
Cited alongside, same era.
Affective and contextual embedding for sarcasm detection
Nastaran Babanejad, Heidar Davoudi, Aijun An, and Manos Papagelis. 2020 · 2020
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al. 2020 · 2020
Cited alongside, same era.
Unsupervised evaluation of interactive dialog with dialogpt
Shikib Mehri and Maxine Eskenazi. 2020 · 2020
Cited alongside, same era.
Modeling intra and inter-modality incongruity for multi-modal sarcasm detection
Hongliang Pan, Zheng Lin, Peng Fu, Yatao Qi, and Weiping Wang. 2020 · 2020
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed H Chi, Quoc V Le, Denny Zhou, et al. 2022 · 2022
Later among the works it cites.
Glm-130b: An open bilingual pre-trained model
Aohan Zeng, Xiao Liu, Zhengxiao Du, Zihan Wang, Hanyu Lai, Ming Ding, Zhuoyi Yang, Yifan Xu, Wendi Zheng, Xiao Xia, et al. 2022 · 2022
Later among the works it cites.
Dino: Detr with improved denoising anchor boxes for end-to-end object detection
Hao Zhang, Feng Li, Shilong Liu, Lei Zhang, Hang Su, Jun Zhu, Lionel M Ni, and Heung-Yeung Shum. 2022 · 2022
Later among the works it cites.
Qwen-vl: A frontier large vision-language model with versatile abilities
Jinze Bai, Shuai Bai, Shusheng Yang, Shijie Wang, Sinan Tan, Peng Wang, Junyang Lin, Chang Zhou, and Jingren Zhou. 2023 · 2023
Later among the works it cites.
Instructblip: Towards general-purpose vision-language models with instruction tuning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Cited alongside, same era.
Reasoning with multimodal sarcastic tweets via modeling cross-modality contrast and semantic association
Nan Xu, Zhixiong Zeng, and Wenji Mao. 2020 · 2020
Cited alongside, same era.
Summeval: Re-evaluating summarization evaluation
Alexander R Fabbri, Wojciech Kryściński, Bryan McCann, Caiming Xiong, Richard Socher, and Dragomir Radev. 2021 · 2021
Cited alongside, same era.
Multi-modal sarcasm detection with interactive in-modal and cross-modal graphs
Bin Liang, Chenwei Lou, Xiang Li, Lin Gui, Min Yang, and Ruifeng Xu. 2021 · 2021
Cited alongside, same era.
Rumor detection on twitter with claim-guided hierarchical graph attention networks
Hongzhan Lin, Jing Ma, Mingfei Cheng, Zhiwei Yang, Liangliang Chen, and Guang Chen. 2021 · 2021
Cited alongside, same era.
Swin transformer: Hierarchical vision transformer using shifted windows
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo. 2021 · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al. 2021 · 2021
Cited alongside, same era.
Wenliang Dai, Junnan Li, Dongxu Li, Anthony Meng Huat Tiong, Junqi Zhao, Weisheng Wang, Boyang Albert Li, Pascale Fung, and Steven C. H. Hoi. 2023 · 2023
Later among the works it cites.
Bad actor, good advisor: Exploring the role of large language models in fake news detection
Beizhe Hu, Qiang Sheng, Juan Cao, Yuhui Shi, Yang Li, Danding Wang, and Peng Qi. 2023 · 2023
Later among the works it cites.
Is chatgpt better than human annotators? potential and limitations of chatgpt in explaining implicit hate speech
Fan Huang, Haewoon Kwak, and Jisun An. 2023 · 2023
Later among the works it cites.
Beneath the surface: Unveiling harmful memes with multimodal reasoning distilled from large language models
Hongzhan Lin, Ziyang Luo, Jing Ma, and Long Chen. 2023a · 2023
Later among the works it cites.
Wizardcoder: Empowering code large language models with evol-instruct
Ziyang Luo, Can Xu, Pu Zhao, Qingfeng Sun, Xiubo Geng, Wenxiang Hu, Chongyang Tao, Jing Ma, Qingwei Lin, and Daxin Jiang. 2023 · 2023
Later among the works it cites.
OpenAI. 2023 · 2023
Later among the works it cites.
Mmsd2. 0: Towards a reliable multi-modal sarcasm detection system
Libo Qin, Shijue Huang, Qiguang Chen, Chenran Cai, Yudi Zhang, Bin Liang, Wanxiang Che, and Ruifeng Xu. 2023 · 2023
Later among the works it cites.
Gemini: A family of highly capable multimodal models
Gemini Team, Rohan Anil, Sebastian Borgeaud, Yonghui Wu, Jean-Baptiste Alayrac, Jiahui Yu, Radu Soricut, Johan Schalkwyk, Andrew M Dai, Anja Hauth, et al. 2023 · 2023
Later among the works it cites.
A comprehensive review of yolo: From yolov1 to yolov8 and beyond
Juan Terven and Diana Cordova-Esparza. 2023 · 2023
Later among the works it cites.
Dynamic routing transformer network for multimodal sarcasm detection
Yuan Tian, Nan Xu, Ruike Zhang, and Wenji Mao. 2023 · 2023
Later among the works it cites.
Cogvlm: Visual expert for pretrained language models
Weihan Wang, Qingsong Lv, Wenmeng Yu, Wenyi Hong, Ji Qi, Yan Wang, Junhui Ji, Zhuoyi Yang, Lei Zhao, Xixuan Song, et al. 2023 · 2023
Later among the works it cites.
Wizardlm: Empowering large language models to follow complex instructions
Can Xu, Qingfeng Sun, Kai Zheng, Xiubo Geng, Pu Zhao, Jiazhan Feng, Chongyang Tao, and Daxin Jiang. 2023 · 2023
Later among the works it cites.
Wsdms: Debunk fake news via weakly supervised detection of misinforming sentences with contextualized social wisdom
Ruichao Yang, Wei Gao, Jing Ma, Hongzhan Lin, and Zhiwei Yang. 2023a · 2023
Later among the works it cites.
Towards explainable harmful meme detection through multimodal debate between large language models
Hongzhan Lin, Ziyang Luo, Wei Gao, Jing Ma, Bo Wang, and Ruichao Yang. 2024a · 2024
Closest in time.
Framewise phoneme classification with bidirectional lstm networks
Alex Graves and Jürgen Schmidhuber. 2005 · 2052
Closest in time.