Fetching the paper…
Reading the bibliography…
Internet memes have become a powerful means for individuals to express emotions, thoughts, and perspectives on social media.
Visualbert: A simple and performant baseline for vision and language
Liunian Harold Li, Mark Yatskar, Da Yin, Cho-Jui Hsieh, and Kai-Wei Chang. 2019 · 1908
Earlier work this paper cites.
Supervised multimodal bitransformers for classifying images and text
Douwe Kiela, Suvrat Bhooshan, Hamed Firooz, Ethan Perez, and Davide Testuggine. 2019 · 1909
Earlier work this paper cites.
Hate speech in pixels: Detection of offensive memes towards automatic moderation
Benet Oriol Sabat, Cristian Canton Ferrer, and Xavier Giro-i Nieto. 2019 · 1910
Earlier work this paper cites.
A coefficient of agreement for nominal scales
Jacob Cohen. 1960 · 1960
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al. 2020 · 2010
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Information retrieval using cosine and jaccard similarity measures in vector space model
Abhishek Jain, Aman Jain, Nihal Chauhan, Vikrant Singh, and Narina Thakur. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Jiasen Lu, Dhruv Batra, Devi Parikh, and Stefan Lee. 2019 · 2019
Earlier work this paper cites.
A simple framework for contrastive learning of visual representations
Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey Hinton. 2020 · 2020
Earlier work this paper cites.
Unsupervised cross-lingual representation learning at scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Edouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov. 2020 · 2020
Earlier work this paper cites.
Exploring hate speech detection in multimodal publications
Raul Gomez, Jaume Gibert, Lluis Gomez, and Dimosthenis Karatzas. 2020 · 2020
Earlier work this paper cites.
The hateful memes challenge: Detecting hate speech in multimodal memes
Douwe Kiela, Hamed Firooz, Aravind Mohan, Vedanuj Goswami, Amanpreet Singh, Pratik Ringshia, and Davide Testuggine. 2020 · 2020
Cited alongside, same era.
Banglabert: Bengali mask language model for bengali language understanding
Sagor Sarker. 2020 · 2020
Cited alongside, same era.
Multimodal meme dataset (multioff) for identifying offensive content in image and text
Shardul Suryawanshi, Bharathi Raja Chakravarthi, Mihael Arcan, and Paul Buitelaar. 2020 · 2020
Cited alongside, same era.
“subverting the jewtocracy”: Online antisemitism detection using multimodal deep learning
Mohit Chandra, Dheeraj Pailla, Himanshu Bhatia, Aadilmehdi Sanchawala, Manish Gupta, Manish Shrivastava, and Ponnurangam Kumaraguru. 2021 · 2021
Cited alongside, same era.
Disentangling hate in online memes
Roy Ka-Wei Lee, Rui Cao, Ziqing Fan, Jing Jiang, and Wen-Haw Chong. 2021 · 2021
Cited alongside, same era.
Prompting for multimodal hateful meme classification
Rui Cao, Roy Ka-Wei Lee, Wen-Haw Chong, and Jing Jiang. 2022 · 2022
Later among the works it cites.
Adaptivity without compromise: a momentumized, adaptive, dual averaged gradient method for stochastic optimization
Aaron Defazio and Samy Jelassi. 2022 · 2022
Later among the works it cites.
MUTE: A multimodal dataset for detecting hateful memes
Eftekhar Hossain, Omar Sharif, and Mohammed Moshiul Hoque. 2022 · 2022
Later among the works it cites.
Multimodal hate speech detection from bengali memes and texts
Md Rezaul Karim, Sumon Kanti Dey, Tanhim Islam, Md Shajalal, and Bharathi Raja Chakravarthi. 2022 · 2022
Later among the works it cites.
Gokul Karthik Kumar and Karthik Nanadakumar. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Align before fuse: Vision and language representation learning with momentum distillation
Junnan Li, Ramprasaath Selvaraju, Akhilesh Gotmare, Shafiq Joty, Caiming Xiong, and Steven Chu Hong Hoi. 2021 · 2021
Cited alongside, same era.
Swin transformer: Hierarchical vision transformer using shifted windows
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo. 2021 · 2021
Cited alongside, same era.
Multimodal hate speech detection in greek social media
Konstantinos Perifanos and Dionysis Goutsos. 2021 · 2021
Cited alongside, same era.
MOMENTA: A multimodal framework for detecting harmful memes and their targets
Shraman Pramanick, Shivam Sharma, Dimitar Dimitrov, Md. Shad Akhtar, Preslav Nakov, and Tanmoy Chakraborty. 2021c · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al. 2021 · 2021
Cited alongside, same era.
Hyperextended lightface: A facial attribute analysis framework
Sefik Ilkin Serengil and Alper Ozpinar. 2021 · 2021
Cited alongside, same era.
Aomd: An analogy-aware approach to offensive meme detection on social media
Lanyu Shang, Yang Zhang, Yuheng Zha, Yingxi Chen, Christina Youn, and Dong Wang. 2021 · 2021
Cited alongside, same era.
Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Junnan Li, Dongxu Li, Caiming Xiong, and Steven Hoi. 2022 · 2022
Later among the works it cites.
Few-shot learning with multilingual generative language models
Xi Victoria Lin, Todor Mihaylov, Mikel Artetxe, Tianlu Wang, Shuohui Chen, Daniel Simig, Myle Ott, Naman Goyal, Shruti Bhosale, Jingfei Du, et al. 2022 · 2022
Later among the works it cites.
A convnet for the 2020s
Zhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer, Trevor Darrell, and Saining Xie. 2022 · 2022
Later among the works it cites.
Tackling cyber-aggression: Identification and fine-grained categorization of aggressive texts on social media using weighted ensemble of transformers
Omar Sharif and Mohammed Moshiul Hoque. 2022 · 2022
Later among the works it cites.
Banglaabusememe: A dataset for bengali abusive meme classification
Mithun Das and Animesh Mukherjee. 2023 · 2023
Later among the works it cites.
Emoffmeme: identifying offensive memes by leveraging underlying emotions
Gitanjali Kumari, Dibyanayan Bandyopadhyay, and Asif Ekbal. 2023 · 2023
Later among the works it cites.
Characterizing the entities in harmful memes: Who is the hero, the villain, the victim?
Shivam Sharma, Atharva Kulkarni, Tharun Suresh, Himanshi Mathur, Preslav Nakov, Md Shad Akhtar, and Tanmoy Chakraborty. 2023 · 2023
Later among the works it cites.