Fetching the paper…
Reading the bibliography…
Visual Instruction Tuning represents a novel learning paradigm involving the fine-tuning of pre-trained language models using task-specific instructions.
F. Alpher, “Frobnication,”
2002
Earlier work this paper cites.
F. Alpher and F. Fotheringham-Smythe, “Frobnication revisited,”
2003
Earlier work this paper cites.
F. Alpher, F. Fotheringham-Smythe, and F. Gamow, “Can a machine frobnicate?”
2004
Earlier work this paper cites.
F. Alpher and F. Gamow, “Can a computer frobnicate?” in
2005
Earlier work this paper cites.
J. A. Mikels, B. L. Fredrickson, G. R. Larkin, C. M. Lindberg, S. J. Maglio, and P. A. Reuter-Lorenz, “Emotional category data on images from the international affective picture system,”
2005
Earlier work this paper cites.
J. Machajdik and A. Hanbury, “Affective image classification using features inspired by psychology and art theory,” in
2010
Earlier work this paper cites.
F. LastName, “The frobnicatable foo filter,” 2014, face and Gesture submission ID 324. Supplied as supplemental material
2014
Earlier work this paper cites.
——, “Frobnication tutorial,” 2014, supplied as supplemental material
2014
Earlier work this paper cites.
K.-C. Peng, T. Chen, A. Sadovnik, and A. C. Gallagher, “A mixed bag of emotions: Model, predict, and transfer emotion distributions,” in
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Earlier work this paper cites.
Q. You, J. Luo, H. Jin, and J. Yang, “Building a large scale dataset for image emotion recognition: The fine print and the benchmark,” in
2016
Earlier work this paper cites.
K.-C. Peng, A. Sadovnik, A. Gallagher, and T. Chen, “Where do emotions come from? predicting the emotion stimuli map,” in
2016
Earlier work this paper cites.
R. Panda, J. Zhang, H. Li, J.-Y. Lee, X. Lu, and A. K. Roy-Chowdhury, “Contemplating visual emotions: Understanding and overcoming dataset bias,” in
2018
Earlier work this paper cites.
D. She, J. Yang, M.-M. Cheng, Y.-K. Lai, P. L. Rosin, and L. Wang, “Wscnet: Weakly supervised coupled networks for visual sentiment classification and detection,”
2019
Cited alongside, same era.
S. Zhao, Z. Jia, H. Chen, L. Li, G. Ding, and K. Keutzer, “Pdanet: Polarity-consistent deep attention network for fine-grained visual emotion regression,” in
2019
Cited alongside, same era.
W. Zhang, X. He, and W. Lu, “Exploring discriminative representations for image emotion recognition with cnns,”
2019
Cited alongside, same era.
H. Zhang and M. Xu, “Weakly supervised emotion intensity prediction for recognition of emotions in images,”
2020
Cited alongside, same era.
H.-X. Xie, L. Lo, H.-H. Shuai, and W.-H. Cheng, “Au-assisted graph attention convolutional network for micro-expression recognition,” in
2020
Cited alongside, same era.
W. Dai, J. Li, D. Li, A. M. H. Tiong, J. Zhao, W. Wang, B. Li, P. Fung, and S. Hoi, “Instructblip: Towards general-purpose vision-language models with instruction tuning,” in
2023
Later among the works it cites.
J. Yang, Q. Huang, T. Ding, D. Lischinski, D. Cohen-Or, and H. Huang, “Emoset: A large-scale visual emotion dataset with rich attributes,” in
2023
Later among the works it cites.
J. Li, D. Li, S. Savarese, and S. Hoi, “Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,” in
2023
Later among the works it cites.
2023
Later among the works it cites.
H. Xie, M.-X. Lee, T.-J. Chen, H.-J. Chen, H.-I. Liu, H.-H. Shuai, and W.-H. Cheng, “Most important person-guided dual-branch cross-patch attention for group affect recognition,” in
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Yang, J. Li, X. Wang, Y. Ding, and X. Gao, “Stimuli-aware visual emotion analysis,”
2021
Cited alongside, same era.
J.-B. Alayrac, J. Donahue, P. Luc, A. Miech, I. Barr, Y. Hasson, K. Lenc, A. Mensch, K. Millican, M. Reynolds
2022
Cited alongside, same era.
L. Xu, Z. Wang, B. Wu, and S. Lui, “Mdan: Multi-level dependent attention network for visual emotion analysis,” in
2022
Cited alongside, same era.
M. Jia, L. Tang, B.-C. Chen, C. Cardie, S. Belongie, B. Hariharan, and S.-N. Lim, “Visual prompt tuning,” in
2022
Cited alongside, same era.
——, “An overview of facial micro-expression analysis: Data, methodology and challenge,”
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2023
Later among the works it cites.
S. Zhang, L. Dong, X. Li, S. Zhang, X. Sun, S. Wang, J. Li, R. Hu, T. Zhang, F. Wu
2023
Later among the works it cites.
2023
Later among the works it cites.
M. Dehghani, J. Djolonga, B. Mustafa, P. Padlewski, J. Heek, J. Gilmer, A. P. Steiner, M. Caron, R. Geirhos, I. Alabdulmohsin
2023
Later among the works it cites.
OpenAI, “Gpt-4 technical report,” Tech. Rep., 2023
2023
Later among the works it cites.
E. Parliament, “Eu ai act: first regulation on artificial intelligence,”
2023
Later among the works it cites.
R. Li, S. Sun, M. Elhoseiny, and P. Torr, “Oxfordtvg-hic: Can machine make humorous captions from images?” in
2023
Later among the works it cites.
H. Xie, H. Chung, H.-H. Shuai, and W.-H. Cheng, “Learning to prompt for vision-language emotion recognition,” in
2023
Later among the works it cites.