Fetching the paper…
Reading the bibliography…
In AI-facilitated teaching, leveraging various query styles to interpret abstract text descriptions is crucial for ensuring high-quality teaching.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan. 2014 · 2014
Earlier work this paper cites.
Image retrieval using scene graphs
Justin Johnson, Ranjay Krishna, Michael Stark, Li-Jia Li, David Shamma, Michael Bernstein, and Li Fei-Fei. 2015 · 2015
Earlier work this paper cites.
Trace norm regularised deep multi-task learning
Yongxin Yang and Timothy M Hospedales. 2016 · 2016
Earlier work this paper cites.
Joint embeddings of scene graphs and images
Eugene Belilovsky, Matthew Blaschko, Jamie Ryan Kiros, Raquel Urtasun, and Richard Zemel. 2017 · 2017
Earlier work this paper cites.
Ubernet: Training a universal convolutional neural network for low-, mid-, and high-level vision using diverse datasets and limited memory
Iasonas Kokkinos. 2017 · 2017
Earlier work this paper cites.
Generating holistic 3d scene abstractions for text-based image retrieval
Ang Li, Jin Sun, Joe Yue-Hei Ng, Ruichi Yu, Vlad I Morariu, and Larry S Davis. 2017 · 2017
Earlier work this paper cites.
Spatial-semantic image search by visual feature synthesis
Long Mai, Hailin Jin, Zhe Lin, Chen Fang, Jonathan Brandt, and Feng Liu. 2017 · 2017
Earlier work this paper cites.
Global dataset on education quality: A review and update (2000-2017)
Harry A Patrinos and Noam Angrist. 2018 · 2017
Earlier work this paper cites.
An overview of multi-task learning in deep neural networks
S Ruder. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Crowdsourcing multiple choice science questions
Johannes Welbl, Nelson F Liu, and Matt Gardner. 2017 · 2017
Earlier work this paper cites.
Moment matching for multi-source domain adaptation
Xingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang, Kate Saenko, and Bo Wang. 2019 · 2019
Earlier work this paper cites.
Language models are few-shot learners
T Brown, B Mann, N Ryder, M Subbiah, JD Kaplan, P Dhariwal, A Neelakantan, P Shyam, G Sastry, A Askell, et al. 2020 · 2020
Earlier work this paper cites.
Ednet: A large-scale hierarchical dataset in education
Youngduck Choi, Youngnam Lee, Dongmin Shin, Junghyun Cho, Seoyon Park, et al. 2020 · 2020
Earlier work this paper cites.
A brief review of domain adaptation
Abolfazl Farahani, Sahar Voghoei, Khaled Rasheed, and Hamid R Arabnia. 2021 · 2020
Earlier work this paper cites.
Vision, challenges, roles and research issues of artificial intelligence in education
Gwo-Jen Hwang, Haoran Xie, Benjamin W Wah, and Dragan Gašević. 2020 · 2020
Earlier work this paper cites.
Cross-modal scene graph matching for relationship-aware image-text retrieval
Sijin Wang, Ruiping Wang, Ziwei Yao, Shiguang Shan, and Xilin Chen. 2020 · 2020
Earlier work this paper cites.
GPT-Neo: Large Scale Autoregressive Language Modeling with Mesh-Tensorflow
Sid Black, Gao Leo, Phil Wang, Connor Leahy, and Stella Biderman. 2021 · 2021
Earlier work this paper cites.
Measuring mathematical problem solving with the math dataset
Dan Hendrycks, Collin Burns, Saurav Kadavath, Akul Arora, Steven Basart, et al. 2021 · 2021
Earlier work this paper cites.
Structured visual search via composition-aware learning
Mert Kilickaya and Arnold WM Smeulders. 2021 · 2021
Earlier work this paper cites.
The power of scale for parameter-efficient prompt tuning
Brian Lester, Rami Al-Rfou, and Noah Constant. 2021b · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, et al. 2021 · 2021
Earlier work this paper cites.
Deep learning for instance retrieval: A survey
Wei Chen, Yu Liu, Weiping Wang, Erwin M Bakker, Theodoros Georgiou, Paul Fieguth, Li Liu, and Michael S Lew. 2022 · 2022
Cited alongside, same era.
Fs-coco: Towards understanding of freehand sketches of common objects in context
Pinaki Nath Chowdhury, Aneeshan Sain, Ayan Kumar Bhunia, Tao Xiang, Yulia Gryaditskaya, and Yi-Zhe Song. 2022 · 2022
Cited alongside, same era.
Visual prompt tuning
Menglin Jia, Luming Tang, Bor-Chun Chen, Claire Cardie, Serge Belongie, et al. 2022 · 2022
Cited alongside, same era.
Probabilistic compositional embeddings for multimodal image retrieval
Andrei Neculai, Yanbei Chen, and Zeynep Akata. 2022 · 2022
Cited alongside, same era.
Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering
Ankit Pal, Logesh Kumar Umapathi, and Malaikannan Sankarasubbu. 2022 · 2022
Cited alongside, same era.
Image style transfer based on vgg neural network model
Composed image retrieval via cross relation network with hierarchical aggregation transformer
Qu Yang, Mang Ye, Zhaohui Cai, Kehua Su, and Bo Du. 2023 · 2023
Later among the works it cites.
Prompt as triggers for backdoor attack: Examining the vulnerability in language models
Shuai Zhao, Jinming Wen, Anh Luu, Junbo Zhao, and Jie Fu. 2023 · 2023
Later among the works it cites.
Bin Zhu, Bin Lin, Munan Ning, Yang Yan, Jiaxi Cui, et al. 2023 · 2023
Later among the works it cites.
Enhancing document image retrieval in education: Leveraging ensemble-based document image retrieval systems for improved precision
Yehia Ibrahim Alzoubi, Ahmet Ercan Topcu, and Erdem Ozdemir. 2024 · 2024
Later among the works it cites.
Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Zhe Chen, Jiannan Wu, Wenhai Wang, Weijie Su, Guo Chen, Sen Xing, Muyan Zhong, et al. 2024 · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yilin Tao. 2022 · 2022
Cited alongside, same era.
Learning to prompt for continual learning
Zifeng Wang, Zizhao Zhang, Chen-Yu Lee, Han Zhang, Ruoxi Sun, Xiaoqi Ren, Guolong Su, Vincent Perot, Jennifer Dy, and Tomas Pfister. 2022 · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al. 2022 · 2022
Cited alongside, same era.
A framework for enabling unpaired multi-modal learning for deep cross-modal hashing retrieval
Mikel Williams-Lekuona, Georgina Cosma, and Iain Phillips. 2022 · 2022
Cited alongside, same era.
A survey on multi-task learning
Yu Zhang and Qiang Yang. 2022 · 2022
Cited alongside, same era.
Distribution-aware prompt tuning for vision-language models
Eulrang Cho, Jooyeon Kim, and Hyunwoo J Kim. 2023 · 2023
Cited alongside, same era.
Domain adaptation via prompt learning
Chunjiang Ge, Rui Huang, Mixue Xie, Zihang Lai, Shiji Song, Shuang Li, and Gao Huang. 2023 · 2023
Cited alongside, same era.
Later among the works it cites.
A survey on in-context learning
Qingxiu Dong, Lei Li, Damai Dai, Ce Zheng, Jingyuan Ma, Rui Li, Heming Xia, Jingjing Xu, Zhiyong Wu, Baobao Chang, et al. 2024 · 2024
Later among the works it cites.
Enhancing medical image retrieval with umls-integrated cnn-based text indexing
Karim Gasmi, Hajer Ayadi, and Mouna Torjmen. 2024 · 2024
Later among the works it cites.
E-eval: A comprehensive chinese k-12 education evaluation benchmark for large language models
Jinchang Hou, Chang Ao, Haihong Wu, Xiangtao Kong, et al. 2024 · 2024
Later among the works it cites.
Aaron Hurst, Adam Lerer, Adam P Goucher, Adam Perelman, Aditya Ramesh, Aidan Clark, et al. 2024 · 2024
Later among the works it cites.
Biped: Pedagogically informed tutoring system for esl education
Soonwoo Kwon, Sojung Kim, Minju Park, Seunghyun Lee, and Kyuseok Kim. 2024 · 2024
Later among the works it cites.
Alleviating the inconsistency of multimodal data in cross-modal retrieval
Tieying Li, Xiaochun Yang, Yiping Ke, Bin Wang, Yinan Liu, and Jiaxing Xu. 2024a · 2024
Later among the works it cites.
Prompt optimization via adversarial in-context learning
Do Long, Yiran Zhao, Hannah Brown, Yuxi Xie, James Zhao, Nancy Chen, Kenji Kawaguchi, Michael Shieh, and Junxian He. 2024 · 2024
Later among the works it cites.
Unified-io 2: Scaling autoregressive multimodal models with vision language audio and action
Jiasen Lu, Christopher Clark, Sangho Lee, Zichen Zhang, Savya Khosla, Ryan Marten, Derek Hoiem, and Aniruddha Kembhavi. 2024 · 2024
Later among the works it cites.
Multitask vision-language prompt tuning
Sheng Shen, Shijia Yang, Tianjun Zhang, Bohan Zhai, Joseph E Gonzalez, Kurt Keutzer, and Trevor Darrell. 2024 · 2024
Later among the works it cites.
Learning to (learn at test time): Rnns with expressive hidden states
Yu Sun, Xinhao Li, Karan Dalal, Jiarui Xu, Arjun Vikram, Genghan Zhang, Yann Dubois, Xinlei Chen, Xiaolong Wang, et al. 2024 · 2024
Later among the works it cites.
Unsupervised multimodal graph contrastive semantic anchor space dynamic knowledge distillation network for cross-media hash retrieval
Yang Yu, Meiyu Liang, Mengran Yin, Kangkang Lu, Junping Du, and Zhe Xue. 2024 · 2024
Later among the works it cites.
Universal vulnerabilities in large language models: Backdoor attacks for in-context learning
Shuai Zhao, Meihuizi Jia, Luu Anh Tuan, Fengjun Pan, and Jinming Wen. 2024a · 2024
Later among the works it cites.
Knowledge graph enhanced multimodal transformer for image-text retrieval
Juncheng Zheng, Meiyu Liang, Yang Yu, Yawen Li, and Zhe Xue. 2024 · 2024
Later among the works it cites.
Freestyleret: Retrieving images from style-diversified queries
Hao Li, Yanhao Jia, Peng Jin, Zesen Cheng, Kehan Li, Jialu Sui, Chang Liu, and Li Yuan. 2025 · 2025
Closest in time.
Exploring cognitive and aesthetic causality for multimodal aspect-based sentiment analysis
Luwei Xiao, Rui Mao, Shuai Zhao, Qika Lin, Yanhao Jia, Liang He, and Erik Cambria. 2025 · 2025
Closest in time.
Text-based image retrieval using progressive multi-instance learning
Wen Li, Lixin Duan, Dong Xu, and Ivor Wai-Hung Tsang. 2011 · 2055
Closest in time.