Fetching the paper…
Reading the bibliography…
Image search stands as a pivotal task in multimedia and computer vision, finding applications across diverse domains, ranging from internet search to medical diagnostics.
Relevance feedback in information retrieval
J. J. Rocchio. 1971 · 1971
Earlier work this paper cites.
Local Feedback in Full-Text Retrieval Systems
R. Attar and Aviezri S. Fraenkel. 1977 · 1977
Earlier work this paper cites.
On Term Selection for Query Expansion
Stephen E. Robertson. 1991 · 1991
Earlier work this paper cites.
Incremental relevance feedback for information filtering. In Annual International ACM SIGIR Conference on Research and Development in Information Retrieval
James Allan. 1996 · 1996
Earlier work this paper cites.
Training algorithms for linear text classifiers. In Annual International ACM SIGIR Conference on Research and Development in Information Retrieval
David D. Lewis, Robert E. Schapire, Jamie Callan, and Ron Papka. 1996 · 1996
Earlier work this paper cites.
Query expansion using local and global document analysis. In Annual International ACM SIGIR Conference on Research and Development in Information Retrieval
Jinxi Xu and W. Bruce Croft. 1996 · 1996
Earlier work this paper cites.
A Probabilistic Analysis of the Rocchio Algorithm with TFIDF for Text Categorization. In International Conference on Machine Learning
Thorsten Joachims. 1997 · 1997
Earlier work this paper cites.
Content-based image retrieval with relevance feedback in MARS
Yong Rui, Thomas S. Huang, and Sharad Mehrotra. 1997 · 1997
Earlier work this paper cites.
A novel relevance feedback technique in image retrieval. In MULTIMEDIA ’99
Yong Rui and Thomas S. Huang. 1999 · 1999
Earlier work this paper cites.
Relevance feedback with a small number of relevance judgements: incremental relevance feedback vs. document clustering. In Annual International ACM SIGIR Conference on Research and Development in Information Retrieval
Makoto Iwayama. 2000 · 2000
Earlier work this paper cites.
Content-based image retrieval at the end of the early years
A.W.M. Smeulders, M. Worring, S. Santini, A. Gupta, and R. Jain. 2000 · 2000
Earlier work this paper cites.
Model-based feedback in the language modeling approach to information retrieval. In International Conference on Information and Knowledge Management
ChengXiang Zhai and John D. Lafferty. 2001 · 2001
Earlier work this paper cites.
Probabilistic models of information retrieval based on measuring the divergence from randomness
Gianni Amati and C J Van Rijsbergen. 2002 · 2002
Earlier work this paper cites.
The IIR evaluation model: a framework for evaluation of interactive information retrieval systems
Pia Borlund. 2003 · 2003
Earlier work this paper cites.
Context-sensitive information retrieval using implicit feedback. In Annual International ACM SIGIR Conference on Research and Development in Information Retrieval
Xuehua Shen, Bin Tan, and Chengxiang Zhai. 2005 · 2005
Earlier work this paper cites.
Extending Faceted Navigation for RDF Data. In International Workshop on the Semantic Web
Eyal Oren, Renaud Delbru, and Stefan Decker. 2006 · 2006
Earlier work this paper cites.
IM2GPS: estimating geographic information from a single image
James Hays and Alexei A. Efros. 2008 · 2008
Earlier work this paper cites.
Active Learning for Interactive Multimedia Retrieval
Thomas S. Huang, Charlie K. Dagli, Shyamsundar Rajaram, Edward Y. Chang, Michael I. Mandel, Graham E. Poliner, and Daniel P. W. Ellis. 2008 · 2008
Earlier work this paper cites.
The retrieval effectiveness of web search engines: considering results descriptions
Dirk Lewandowski. 2008 · 2008
Earlier work this paper cites.
A study of methods for negative relevance feedback. In Annual International ACM SIGIR Conference on Research and Development in Information Retrieval
Xuanhui Wang, Hui Fang, and ChengXiang Zhai. 2008 · 2008
Earlier work this paper cites.
Automatic tagging and geotagging in video collections and communities. In Proceedings of the 1st ACM International Conference on Multimedia Retrieval (ICMR ’11) . Association for Computing Machinery, New York, NY, USA, Article 51, 8 pages
Martha Larson, Mohammad Soleymani, Pavel Serdyukov, Stevan Rudinac, Christian Wartena, Vanessa Murdock, Gerald Friedland, Roeland Ordelman, and Gareth J. F. Jones. 2011 · 2011
Earlier work this paper cites.
Relative attributes
Devi Parikh and Kristen Grauman. 2011 · 2011
Earlier work this paper cites.
Introduction to Recommender Systems Handbook. In Recommender Systems Handbook
Francesco Ricci, Lior Rokach, and Bracha Shapira. 2011 · 2011
Earlier work this paper cites.
WhittleSearch: Image search with relative attribute feedback
Adriana Kovashka, Devi Parikh, and Kristen Grauman. 2012 · 2012
Earlier work this paper cites.
Leveraging visual concepts and query performance prediction for semantic-theme-based video retrieval
Stevan Rudinac, Martha Larson, and Alan Hanjalic. 2012 · 2012
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
Accurate and effective latent concept modeling for ad hoc information retrieval
Romain Deveaud, Eric SanJuan, and Patrice Bellot. 2014 · 2014
Earlier work this paper cites.
Learning a Deep Convolutional Network for Image Super-Resolution. In European Conference on Computer Vision
Chao Dong, Chen Change Loy, Kaiming He, and Xiaoou Tang. 2014 · 2014
Earlier work this paper cites.
Microsoft COCO: Common Objects in Context. In European Conference on Computer Vision
Tsung-Yi Lin, Michael Maire, Serge J. Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C. Lawrence Zitnick. 2014 · 2014
Earlier work this paper cites.
Bridging the Ultimate Semantic Gap: A Semantic Search Engine for Internet Videos
Lu Jiang, Shoou-I Yu, Deyu Meng, Teruko Mitamura, and Alexander Hauptmann. 2015 · 2015
Earlier work this paper cites.
Learning deep representations for ground-to-aerial geolocalization
Tsung-Yi Lin, Yin Cui, Serge J. Belongie, and James Hays. 2015 · 2015
Earlier work this paper cites.
Deep Face Recognition. In British Machine Vision Conference
Omkar M. Parkhi, Andrea Vedaldi, and Andrew Zisserman. 2015 · 2015
Cited alongside, same era.
Flickr30k Entities: Collecting Region-to-Phrase Correspondences for Richer Image-to-Sentence Models
Bryan A. Plummer, Liwei Wang, Christopher M. Cervantes, Juan C. Caicedo, J. Hockenmaier, and Svetlana Lazebnik. 2015 · 2015
Cited alongside, same era.
FaceNet: A unified embedding for face recognition and clustering
Florian Schroff, Dmitry Kalenichenko, and James Philbin. 2015 · 2015
Cited alongside, same era.
Unsupervised learning of visual representations using videos. In Proceedings of the IEEE international conference on computer vision . 2794–2802
Xiaolong Wang and Abhinav Gupta. 2015 · 2015
Cited alongside, same era.
Analytic Quality: Evaluation of Performance and Insight in Multimedia Collection Analysis
Jan Zahálka, Stevan Rudinac, and Marcel Worring. 2015 · 2015
Cited alongside, same era.
Scaling up visual and vision-language representation learning with noisy text supervision. In International conference on machine learning . PMLR, 4904–4916
Chao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen, Zarana Parekh, Hieu Pham, Quoc Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig. 2021 · 2021
Later among the works it cites.
Supervision Exists Everywhere: A Data Efficient Contrastive Language-Image Pre-training Paradigm
Yangguang Li, Feng Liang, Lichen Zhao, Yufeng Cui, Wanli Ouyang, Jing Shao, Fengwei Yu, and Junjie Yan. 2021a · 2021
Later among the works it cites.
Xiaopeng Lu, Tiancheng Zhao, and Kyusong Lee. 2021 · 2021
Later among the works it cites.
Learning Transferable Visual Models From Natural Language Supervision. In International Conference on Machine Learning
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
MSR-VTT: A Large Video Description Dataset for Bridging Video and Language
Jun Xu, Tao Mei, Ting Yao, and Yong Rui. 2016 · 2016
Cited alongside, same era.
Summary transfer: Exemplar-based subset selection for video summarization. In Proceedings of the IEEE conference on computer vision and pattern recognition . 1059–1067
Ke Zhang, Wei-Lun Chao, Fei Sha, and Kristen Grauman. 2016 · 2016
Cited alongside, same era.
Recent Advance in Content-based Image Retrieval: A Literature Survey
Wen gang Zhou, Houqiang Li, and Qi Tian. 2017 · 2017
Cited alongside, same era.
Automatic Spatially-Aware Fashion Concept Discovery
Xintong Han, Zuxuan Wu, Phoenix X. Huang, Xiao Zhang, Menglong Zhu, Yuan Li, Yang Zhao, and Larry S. Davis. 2017 · 2017
Cited alongside, same era.
Robustness Analysis of Visual Question Answering Models by Basic Questions
Jia-Hong Huang. 2017 · 2017
Cited alongside, same era.
VQABQ: Visual Question Answering by Basic Questions
Jia-Hong Huang, Modar Alfadly, and Bernard Ghanem. 2017 · 2017
Cited alongside, same era.
Large-scale image retrieval with attentive deep local features. In Proceedings of the IEEE international conference on computer vision . 3456–3465
Hyeonwoo Noh, Andre Araujo, Jack Sim, Tobias Weyand, and Bohyung Han. 2017 · 2017
Cited alongside, same era.
LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs
Christoph Schuhmann, Richard Vencu, Romain Beaumont, Robert Kaczmarczyk, Clayton Mullis, Aarush Katta, Theo Coombes, Jenia Jitsev, and Aran Komatsuzaki. 2021 · 2021
Later among the works it cites.
Lightningdot: Pre-training visual-semantic embeddings for real-time image-text retrieval. In Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . 982–997
Siqi Sun, Yen-Chun Chen, Linjie Li, Shuohang Wang, Yuwei Fang, and Jingjing Liu. 2021 · 2021
Later among the works it cites.
Controlling the Risk of Conversational Search via Reinforcement Learning
Zhenduo Wang and Qingyao Ai. 2021 · 2021
Later among the works it cites.
Scaling Instruction-Finetuned Language Models
Hyung Won Chung, Le Hou, S. Longpre, Barret Zoph, Yi Tay, and William Fedus et al. 2022 · 2022
Later among the works it cites.
X-mir: Explainable medical image retrieval. In Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision . 440–450
Brian Hu, Bhavan Vasu, and Anthony Hoogs. 2022 · 2022
Later among the works it cites.
Large Language Models are Zero-Shot Reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa. 2022 · 2022
Later among the works it cites.
COTS: Collaborative two-stream vision-language pre-training model for cross-modal retrieval. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 15692–15701
Haoyu Lu, Nanyi Fei, Yuqi Huo, Yizhao Gao, Zhiwu Lu, and Ji-Rong Wen. 2022 · 2022
Later among the works it cites.
Revisiting Open Domain Query Facet Extraction and Generation
Chris Samarinas and Hamed Zamani. 2022 · 2022
Later among the works it cites.
Chain of Thought Prompting Elicits Reasoning in Large Language Models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Ed Huai hsin Chi, F. Xia, Quoc Le, and Denny Zhou. 2022 · 2022
Later among the works it cites.
Contrastive Learning of Medical Visual Representations from Paired Images and Text. In Proceedings of the 7th Machine Learning for Healthcare Conference (Proceedings of Machine Learning Research) , Zachary Lipton, Rajesh Ranganath, Mark Sendak, Michael Sjoding, and Serena Yeung (Eds.), Vol. 182. PMLR, 2–25
Yuhao Zhang, Hang Jiang, Yasuhide Miura, Christopher D. Manning, and Curtis P. Langlotz. 2022 · 2022
Later among the works it cites.
InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning
Wenliang Dai, Junnan Li, Dongxu Li, Anthony Meng Huat Tiong, Junqi Zhao, Weisheng Wang, Boyang Albert Li, Pascale Fung, and Steven C. H. Hoi. 2023 · 2023
Later among the works it cites.
CIEM: Contrastive Instruction Evaluation Method for Better Instruction Tuning
Hongyu Hu, Jiyuan Zhang, Minyi Zhao, and Zhenbang Sun. 2023 · 2023
Later among the works it cites.
Jia-Hong Huang, Modar Alfadly, Bernard Ghanem, and Marcel Worring. 2023 · 2023
Later among the works it cites.
Query Expansion by Prompting Large Language Models
Rolf Jagerman, Honglei Zhuang, Zhen Qin, Xuanhui Wang, and Michael Bendersky. 2023a · 2023
Later among the works it cites.
Query Expansion by Prompting Large Language Models
Rolf Jagerman, Honglei Zhuang, Zhen Qin, Xuanhui Wang, and Michael Bendersky. 2023b · 2023
Later among the works it cites.
Evaluating Object Hallucination in Large Vision-Language Models
Yifan Li, Yifan Du, Kun Zhou, Jinpeng Wang, Wayne Xin Zhao, and Ji rong Wen. 2023a · 2023
Later among the works it cites.
Generative Relevance Feedback with Large Language Models
Iain Mackie, Shubham Chatterjee, and Jeffrey Dalton. 2023 · 2023
Later among the works it cites.
Kelong Mao, Zhicheng Dou, Haonan Chen, Fengran Mo, and Hongjin Qian. 2023a · 2023
Later among the works it cites.
Search-oriented conversational query editing. In Findings of the Association for Computational Linguistics: ACL 2023 . 4160–4172
Kelong Mao, Zhicheng Dou, Bang Liu, Hongjin Qian, Fengran Mo, Xiangli Wu, Xiaohua Cheng, and Zhao Cao. 2023b · 2023
Later among the works it cites.
Text-to-Image Fashion Retrieval with Fabric Textures
Daichi Suzuki, Go Irie, and Kiyoharu Aizawa. 2023 · 2023
Later among the works it cites.
Llama 2: Open Foundation and Fine-Tuned Chat Models
Hugo Touvron, Louis Martin, Kevin R. Stone, Peter Albert, and Amjad Almahairi et al. 2023 · 2023
Later among the works it cites.
Element-aware Summarization with Large Language Models: Expert-aligned Evaluation and Chain-of-Thought Method. In Annual Meeting of the Association for Computational Linguistics
Yiming Wang, Zhuosheng Zhang, and Rui Wang. 2023 · 2023
Later among the works it cites.
Enhancing conversational search: Large language model-aided informative query rewriting
Fanghua Ye, Meng Fang, Shenghui Li, and Emine Yilmaz. 2023 · 2023
Later among the works it cites.
Bohan Zhai, Shijia Yang, Xiangchen Zhao, Chenfeng Xu, Sheng Shen, Dongdi Zhao, Kurt Keutzer, Manling Li, Tan Yan, and Xiangjun Fan. 2023 · 2023
Later among the works it cites.
Judging LLM-as-a-judge with MT-Bench and Chatbot Arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric P. Xing, Haotong Zhang, Joseph Gonzalez, and Ion Stoica. 2023 · 2023
Later among the works it cites.
Chatting makes perfect: Chat-based image retrieval
Matan Levy, Rami Ben-Ari, Nir Darshan, and Dani Lischinski. 2024 · 2024
Closest in time.
Prototype-Enhanced Hypergraph Learning for Heterogeneous Information Networks. In International Conference on Multimedia Modeling . Springer, 462–476
Shuai Wang, Jiayi Shen, Athanasios Efthymiou, Stevan Rudinac, Monika Kackovic, Nachoem Wijnberg, and Marcel Worring. 2024 · 2024
Closest in time.