Fetching the paper…
Reading the bibliography…
Recent multimodal large language models (MLLMs) have demonstrated significant potential in open-ended conversation, generating more accurate and personalized responses.
The relative importance of work: Models and measures for meaningful data
Donald E Super. 1982 · 1982
Earlier work this paper cites.
The dialogic mind: A dialogic approach to the higher mental functions
Charles Fernyhough. 1996 · 1996
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Dosovitskiy Alexey. 2020 · 2010
Earlier work this paper cites.
Open-ended elaborations in creative metaphor
John Barnden. 2014 · 2014
Earlier work this paper cites.
Memorybank: Enhancing large language models with long-term memory
Wanjun Zhong, Lianghong Guo, Qiqi Gao, He Ye, and Yanlin Wang. 2024 · 2014
Earlier work this paper cites.
The importance of open data and software: Is energy research lagging behind?
Stefan Pfenninger, Joseph DeCarolis, Lion Hirth, Sylvain Quoilin, and Iain Staffell. 2017 · 2017
Earlier work this paper cites.
Dialog-context aware end-to-end speech recognition
Suyoun Kim and Florian Metze. 2018 · 2018
Earlier work this paper cites.
What makes a good conversation? challenges in designing truly conversational agents
Leigh Clark, Nadia Pantidi, Orla Cooney, Philip Doyle, Diego Garaialde, Justin Edwards, Brendan Spillane, Emer Gilmartin, Christine Murad, Cosmin Munteanu, et al. 2019 · 2019
Earlier work this paper cites.
Big data privacy and ethical challenges
Paulette Lacroix. 2019 · 2019
Earlier work this paper cites.
A review of the research on dialogue management of task-oriented systems
Yin Jiang Zhao, Yan Ling Li, and Min Lin. 2019 · 2019
Earlier work this paper cites.
Tell me about yourself: Using an ai-powered chatbot to conduct conversational surveys with open-ended questions
Ziang Xiao, Michelle X Zhou, Q Vera Liao, Gloria Mark, Changyan Chi, Wenxi Chen, and Huahai Yang. 2020 · 2020
Earlier work this paper cites.
A teacher questioning activity: The use of oral open-ended questions in mathematics classroom
Mela Aziza. 2021 · 2021
Earlier work this paper cites.
An introduction to conversation analysis
Anthony J Liddicoat. 2021 · 2021
Earlier work this paper cites.
Ethical issues in consent for the reuse of data in health data platforms
Alex McKeown, Miranda Mourby, Paul Harrison, Sophie Walker, Mark Sheehan, and Ilina Singh. 2021 · 2021
Earlier work this paper cites.
Data and its (dis) contents: A survey of dataset development and use in machine learning research
Amandalynne Paullada, Inioluwa Deborah Raji, Emily M Bender, Emily Denton, and Alex Hanna. 2021 · 2021
Earlier work this paper cites.
Scholarly text classification with sentence bert and entity embeddings
Guangyuan Piao. 2021 · 2021
Earlier work this paper cites.
Chatbots language design: The influence of language variation on user experience with tourist assistant chatbots
Ana Paula Chaves, Jesse Egbert, Toby Hocking, Eck Doerry, and Marco Aurelio Gerosa. 2022 · 2022
Earlier work this paper cites.
Topic Shifts: Preserving Comprehension in Conversation
Amandine Decker. 2022 · 2022
Earlier work this paper cites.
What do we hear in the voice? an open-ended judgment study of emotional speech prosody
Hillary Anger Elfenbein, Petri Laukka, Jean Althoff, Wanda Chui, Frederick K Iraki, Thomas Rockstuhl, and Nutankumar S Thingujam. 2022 · 2022
Earlier work this paper cites.
Implicit data crimes: Machine learning bias arising from misuse of public data
Efrat Shimron, Jonathan I Tamir, Ke Wang, and Michael Lustig. 2022 · 2022
Earlier work this paper cites.
The relationship between collaborative problem solving behaviors and solution outcomes in a game-based learning environment
Chen Sun, Valerie J Shute, Angela EB Stewart, Quinton Beck-White, Caroline R Reinhardt, Guojing Zhou, Nicholas Duran, and Sidney K D’Mello. 2022 · 2022
Earlier work this paper cites.
Stepwise feature fusion: Local guides global
Jinfeng Wang, Qiming Huang, Feilong Tang, Jia Meng, Jionglong Su, and Sifan Song. 2022 · 2022
Earlier work this paper cites.
Clip pre-trained models for cross-modal retrieval in newsimages 2022
Yang Zhang, Yi Shao, Wenbo Wan, Jing Li, and Jiande Sun. 2022 · 2022
Earlier work this paper cites.
Ethical considerations for responsible data curation
Jerone Andrews, Dora Zhao, William Thong, Apostolos Modas, Orestis Papakyriakopoulos, and Alice Xiang. 2023 · 2023
Earlier work this paper cites.
Qwen-vl: A frontier large vision-language model with versatile abilities
Jinze Bai, Shuai Bai, Shusheng Yang, Shijie Wang, Sinan Tan, Peng Wang, Junyang Lin, Chang Zhou, and Jingren Zhou. 2023 · 2023
Earlier work this paper cites.
Fine-tuning multimodal llms to follow zero-shot demonstrative instructions
Juncheng Li, Kaihang Pan, Zhiqi Ge, Minghe Gao, Wei Ji, Wenqiao Zhang, Tat-Seng Chua, Siliang Tang, Hanwang Zhang, and Yueting Zhuang. 2023 · 2023
Earlier work this paper cites.
Improved baselines with visual instruction tuning
Haotian Liu, Chunyuan Li, Yuheng Li, and Yong Jae Lee. 2023 · 2023
Earlier work this paper cites.
Dialogbench: Evaluating llms as human-like dialogue systems
Jiao Ou, Junda Lu, Che Liu, Yihong Tang, Fuzheng Zhang, Di Zhang, and Kun Gai. 2023 · 2023
Cited alongside, same era.
Duat: Dual-aggregation transformer network for medical image segmentation
Feilong Tang, Zhongxing Xu, Qiming Huang, Jinfeng Wang, Xianxu Hou, Jionglong Su, and Jingxin Liu. 2023 · 2023
Cited alongside, same era.
Rongwu Xu, Brian S Lin, Shujian Yang, Tianqi Zhang, Weiyan Shi, Tianwei Zhang, Zhixuan Fang, Wei Xu, and Han Qiu. 2023 · 2023
Cited alongside, same era.
Sigmoid loss for language image pre-training
Xiaohua Zhai, Basil Mustafa, Alexander Kolesnikov, and Lucas Beyer. 2023 · 2023
Cited alongside, same era.
Qwen2 technical report
2024 · 2024
Cited alongside, same era.
Mitigating dialogue hallucination for large multi-modal models via adversarial instruction tuning
Dongmin Park, Zhaofang Qian, Guangxing Han, and Ser-Nam Lim. 2024 · 2024
Later among the works it cites.
What is research data “misuse”? and how can it be prevented or mitigated?
Irene V Pasquetto, Zoë Cullen, Andrea Thomer, and Morgan Wofford. 2024 · 2024
Later among the works it cites.
Data privacy and ethical considerations in database management
Eduardo Pina, José Ramos, Henrique Jorge, Paulo Váz, José Silva, Cristina Wanzeller, Maryam Abbasi, and Pedro Martins. 2024 · 2024
Later among the works it cites.
Chacha: Leveraging large language models to prompt children to share their emotions about personal events
Woosuk Seo, Chanmo Yang, and Young-Ho Kim. 2024 · 2024
Later among the works it cites.
Parrot: Enhancing multi-turn instruction following for large language models
Yuchong Sun, Che Liu, Kun Zhou, Jinwen Huang, Ruihua Song, Wayne Xin Zhao, Fuzheng Zhang, Di Zhang, and Kun Gai. 2024 · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A review of dialogue systems: current trends and future directions
Atheer Algherairy and Moataz Ahmed. 2024 · 2024
Cited alongside, same era.
Mt-bench-101: A fine-grained benchmark for evaluating large language models in multi-turn dialogues
Ge Bai, Jie Liu, Xingyuan Bu, Yancheng He, Jiaheng Liu, Zhanhui Zhou, Zhuoran Lin, Wenbo Su, Tiezheng Ge, Bo Zheng, et al. 2024 · 2024
Cited alongside, same era.
Meta curvature-aware minimization for domain generalization
Ziyang Chen, Yiwen Ye, Feilong Tang, Yongsheng Pan, and Yong Xia. 2024 · 2024
Cited alongside, same era.
A survey on multimodal large language models for autonomous driving
Can Cui, Yunsheng Ma, Xu Cao, Wenqian Ye, Yang Zhou, Kaizhao Liang, Jintai Chen, Juanwu Lu, Zichong Yang, Kuei-Da Liao, et al. 2024 · 2024
Cited alongside, same era.
Abhimanyu Dubey, Abhinav Jauhri, Abhinav Pandey, Abhishek Kadian, Ahmad Al-Dahle, Aiesha Letman, Akhil Mathur, Alan Schelten, Amy Yang, Angela Fan, et al. 2024 · 2024
Cited alongside, same era.
From multimodal llm to human-level ai: Modality, instruction, reasoning and beyond
Hao Fei, Xiangtai Li, Haotian Liu, Fuxiao Liu, Zhuosheng Zhang, Hanwang Zhang, and Shuicheng Yan. 2024 · 2024
Cited alongside, same era.
Mme-survey: A comprehensive survey on evaluation of multimodal llms
Chaoyou Fu, Yi-Fan Zhang, Shukang Yin, Bo Li, Xinyu Fang, Sirui Zhao, Haodong Duan, Xing Sun, Ziwei Liu, Liang Wang, et al. 2024 · 2024
Cited alongside, same era.
Later among the works it cites.
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Gemini Team, Petko Georgiev, Ving Ian Lei, Ryan Burnell, Libin Bai, Anmol Gulati, Garrett Tanzer, Damien Vincent, Zhufeng Pan, Shibo Wang, et al. 2024 · 2024
Later among the works it cites.
Sight for sore heads–using cnns to diagnose migraines
Matt Trinh, Feilong Tang, Angelica Ly, Annita Duong, Fiona Stapleton, Zongyuan Ge, and Imran Razzak. 2024 · 2024
Later among the works it cites.
Demonstrative instruction following in multimodal llms via integrating low-rank adaptation with ensemble learning
Jingyu Wei, Yi Su, Kele Xu, Lingbin Zeng, Bo Liu, and Huaimin Wang. 2024 · 2024
Later among the works it cites.
Longmemeval: Benchmarking chat assistants on long-term interactive memory
Di Wu, Hongwei Wang, Wenhao Yu, Yuwei Zhang, Kai-Wei Chang, and Dong Yu. 2024 · 2024
Later among the works it cites.
Sam2-unet: Segment anything 2 makes strong encoder for natural and medical image segmentation
Xinyu Xiong, Zihuang Wu, Shuangyi Tan, Wenxue Li, Feilong Tang, Ying Chen, Siying Li, Jie Ma, and Guanbin Li. 2024 · 2024
Later among the works it cites.
Mlevlm: Improve multi-level progressive capabilities based on multimodal large language model for medical visual question answering
Dexuan Xu, Yanyuan Chen, Jieyi Wang, Yue Huang, Hanpin Wang, Zhi Jin, Hongxing Wang, Weihua Yue, Jing He, Hang Li, et al. 2024a · 2024
Later among the works it cites.
Minicpm-v: A gpt-4v level mllm on your phone
Yuan Yao, Tianyu Yu, Ao Zhang, Chongyi Wang, Junbo Cui, Hongji Zhu, Tianchi Cai, Haoyu Li, Weilin Zhao, Zhihui He, et al. 2024 · 2024
Later among the works it cites.
Justice or prejudice? quantifying biases in llm-as-a-judge
Jiayi Ye, Yanbo Wang, Yue Huang, Dongping Chen, Qihui Zhang, Nuno Moniz, Tian Gao, Werner Geyer, Chao Huang, Pin-Yu Chen, et al · 2024
Later among the works it cites.
mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration
Qinghao Ye, Haiyang Xu, Jiabo Ye, Ming Yan, Anwen Hu, Haowei Liu, Qi Qian, Ji Zhang, and Fei Huang. 2024 · 2024
Later among the works it cites.
Synlogue with aizuchi-bot: Investigating the co-adaptive and open-ended interaction paradigm
Kazumi Yoshimura, Dominique Chen, and Olaf Witkowski. 2024 · 2024
Later among the works it cites.
Contextual object detection with multimodal large language models
Yuhang Zang, Wei Li, Jun Han, Kaiyang Zhou, and Chen Change Loy. 2024 · 2024
Later among the works it cites.
Improving accuracy and generalizability via multi-modal large language models collaboration
Shuili Zhang, Hongzhang Mu, and Tingwen Liu. 2024 · 2024
Later among the works it cites.
Sfc: Shared feature calibration in weakly supervised semantic segmentation
Xinqiao Zhao, Feilong Tang, Xiaoyang Wang, and Jimin Xiao. 2024 · 2024
Later among the works it cites.
Ophnet: A large-scale video benchmark for ophthalmic surgical workflow understanding
Ming Hu, Peng Xia, Lin Wang, Siyuan Yan, Feilong Tang, Zhongxing Xu, Yimin Luo, Kaimin Song, Jurgen Leitner, Xuelian Cheng, et al. 2025 · 2025
Closest in time.
Massive values in self-attention modules are the key to contextual knowledge understanding
Mingyu Jin, Kai Mei, Wujiang Xu, Mingjie Sun, Ruixiang Tang, Mengnan Du, Zirui Liu, and Yongfeng Zhang. 2025 · 2025
Closest in time.
A survey of multilingual large language models
Libo Qin, Qiguang Chen, Yuhang Zhou, Zhi Chen, Yinghui Li, Lizi Liao, Min Li, Wanxiang Che, and S Yu Philip. 2025 · 2025
Closest in time.
Can large language models understand preferences in personalized recommendation?
Zhaoxuan Tan, Zinan Zeng, Qingkai Zeng, Zhenyu Wu, Zheyuan Liu, Fengran Mo, and Meng Jiang. 2025 · 2025
Closest in time.
Neighbor does matter: Density-aware contrastive learning for medical semi-supervised segmentation
Feilong Tang, Zhongxing Xu, Ming Hu, Wenxue Li, Peng Xia, Yiheng Zhong, Hanjun Wu, Jionglong Su, and Zongyuan Ge. 2025 · 2025
Closest in time.
Dettoolchain: A new prompting paradigm to unleash detection ability of mllm
Yixuan Wu, Yizhou Wang, Shixiang Tang, Wenhao Wu, Tong He, Wanli Ouyang, Philip Torr, and Jian Wu. 2025 · 2025
Closest in time.
Toward modality gap: Vision prototype learning for weakly-supervised semantic segmentation with clip
Zhongxing Xu, Feilong Tang, Zhe Chen, Yingxue Su, Zhiyi Zhao, Ge Zhang, Jionglong Su, and Zongyuan Ge. 2025 · 2025
Closest in time.
Recent advances and challenges in task-oriented dialog systems
Zheng Zhang, Ryuichi Takanobu, Qi Zhu, MinLie Huang, and XiaoYan Zhu. 2020 · 2027
Closest in time.