Fetching the paper…
Reading the bibliography…
While multimodal foundation models can now natively work with data beyond text, they remain underutilized in analyzing the considerable amounts of multi-dimensional time-series data in fields like healthcare, finance, and social sciences, representing a missed opportunity for richer, data-driven insights.
Individual comparisons by ranking methods
Frank Wilcoxon · 1945
Earlier work this paper cites.
Multiple significance tests: the bonferroni method
J Martin Bland and Douglas G Altman · 1995
Earlier work this paper cites.
Readings in information visualization: using vision to think
Stuart K Card, Jock Mackinlay, and Ben Shneiderman · 1999
Earlier work this paper cites.
Smart devices are different: Assessing and mitigatingmobile sensing heterogeneities for activity recognition
Allan Stisen, Henrik Blunck, Sourav Bhattacharya, Thor Siiger Prentow, Mikkel Baun Kjærgaard, Anind Dey, Tobias Sonne, and Mads Møller Jensen · 2015
Earlier work this paper cites.
Cognitive stages in visual data exploration
M Adil Yalçin, Niklas Elmqvist, and Benjamin B Bederson · 2016
Earlier work this paper cites.
A comparison of accuracy of fall detection algorithms (threshold-based vs. machine learning) using waist-mounted tri-axial accelerometer signals from a comprehensive set of falls and non-fall trials
Omar Aziz, Magnus Musngi, Edward J Park, Greg Mori, and Stephen N Robinovitch · 2017
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Earlier work this paper cites.
Deeptranshhar: Inter-subjects heterogeneous activity recognition approach in the non-identical environment using wearable sensors
Prabhat Kumar and Suresh Selvam · 2022
Earlier work this paper cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Earlier work this paper cites.
Jolt: Jointly learned representations of language and time-series
Yifu Cai, Mononito Goswami, Arjun Choudhry, Arvind Srinivasan, and Artur Dubrawski · 2023
Earlier work this paper cites.
A decoder-only foundation model for time-series forecasting
Abhimanyu Das, Weihao Kong, Rajat Sen, and Yichen Zhou · 2023
Earlier work this paper cites.
Gemini: a family of highly capable multimodal models
Gemini Team, Rohan Anil, Sebastian Borgeaud, Yonghui Wu, Jean-Baptiste Alayrac, Jiahui Yu, Radu Soricut, Johan Schalkwyk, Andrew M Dai, Anja Hauth, et al · 2023
Cited alongside, same era.
Imagebind: One embedding space to bind them all
Rohit Girdhar, Alaaeldin El-Nouby, Zhuang Liu, Mannat Singh, Kalyan Vasudev Alwala, Armand Joulin, and Ishan Misra · 2023
Cited alongside, same era.
DePlot: One-shot visual language reasoning by plot-to-table translation
Fangyu Liu, Julian Eisenschlos, Francesco Piccinno, Syrine Krichene, Chenxi Pang, Kenton Lee, Mandar Joshi, Wenhu Chen, Nigel Collier, and Yasemin Altun · 2023
Cited alongside, same era.
IMU2CLIP: Language-grounded Motion Sensor Translation with Multimodal Contrastive Learning
Seungwhan Moon, Andrea Madotto, Zhaojiang Lin, Aparajita Saraf, Amy Bearman, and Babak Damavandi · 2023
Cited alongside, same era.
Leveraging vision-language models for granular market change prediction
Moment: A family of open time-series foundation models
Mononito Goswami, Konrad Szafer, Arjun Choudhry, Yifu Cai, Shuo Li, and Artur Dubrawski · 2024
Closest in time.
Large language models are zero-shot time series forecasters
Nate Gruver, Marc Finzi, Shikai Qiu, and Andrew G Wilson · 2024
Closest in time.
OpenAI API Pricing
OpenAI · 2024
Closest in time.
Llm evaluators recognize and favor their own generations
Arjun Panickssery, Samuel R Bowman, and Shi Feng · 2024
Closest in time.
Spiqa: A dataset for multimodal question answering on scientific papers
Shraman Pramanick, Rama Chellappa, and Subhashini Venugopalan · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Christopher Wimmer and Navid Rekabsaz · 2023
Cited alongside, same era.
Chronos: Learning the language of time series
Abdul Fatir Ansari, Lorenzo Stella, Caner Turkmen, Xiyuan Zhang, Pedro Mercado, Huibin Shen, Oleksandr Shchur, Syama Sundar Rangapuram, Sebastian Pineda Arango, Shubham Kapoor, et al · 2024
Cited alongside, same era.
Timeseriesexam: A time series understanding exam
Yifu Cai, Arjun Choudhry, Mononito Goswami, and Artur Dubrawski · 2024
Cited alongside, same era.
Medtsllm: Leveraging llms for multimodal medical time series analysis
Nimeesha Chan, Felix Parker, William Bennett, Tianyi Wu, Mung Yao Jia, James Fackler, and Kimia Ghobadi · 2024
Cited alongside, same era.
VLMs Aren’t Blind
Daniel Corin · 2024
Cited alongside, same era.
Towards a personal health large language model
Justin Cosentino, Anastasiya Belyaeva, Xin Liu, Nicholas A Furlotte, Zhun Yang, Chace Lee, Erik Schenck, Yojan Patel, Jian Cui, Logan Douglas Schneider, et al · 2024
Cited alongside, same era.
Gemini API Cost
Google · 2024
Cited alongside, same era.
Pooyan Rahmanzadehgervi, Logan Bolton, Mohammad Reza Taesiri, and Anh Totti Nguyen · 2024
Closest in time.
The first step is the hardest: pitfalls of representing and tokenizing temporal data for large language models
Dimitris Spathis and Fahim Kawsar · 2024
Closest in time.
Charxiv: Charting gaps in realistic chart understanding in multimodal llms
Zirui Wang, Mengzhou Xia, Luxi He, Howard Chen, Yitao Liu, Richard Zhu, Kaiqu Liang, Xindi Wu, Haotian Liu, Sadhika Malladi, et al · 2024
Closest in time.
Unified training of universal time series forecasting transformers
Gerald Woo, Chenghao Liu, Akshat Kumar, Caiming Xiong, Silvio Savarese, and Doyen Sahoo · 2024
Closest in time.
Large language models for time series: A survey
Xiyuan Zhang, Ranak Roy Chowdhury, Rajesh K. Gupta, and Jingbo Shang · 2024
Closest in time.