Fetching the paper…
Reading the bibliography…
Batch data analytics is a growing application for Large Language Models (LLMs).
Multi-relational data mining: an introduction
Džeroski, S · 2003
Earlier work this paper cites.
Bootstrap confidence interval
Wilcox, R. R · 2003
Earlier work this paper cites.
Cords: Automatic discovery of correlations and soft functional dependencies
Ilyas, I. F., Markl, V., Haas, P., Brown, P., and Aboulnaga, A · 2004
Earlier work this paper cites.
Seeing stars: Exploiting class relationships for sentiment categorization with respect to rating scales
Pang, B. and Lee, L · 2005
Earlier work this paper cites.
Database cracking
Idreos, S., Kersten, M. L., Manegold, S., et al · 2007
Earlier work this paper cites.
Reordering columns for smaller indexes
Lemire, D. and Kaser, O · 2011
Earlier work this paper cites.
Model-based multidimensional clustering of categorical data
Chen, T., Zhang, N. L., Liu, T., Poon, K. M., and Wang, Y · 2012
Earlier work this paper cites.
Learning attitudes and attributes from multi-aspect reviews, 2012
McAuley, J., Leskovec, J., and Jurafsky, D · 2012
Earlier work this paper cites.
Resilient distributed datasets: A { \{ Fault-Tolerant } \} abstraction for { \{ In-Memory } \} cluster computing
Zaharia, M., Chowdhury, M., Das, T., Dave, A., Ma, J., McCauly, M., Franklin, M. J., Shenker, S., and Stoica, I · 2012
Earlier work this paper cites.
Spark sql: Relational data processing in spark
Armbrust, M., Xin, R. S., Lian, C., Huai, Y., Liu, D., Bradley, J. K., Meng, X., Kaftan, T., Franklin, M. J., Ghodsi, A., and Zaharia, M · 2015
Earlier work this paper cites.
Ups and downs: Modeling the visual evolution of fashion trends with one-class collaborative filtering
He, R. and McAuley, J · 2016
Earlier work this paper cites.
Squad: 100,000+ questions for machine comprehension of text, 2016
Rajpurkar, P., Zhang, J., Lopyrev, K., and Liang, P · 2016
Earlier work this paper cites.
Noscope: optimizing neural network queries over video at scale
Kang, D., Emmons, J., Abuzaid, F., Bailis, P., and Zaharia, M · 2017
Earlier work this paper cites.
Accelerating machine learning inference with probabilistic predicates
Lu, Y., Chowdhery, A., Kandula, S., and Chaudhuri, S · 2018
Cited alongside, same era.
C-store: a column-oriented dbms
Stonebraker, M., Abadi, D. J., Batkin, A., Chen, X., Cherniack, M., Ferreira, M., Lau, E., Lin, A., Madden, S., O’Neil, E., et al · 2018
Cited alongside, same era.
Fever: a large-scale dataset for fact extraction and verification, 2018
Thorne, J., Vlachos, A., Christodoulopoulos, C., and Mittal, A · 2018
Cited alongside, same era.
Billion-scale similarity search with GPUs
Johnson, J., Douze, M., and Jégou, H · 2019
Cited alongside, same era.
Delta lake: high-performance acid table storage over cloud object stores
Armbrust, M., Das, T., Sun, L., Yavuz, B., Zhu, S., Murthy, M., Torres, J., van Hovell, H., Ionescu, A., Łuszczak, A., et al · 2020
Cited alongside, same era.
https://aws.amazon.com/blogs/big-data/large-language-models-for-sentiment-analysis-with-amazon-redshift-ml-preview/
Large Language Models for sentiment analysis with Amazon Redshift ML (Preview) — Amazon Web Services — aws.amazon.com · 2024
Closest in time.
https://docs.databricks.com/en/large-language-models/ai-functions.html
AI Functions on Databricks — docs.databricks.com · 2024
Closest in time.
https://cloud.google.com/blog/products/ai-machine-learning/llm-with-vertex-ai-only-using-sql-queries-in-bigquery
LLM with Vertex AI only using SQL queries in BigQuery — Google Cloud Blog — cloud.google.com · 2024
Closest in time.
https://www.anthropic.com/news/prompt-caching , 2024
Prompt caching with claude · 2024
Closest in time.
https://ai.google.dev/gemini-api/docs/caching?lang=python , 2024
Context caching · 2024
Closest in time.
URL https://ai.meta.com/blog/meta-llama-3/
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lewis, P., Perez, E., Piktus, A., Petroni, F., Karpukhin, V., Goyal, N., Küttler, H., Lewis, M., tau Yih, W., Rocktäschel, T., Riedel, S., and Kiela, D · 2021
Cited alongside, same era.
LangChain, October 2022
Chase, H · 2022
Cited alongside, same era.
Orca: A distributed serving system for Transformer-Based generative models
Yu, G.-I., Jeong, J. S., Kim, G.-W., Kim, S., and Chun, B.-G · 2022
Cited alongside, same era.
Text Generation Inference, 2023
Huggingface · 2023
Cited alongside, same era.
Efficient memory management for large language model serving with pagedattention
Kwon, W., Li, Z., Zhuang, S., Sheng, Y., Zheng, L., Yu, C. H., Gonzalez, J., Zhang, H., and Stoica, I · 2023
Cited alongside, same era.
Towards general text embeddings with multi-stage contrastive learning
Li, Z., Zhang, X., Zhang, Y., Long, D., Xie, P., and Zhang, M · 2023
Cited alongside, same era.
Attention is all you need, 2023
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, L., and Polosukhin, I · 2023
Cited alongside, same era.
Apr 2024 · 2024
Closest in time.
Prompt cache: Modular attention reuse for low-latency inference, 2024
Gim, I., Chen, G., seob Lee, S., Sarda, N., Khandelwal, A., and Zhong, L · 2024
Closest in time.
Hydragen: High-throughput llm inference with shared prefixes, 2024
Juravsky, J., Brown, B., Ehrlich, R., Fu, D. Y., Ré, C., and Mirhoseini, A · 2024
Closest in time.
Can llm already serve as a database interface? a big bench for large-scale database grounded text-to-sqls
Li, J., Hui, B., Qu, G., Yang, J., Li, B., Li, B., Wang, B., Qin, B., Geng, R., Huo, N., et al · 2024
Closest in time.
Pdmx: A large-scale public domain musicxml dataset for symbolic music processing, 2024
Long, P., Novack, Z., Berg-Kirkpatrick, T., and McAuley, J · 2024
Closest in time.
Pricing — openai.com
OpenAI · 2024
Closest in time.
Lotus: Enabling semantic queries with llms over tables of unstructured and structured data, 2024
Patel, L., Jha, S., Guestrin, C., and Zaharia, M · 2024
Closest in time.
Cascade inference: Memory bandwidth efficient shared prefix batch decoding, February 2024
Ye, Z., Lai, R., Lu, B.-R., Lin, C.-Y., Zheng, S., Chen, L., Chen, T., and Ceze, L · 2024
Closest in time.