Fetching the paper…
Reading the bibliography…
Matrices are exceptionally useful in various fields of study as they provide a convenient framework to organize and manipulate data in a structured manner.
Searching for mobilenetv3, 2019
A. Howard, M. Sandler, G. Chu, L.-C. Chen, B. Chen, M. Tan, W. Wang, Y. Zhu, R. Pang, V. Vasudevan, Q. V. Le, and H. Adam · 1905
Earlier work this paper cites.
On the effectiveness of low-rank matrix factorization for LSTM model compression, 2019
G. I. Winata, A. Madotto, J. Shin, E. J. Barezi, and P. Fung · 1908
Earlier work this paper cites.
The approximation of one matrix by another of lower rank
C. Eckart and G. Young · 1936
Earlier work this paper cites.
Second moments of inverse wishart-matrix elements
V. Siskind · 1972
Earlier work this paper cites.
Multivariate Analysis
K. Mardia, J. Kent, and J. Bibby · 1979
Earlier work this paper cites.
Dithered quantizers
R. Gray and T. Stockham · 1993
Earlier work this paper cites.
Anti-hadamard matrices, coin weighing, threshold gates, and indecomposable hypergraphs
N. Alon and V. H. Vu · 1997
Earlier work this paper cites.
Fast monte carlo algorithms for matrices II: Computing a low-rank approximation to a matrix
P. Drineas, R. Kannan, and M. W. Mahoney · 2006
Earlier work this paper cites.
An improved approximation for the gaussian Q-function
G. K. Karagiannidis and A. S. Lioumpas · 2007
Earlier work this paper cites.
2D & 3D shepp-logan phantom standards for MRI
H. M. Gach, C. Tanase, and F. Boada · 2008
Earlier work this paper cites.
Smallest singular value of a random rectangular matrix
M. Rudelson and R. Vershynin · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Accelerated dynamic MRI exploiting sparsity and low-rank structure: k-t SLR
S. G. Lingala, Y. Hu, E. DiBella, and M. Jacob · 2010
Earlier work this paper cites.
Uncertainty principles and vector quantization
Y. Lyubarskii and R. Vershynin · 2010
Earlier work this paper cites.
A randomized algorithm for the decomposition of matrices
P.-G. Martinsson, V. Rokhlin, and M. Tygert · 2010
Earlier work this paper cites.
Finding structure with randomness: Probabilistic algorithms for constructing approximate matrix decompositions
N. Halko, P. G. Martinsson, and J. A. Tropp · 2011
Earlier work this paper cites.
Introduction to the non-asymptotic analysis of random matrices, 2011
R. Vershynin · 2011
Earlier work this paper cites.
Intrinsic dimensionality explains the effectiveness of language model fine-tuning, 2020
A. Aghajanyan, L. Zettlemoyer, and S. Gupta · 2012
Earlier work this paper cites.
Signal representations with minimum ℓ ∞ \ell_{\infty} -norm
C. Studer, W. Yin, and R. G. Baraniuk · 2012
Earlier work this paper cites.
Learning structured low-rank representations for image classification
Y. Zhang, Z. Jiang, and L. S. Davis · 2013
Earlier work this paper cites.
Speeding up convolutional neural networks with low rank expansions, 2014
M. Jaderberg, A. Vedaldi, and A. Zisserman · 2014
Earlier work this paper cites.
Dimension reduction by random hyperplane tessellations
Y. Plan and R. Vershynin · 2014
Earlier work this paper cites.
Low-rank modeling and its applications in image analysis
X. Zhou, C. Yang, H. Zhao, and W. Yu · 2014
Earlier work this paper cites.
Randomized algorithms for low-rank matrix factorizations: Sharp performance bounds
R. Witten and E. Candès · 2015
Earlier work this paper cites.
RandNLA: Randomized numerical linear algebra
P. Drineas and M. W. Mahoney · 2016
Earlier work this paper cites.
Lecture notes on randomized linear algebra, 2016
M. W. Mahoney · 2016
Earlier work this paper cites.
Convolutional neural networks with low-rank regularization, 2016
C. Tai, T. Xiao, Y. Zhang, X. Wang, and W. E · 2016
Earlier work this paper cites.
QSGD: Communication-efficient SGD via gradient quantization and encoding
D. Alistarh, D. Grubic, J. Li, R. Tomioka, and M. Vojnovic · 2017
Earlier work this paper cites.
Distributed mean estimation with limited communication
A. T. Suresh, F. X. Yu, S. Kumar, and H. B. McMahan · 2017
Cited alongside, same era.
Practical sketching algorithms for low-rank matrix approximation
J. A. Tropp, A. Yurtsever, M. Udell, and V. Cevher · 2017
Cited alongside, same era.
On compressing deep models by low rank and sparse decomposition
X. Yu, T. Liu, X. Wang, and D. Tao · 2017
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2018
Cited alongside, same era.
When and why are pre-trained word embeddings useful for neural machine translation?
Y. Qi, D. S. Sachan, M. Felix, S. J. Padmanabhan, and G. Neubig · 2018
Cited alongside, same era.
Compacter: Efficient low-rank hypercomplex adapter layers
R. Karimi Mahabadi, J. Henderson, and S. Ruder · 2021
Later among the works it cites.
Fast and accurate randomized algorithms for low-rank tensor decompositions
L. Ma and E. Solomonik · 2021
Later among the works it cites.
RATQ: A universal fixed-length quantizer for stochastic optimization
P. Mayekar and H. Tyagi · 2021
Later among the works it cites.
Uncertainty principle for communication compression in distributed and federated learning and the search for an optimal compressor
M. Safaryan, E. Shulgin, and P. Richtárik · 2021
Later among the works it cites.
DRIVE: One-bit distributed mean estimation
S. Vargaftik, R. Ben-Basat, A. Portnoy, G. Mendelson, Y. Ben-Itzhak, and M. Mitzenmacher · 2021
Later among the works it cites.
K-Adapter: Infusing Knowledge into Pre-Trained Models with Adapters
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
High-Dimensional Probability: An Introduction with Applications in Data Science
R. Vershynin · 2018
Cited alongside, same era.
Nonconvex optimization meets low-rank matrix factorization: An overview
Y. Chi, Y. M. Lu, and Y. Chen · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, A. Desmaison, A. Kopf, E. Yang, Z. DeVito, M. Raison, A. Tejani, S. Chilamkurthy, B. Steiner, L. Fang, J. Bai, and S. Chintala · 2019
Cited alongside, same era.
Sentiment analysis based on improved pre-trained word embeddings
S. M. Rezaeinia, R. Rahmani, A. Ghodsi, and H. Veisi · 2019
Cited alongside, same era.
Streaming low-rank matrix approximation with an application to scientific simulation
J. A. Tropp, A. Yurtsever, M. Udell, and V. Cevher · 2019
Cited alongside, same era.
Why are big data matrices approximately low rank?
M. Udell and A. Townsend · 2019
Cited alongside, same era.
Basic tail and concentration bounds , page 21–57
M. J. Wainwright · 2019
Cited alongside, same era.
R. Wang, D. Tang, N. Duan, Z. Wei, X. Huang, J. Ji, G. Cao, D. Jiang, and M. Zhou · 2021
Later among the works it cites.
Global convergence of gradient descent for asymmetric low-rank matrix factorization
T. Ye and S. S. Du · 2021
Later among the works it cites.
Transform quantization for CNN compression
S. Young, Z. Wang, D. Taubman, and B. Girod · 2021
Later among the works it cites.
General low-rank matrix optimization: Geometric analysis and sharper bounds
H. Zhang, Y. Bi, and J. Lavaei · 2021
Later among the works it cites.
LLM.int8 (): 8-bit matrix multiplication for transformers at scale
T. Dettmers, M. Lewis, Y. Belkada, and L. Zettlemoyer · 2022
Later among the works it cites.
Gptq: Accurate post-training quantization for generative pre-trained transformers
E. Frantar, S. Ashkboos, T. Hoefler, and D. Alistarh · 2022
Later among the works it cites.
LoRA: Low-rank adaptation of large language models
E. J. Hu, Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, and W. Chen · 2022
Later among the works it cites.
LlamaIndex, 11 2022
J. Liu · 2022
Later among the works it cites.
The unsurprising effectiveness of pre-trained vision models for control
S. Parisi, A. Rajeswaran, S. Purushwalkam, and A. Gupta · 2022
Later among the works it cites.
FedNL: Making newton-type methods applicable to federated learning
M. Safaryan, R. Islamov, X. Qian, and P. Richtárik · 2022
Later among the works it cites.
Efficient randomized subspace embeddings for distributed optimization under a communication budget
R. Saha, M. Pilanci, and A. J. Goldsmith · 2022
Later among the works it cites.
Correlated quantization for distributed mean estimation and optimization
A. T. Suresh, Z. Sun, J. Ro, and F. Yu · 2022
Later among the works it cites.
KroneckerBERT: Significant compression of pre-trained language models through kronecker decomposition and knowledge distillation
M. Tahaei, E. Charlaix, V. Nia, A. Ghodsi, and M. Rezagholizadeh · 2022
Later among the works it cites.
Eden: Communication-efficient and robust distributed mean estimation for federated learning
S. Vargaftik, R. B. Basat, A. Portnoy, G. Mendelson, Y. B. Itzhak, and M. Mitzenmacher · 2022
Later among the works it cites.
ZeroQuant: Efficient and affordable post-training quantization for large-scale transformers
Z. Yao, R. Yazdani Aminabadi, M. Zhang, X. Wu, C. Li, and Y. He · 2022
Later among the works it cites.
Enterprise Search Engine - Amazon Kendra - AWS, 2023
Amazon AWS · 2023
Closest in time.
Hardware-limited non-uniform task-based quantizers
N. I. Bernardo, J. Zhu, Y. C. Eldar, and J. Evans · 2023
Closest in time.
marqo-ai/marqo: Vector search for humans, 2023
Marqo · 2023
Closest in time.
Vector database for vector search | pinecone
Pinecone · 2023
Closest in time.
Low precision representations for high dimensional models
R. Saha, M. Pilanci, and A. J. Goldsmith · 2023
Closest in time.
Llama: Open and efficient foundation language models
H. Touvron, T. Lavril, G. Izacard, X. Martinet, M.-A. Lachaux, T. Lacroix, B. Rozière, N. Goyal, E. Hambro, F. Azhar, et al · 2023
Closest in time.
M. Valipour, M. Rezagholizadeh, I. Kobyzev, and A. Ghodsi · 2023
Closest in time.