Fetching the paper…
Reading the bibliography…
While convolution and self-attention mechanisms have dominated architectural design in deep learning, this survey examines a fundamental yet understudied primitive: the Hadamard product.
T. Zhang, H.-Y. Tseng, L. Jiang, W. Yang, H. Lee, and I. Essa, “Text as neural operator: Image manipulation by text instruction,” in Proceedings of the 29th ACM International Conference on Multimedia , 2021, pp. 1893–1902
1902
Earlier work this paper cites.
J. Hadamard, Leçons sur la propagation des ondes et les équations de l’hydrodynamique . A. Hermann, 1903
1903
Earlier work this paper cites.
J. Schur, “Bemerkungen zur theorie der beschränkten bilinearformen mit unendlich vielen veränderlichen.” 1911
1911
Earlier work this paper cites.
F. L. Hitchcock, “The expression of a tensor or a polyadic as a sum of products,” J. of Math. and Phys. , vol. 6, no. 1-4, pp. 164–189, 1927
1927
Earlier work this paper cites.
H. Harold, “Relations between two sets of variates,” Biometrika , vol. 28, no. 3/4, p. 321, 1936
1936
Earlier work this paper cites.
R. B. Cattell, ““parallel proportional profiles” and other principles for determining the choice of factors by rotation,” Psychometrika , vol. 9, no. 4, pp. 267–283, 1944
1944
Earlier work this paper cites.
P. R. Halmos, Finite dimensional vector spaces . Princeton University Press, 1948, no. 7
1948
Earlier work this paper cites.
M. H. Stone, “The generalized Weierstrass approximation theorem,” Math. Mag. , vol. 21, no. 5, pp. 237–254, 1948
1948
Earlier work this paper cites.
B. Vinograde, “Canonical positive definite matrices under internal linear transformations,” Proceedings of the American Mathematical Society , vol. 1, no. 2, pp. 159–161, 1950
1950
Earlier work this paper cites.
W. H. Sumby and I. Pollack, “Visual contribution to speech intelligibility in noise,” The journal of the acoustical society of america , vol. 26, no. 2, pp. 212–215, 1954
1954
Earlier work this paper cites.
A. H. Land and A. G. Doig, “An automatic method of solving discrete programming problems,” Econometrica , vol. 28, no. 3, pp. 497–520, 1960
1960
Earlier work this paper cites.
J. H. Ahlberg, E. N. Nilson, and J. L. Walsh, “The theory of splines and their applications,” Mathematics in science and engineering , 1967
1967
Earlier work this paper cites.
C. Khatri and C. R. Rao, “Solutions to some functional equations and their applications to characterization of probability distributions,” Sankhyā: The Indian Journal of Statistics, Series A , pp. 167–180, 1968
1968
Earlier work this paper cites.
C. S. Ballantine, “On the hadamard product,” Mathematische Zeitschrift , vol. 105, no. 5, pp. 365–366, Oct. 1968
1968
Earlier work this paper cites.
J. D. Carroll and J.-J. Chang, “Analysis of individual differences in multidimensional scaling via an n-way generalization of “Eckart-Young” decomposition,” Psychometrika , vol. 35, pp. 283–319, 1970
1970
Earlier work this paper cites.
A. G. Ivakhnenko, “Polynomial theory of complex systems,” Transactions on Systems, Man, and Cybernetics , no. 4, pp. 364–378, 1971
1971
Earlier work this paper cites.
J. R. Kettenring, “Canonical analysis of several sets of variables,” Biometrika , vol. 58, no. 3, pp. 433–451, 1971
1971
Earlier work this paper cites.
G. P. Styan, “Hadamard products and multivariate statistical analysis,” Linear algebra and its applications , vol. 6, pp. 217–240, 1973
1973
Earlier work this paper cites.
H. McGurk and J. MacDonald, “Hearing lips and seeing voices,” Nature , vol. 264, no. 5588, pp. 746–748, 1976
1976
Earlier work this paper cites.
G. E. Hinton and T. J. Sejnowski, “Optimal perceptual inference,” in Conference on Computer Vision and Pattern Recognition (CVPR) , vol. 448, 1983, pp. 448–453
1983
Earlier work this paper cites.
E. D. Petajan, “Automatic lipreading to enhance speech recognition (speech reading),” Ph.D. dissertation, 1984
1984
Earlier work this paper cites.
D. Psaltis and C. H. Park, “Nonlinear discriminant functions and associative memories,” in AIP conference Proceedings , vol. 151, no. 1. American Institute of Physics, 1986, pp. 370–375
1986
Earlier work this paper cites.
T. J. Sejnowski, “Higher-order boltzmann machines,” in AIP Conference Proceedings , vol. 151, no. 1. American Institute of Physics, 1986, pp. 398–403
1986
Earlier work this paper cites.
C. L. Giles and T. Maxwell, “Learning, invariance, and generalization in high-order neural networks,” Applied optics , vol. 26, no. 23, pp. 4972–4978, 1987
1987
Earlier work this paper cites.
B. P. Yuhas, M. H. Goldstein, and T. J. Sejnowski, “Integration of acoustic and visual speech signals using neural networks,” Communications Magazine , vol. 27, no. 11, pp. 65–71, 1989
1989
Earlier work this paper cites.
C. R. Johnson, Matrix theory and applications . American Mathematical Soc., 1990, vol. 40
1990
Earlier work this paper cites.
Y. Shin and J. Ghosh, “The pi-sigma network: An efficient higher-order neural network for pattern classification and function approximation,” in International Joint Conference on Neural Networks (IJCNN) , vol. 1, 1991, pp. 13–18
1991
Earlier work this paper cites.
T. Ando, “Majorization relations for hadamard products,” Linear algebra and its applications , vol. 223, pp. 57–64, 1995
1995
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
S.-K. Oh, W. Pedrycz, and B.-J. Park, “Polynomial neural networks architecture: analysis and design,” Computers & Electrical Engineering , vol. 29, no. 6, pp. 703–725, 2003
2003
Earlier work this paper cites.
H.-B. Li, T.-Z. Huang, S.-Q. Shen, and H. Li, “Lower bounds for the minimum eigenvalue of hadamard product of an m-matrix and its inverse,” Linear Algebra and its applications , vol. 420, no. 1, pp. 235–247, 2007
2007
Earlier work this paper cites.
T. G. Kolda and B. W. Bader, “Tensor decompositions and applications,” SIAM review , vol. 51, no. 3, pp. 455–500, 2009
2009
Earlier work this paper cites.
A. Croitor-Sava, M. Martinez-Bisbal, T. Laudadio, J. Piquer, B. Celda, A. Heerschap, D. Sima, and S. Van Huffel, “Fusing in vivo and ex vivo nmr sources of information for brain tumor classification,” Measurement Science and Technology , vol. 22, no. 11, p. 114012, 2011
2011
Earlier work this paper cites.
2012
Earlier work this paper cites.
F. J. Caro-Lopera, V. Leiva, and N. Balakrishnan, “Connection between the hadamard and matrix products with an application to matrix-variate birnbaum–saunders distributions,” Journal of Multivariate Analysis , vol. 104, no. 1, pp. 126–139, 2012
2012
Earlier work this paper cites.
S. Nikol’skii, Analysis III: Spaces of Differentiable Functions , ser. Encyclopaedia of Mathematical Sciences. Springer Berlin Heidelberg, 2013
2013
Earlier work this paper cites.
L. Wan, M. Zeiler, S. Zhang, Y. Le Cun, and R. Fergus, “Regularization of neural networks using dropconnect,” in International Conference on Machine Learning (ICML) , 2013, pp. 1058–1066
2013
Earlier work this paper cites.
J. Ba and B. Frey, “Adaptive dropout for training deep neural networks,” Advances in neural information processing systems (NeurIPS) , vol. 26, 2013
2013
Earlier work this paper cites.
K. Cho, B. van Merriënboer, D. Bahdanau, and Y. Bengio, “On the properties of neural machine translation: Encoder–decoder approaches,” in Proceedings of SSST-8, Eighth Workshop on Syntax, Semantics and Structure in Statistical Translation , 2014, pp. 103–111
2014
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in Advances in neural information processing systems (NeurIPS) , 2014
2014
Earlier work this paper cites.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: a simple way to prevent neural networks from overfitting,” Journal of Machine Learning Research , vol. 15, no. 1, pp. 1929–1958, 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
R. Livni, S. Shalev-Shwartz, and O. Shamir, “On the computational efficiency of training neural networks,” in Advances in neural information processing systems (NeurIPS) , 2014, pp. 855–863
2014
Earlier work this paper cites.
S. Shalev-Shwartz and S. Ben-David, Understanding machine learning: From theory to algorithms . Cambridge university press, 2014
2014
Earlier work this paper cites.
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. Goodfellow, and R. Fergus, “Intriguing properties of neural networks,” in International Conference on Learning Representations (ICLR) , 2014
2014
Earlier work this paper cites.
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh, “Vqa: Visual question answering,” in International Conference on Computer Vision (ICCV) , 2015, pp. 2425–2433
2015
Earlier work this paper cites.
M. Dalla Mura, S. Prasad, F. Pacifici, P. Gamba, J. Chanussot, and J. A. Benediktsson, “Challenges and opportunities of multimodality and data fusion in remote sensing,” Proceedings of the IEEE , vol. 103, no. 9, pp. 1585–1601, 2015
2015
Earlier work this paper cites.
R. K. Srivastava, K. Greff, and J. Schmidhuber, “Training very deep networks,” Advances in neural information processing systems (NeurIPS) , vol. 28, 2015
2015
Earlier work this paper cites.
D. Lahat, T. Adali, and C. Jutten, “Multimodal data fusion: an overview of methods, challenges, and prospects,” Proceedings of the IEEE , vol. 103, no. 9, pp. 1449–1477, 2015
2015
Earlier work this paper cites.
G. Carneiro, T. Peng, C. Bayer, and N. Navab, “Weakly-supervised structured output learning with flexible and latent graphs using high-order loss functions,” in International Conference on Computer Vision (ICCV) , 2015, pp. 648–656
2015
Earlier work this paper cites.
I. J. Goodfellow, J. Shlens, and C. Szegedy, “Explaining and harnessing adversarial examples,” in International Conference on Learning Representations (ICLR) , 2015
2015
Earlier work this paper cites.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” in International Conference on Machine Learning (ICML) , 2015
2015
Earlier work this paper cites.
D. Pathak, P. Krahenbuhl, J. Donahue, T. Darrell, and A. A. Efros, “Context encoders: Feature learning by inpainting,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 2536–2544
2016
Earlier work this paper cites.
A. Fukui, D. H. Park, D. Yang, A. Rohrbach, T. Darrell, and M. Rohrbach, “Multimodal compact bilinear pooling for visual question answering and visual grounding,” in Empirical Methods in Natural Language Processing (EMNLP) , 2016
2016
Earlier work this paper cites.
C. Xiong, S. Merity, and R. Socher, “Dynamic memory networks for visual and textual question answering,” in International Conference on Machine Learning (ICML) , 2016, pp. 2397–2406
2016
Earlier work this paper cites.
J.-H. Kim, S.-W. Lee, D. Kwak, M.-O. Heo, J. Kim, J.-W. Ha, and B.-T. Zhang, “Multimodal residual learning for visual qa,” in Advances in neural information processing systems (NeurIPS) , vol. 29, 2016
2016
Earlier work this paper cites.
R. Li and J. Jia, “Visual question answering with question representation update (qru),” in Advances in neural information processing systems (NeurIPS) , vol. 29, 2016
2016
Earlier work this paper cites.
L. Drumetz, M.-A. Veganzones, S. Henrot, R. Phlypo, J. Chanussot, and C. Jutten, “Blind hyperspectral unmixing using an extended linear mixing model to address spectral variability,” IEEE Transactions in Image Processing (TIP) , vol. 25, no. 8, pp. 3890–3905, 2016
2016
Earlier work this paper cites.
C. Murdock, Z. Li, H. Zhou, and T. Duerig, “Blockout: Dynamic model selection for hierarchical deep networks,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 770–778
2016
Earlier work this paper cites.
Y. N. Dauphin, A. Fan, M. Auli, and D. Grangier, “Language modeling with gated convolutional networks,” in International Conference on Machine Learning (ICML) , 2017, pp. 933–941
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in Advances in neural information processing systems (NeurIPS) , 2017, pp. 5998–6008
2017
Earlier work this paper cites.
J. Shi, X. Zheng, Y. Li, Q. Zhang, and S. Ying, “Multimodal neuroimaging feature learning with multimodal stacked deep polynomial networks for diagnosis of alzheimer’s disease,” IEEE journal of biomedical and health informatics , vol. 22, no. 1, pp. 173–183, 2017
2017
Earlier work this paper cites.
N. D. Sidiropoulos, L. De Lathauwer, X. Fu, K. Huang, E. E. Papalexakis, and C. Faloutsos, “Tensor decomposition for signal processing and machine learning,” IEEE Transactions on Signal Processing , vol. 65, no. 13, pp. 3551–3582, 2017
2017
Earlier work this paper cites.
Y. Wang, L. Xie, C. Liu, S. Qiao, Y. Zhang, W. Zhang, Q. Tian, and A. Yuille, “Sort: Second-order response transform for visual recognition,” in International Conference on Computer Vision (ICCV) , 2017, pp. 1359–1368
2017
Earlier work this paper cites.
X. Dong, J. Huang, Y. Yang, and S. Yan, “More is less: A more complicated network with less inference complexity,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 5840–5848
2017
Earlier work this paper cites.
X. Huang and S. Belongie, “Arbitrary style transfer in real-time with adaptive instance normalization,” in International Conference on Computer Vision (ICCV) , 2017, pp. 1501–1510
2017
Earlier work this paper cites.
X. Chu, W. Yang, W. Ouyang, C. Ma, A. L. Yuille, and X. Wang, “Multi-context attention for human pose estimation,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 1831–1840
2017
Earlier work this paper cites.
H. Ben-Younes, R. Cadene, M. Cord, and N. Thome, “Mutan: Multimodal tucker fusion for visual question answering,” in International Conference on Computer Vision (ICCV) , 2017, pp. 2612–2620
2017
Earlier work this paper cites.
H. Zhang, Z. Kyaw, J. Yu, and S.-F. Chang, “Ppr-fcn: Weakly supervised visual relation detection via parallel pairwise r-fcn,” in International Conference on Computer Vision (ICCV) , 2017, pp. 4233–4241
2017
Earlier work this paper cites.
H. De Vries, F. Strub, J. Mary, H. Larochelle, O. Pietquin, and A. C. Courville, “Modulating early visual processing by language,” in Advances in neural information processing systems (NeurIPS) , 2017, pp. 6594–6604
2017
Earlier work this paper cites.
H. Nam, J.-W. Ha, and J. Kim, “Dual attention networks for multimodal reasoning and matching,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 299–307
2017
Earlier work this paper cites.
J. Yang, A. Kannan, D. Batra, and D. Parikh, “Lr-gan: Layered recursive generative adversarial networks for image generation,” in International Conference on Learning Representations (ICLR) , 2017
2017
Earlier work this paper cites.
J. G. Zilly, R. K. Srivastava, J. Koutnık, and J. Schmidhuber, “Recurrent highway networks,” in International Conference on Machine Learning (ICML) , 2017, pp. 4189–4198
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
M. Raghu, B. Poole, J. Kleinberg, S. Ganguli, and J. Sohl-Dickstein, “On the expressive power of deep neural networks,” in International Conference on Machine Learning (ICML) , 2017
2017
Earlier work this paper cites.
C. Zhang, S. Bengio, M. Hardt, B. Recht, and O. Vinyals, “Understanding deep learning requires rethinking generalization,” in International Conference on Learning Representations (ICLR) , 2017
2017
Earlier work this paper cites.
M. Cisse, P. Bojanowski, E. Grave, Y. Dauphin, and N. Usunier, “Parseval networks: Improving robustness to adversarial examples,” in International Conference on Machine Learning (ICML) , 2017
2017
Earlier work this paper cites.
D. Kressner and L. Perisa, “Recompression of hadamard products of tensors in tucker format,” SIAM Journal on Scientific Computing , vol. 39, no. 5, pp. A1879–A1902, 2017
2017
Earlier work this paper cites.
J. Hu, L. Shen, and G. Sun, “Squeeze-and-excitation networks,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 7132–7141
2018
Earlier work this paper cites.
J. Yu, Z. Lin, J. Yang, X. Shen, X. Lu, and T. S. Huang, “Generative image inpainting with contextual attention,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 5505–5514
2018
Earlier work this paper cites.
Y. Xu, Q. Kong, W. Wang, and M. D. Plumbley, “Large-scale weakly supervised audio classification using gated convolutional neural network,” in International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2018, pp. 121–125
2018
Earlier work this paper cites.
J. Hu, L. Shen, S. Albanie, G. Sun, and A. Vedaldi, “Gather-excite: Exploiting feature context in convolutional neural networks,” Advances in neural information processing systems (NeurIPS) , vol. 31, 2018
2018
Earlier work this paper cites.
S. Woo, J. Park, J.-Y. Lee, and I. S. Kweon, “Cbam: Convolutional block attention module,” in European Conference on Computer Vision (ECCV) , 2018, pp. 3–19
2018
Earlier work this paper cites.
Y. Zhang, K. Li, K. Li, L. Wang, B. Zhong, and Y. Fu, “Image super-resolution using very deep residual channel attention networks,” in European Conference on Computer Vision (ECCV) , 2018, pp. 286–301
2018
Earlier work this paper cites.
H. Zhang, K. Dana, J. Shi, Z. Zhang, X. Wang, A. Tyagi, and A. Agrawal, “Context encoding for semantic segmentation,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 7151–7160
2018
Earlier work this paper cites.
D. Teney, P. Anderson, X. He, and A. Van Den Hengel, “Tips and tricks for visual question answering: Learnings from the 2017 challenge,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 4223–4232
2018
Earlier work this paper cites.
P. Anderson, X. He, C. Buehler, D. Teney, M. Johnson, S. Gould, and L. Zhang, “Bottom-up and top-down attention for image captioning and visual question answering,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 6077–6086
2018
Earlier work this paper cites.
L. Zhou, Y. Zhou, J. J. Corso, R. Socher, and C. Xiong, “End-to-end dense video captioning with masked transformer,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 8739–8748
2018
Earlier work this paper cites.
Z. Yu, J. Yu, C. Xiang, J. Fan, and D. Tao, “Beyond bilinear: Generalized multimodal factorized high-order pooling for visual question answering,” IEEE Transactions on Neural Networks and Learning Systems (T-NN) , vol. 29, no. 12, pp. 5947–5959, 2018
2018
Earlier work this paper cites.
C. Deng, Q. Wu, Q. Wu, F. Hu, F. Lyu, and M. Tan, “Visual grounding via accumulated attention,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 7746–7755
2018
Earlier work this paper cites.
B. Duke and G. W. Taylor, “Generalized hadamard-product fusion operators for visual question answering,” in 2018 15th Conference on Computer and Robot Vision (CRV) . IEEE, 2018, pp. 39–46
2018
Earlier work this paper cites.
C. Ma, C. Shen, A. Dick, Q. Wu, P. Wang, A. van den Hengel, and I. Reid, “Visual question answering with memory-augmented networks,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 6975–6984
2018
Earlier work this paper cites.
S. M. Kazemi and D. Poole, “Simple embedding for link prediction in knowledge graphs,” Advances in neural information processing systems (NeurIPS) , vol. 31, 2018
2018
Earlier work this paper cites.
S. A. Hasan, Y. Ling, O. Farri, J. Liu, H. Müller, and M. Lungren, “Overview of imageclef 2018 medical domain visual question answering task,” 10-14 September 2018, Tech. Rep., 2018
2018
Earlier work this paper cites.
G. Liu, F. A. Reda, K. J. Shih, T.-C. Wang, A. Tao, and B. Catanzaro, “Image inpainting for irregular holes using partial convolutions,” in European Conference on Computer Vision (ECCV) , 2018, pp. 85–100
2018
Earlier work this paper cites.
A. Radford, K. Narasimhan, T. Salimans, I. Sutskever et al. , “Improving language understanding by generative pre-training,” 2018
2018
Earlier work this paper cites.
X. Chen, J. H. Liew, W. Xiong, C.-K. Chui, and S.-H. Ong, “Focus, segment and erase: an efficient network for multi-label brain tumor segmentation,” in European Conference on Computer Vision (ECCV) , 2018, pp. 654–669
2018
Earlier work this paper cites.
A. Mallya, D. Davis, and S. Lazebnik, “Piggyback: Adapting a single network to multiple tasks by learning to mask weights,” in European Conference on Computer Vision (ECCV) , 2018, pp. 67–82
2018
Earlier work this paper cites.
S. Yan, Y. Xiong, and D. Lin, “Spatial temporal graph convolutional networks for skeleton-based action recognition,” in AAAI Conference on Artificial Intelligence , 2018
2018
Earlier work this paper cites.
Y. Kong and T. Yu, “A graph-embedded deep feedforward network for disease outcome classification and feature selection using gene expression data,” Bioinformatics , vol. 34, no. 21, pp. 3727–3737, 2018
2018
Earlier work this paper cites.
M. Ravanelli, P. Brakel, M. Omologo, and Y. Bengio, “Light gated recurrent units for speech recognition,” IEEE Transactions on Emerging Topics in Computational Intelligence , vol. 2, no. 2, pp. 92–102, 2018
2018
Earlier work this paper cites.
S. Li, W. Li, C. Cook, C. Zhu, and Y. Gao, “Independently recurrent neural network (indrnn): Building a longer and deeper rnn,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 5457–5466
2018
Earlier work this paper cites.
A. Jacot, F. Gabriel, and C. Hongler, “Neural tangent kernel: Convergence and generalization in neural networks,” in Advances in neural information processing systems (NeurIPS) , 2018
2018
Earlier work this paper cites.
Y. Tsuzuku, I. Sato, and M. Sugiyama, “Lipschitz-margin training: Scalable certification of perturbation invariance for deep neural networks,” in Advances in neural information processing systems (NeurIPS) , 2018
2018
Earlier work this paper cites.
A. Virmaux and K. Scaman, “Lipschitz regularity of deep neural networks: analysis and efficient estimation,” in Advances in neural information processing systems (NeurIPS) , 2018
2018
Earlier work this paper cites.
S. Sahoo, C. Lampert, and G. Martius, “Learning equations for extrapolation and control,” in International Conference on Machine Learning (ICML) , 2018, pp. 4442–4450
2018
Earlier work this paper cites.
S. Du and J. Lee, “On the power of over-parametrization in neural networks with quadratic activation,” in International Conference on Machine Learning (ICML) , 2018
2018
Cited alongside, same era.
L. Sun, B. Zheng, J. Zhou, and H. Yan, “Some inequalities for the hadamard product of tensors,” Linear and Multilinear Algebra , vol. 66, no. 6, pp. 1199–1214, 2018
2018
Cited alongside, same era.
T. Karras, S. Laine, and T. Aila, “A style-based generator architecture for generative adversarial networks,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2019
2019
Cited alongside, same era.
T. Park, M.-Y. Liu, T.-C. Wang, and J.-Y. Zhu, “Semantic image synthesis with spatially-adaptive normalization,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 2337–2346
2019
Cited alongside, same era.
A.-A. Liu, Y. Zhai, N. Xu, W. Nie, W. Li, and Y. Zhang, “Region-aware image captioning via interaction learning,” IEEE Transactions on Circuits and Systems for Video Technology , 2021
2021
Later among the works it cites.
C. Rodriguez-Opazo, E. Marrese-Taylor, B. Fernando, H. Li, and S. Gould, “Dori: discovering object relationships for moment localization of a natural language query in a video,” in Winter Conference on Applications of Computer Vision (WACV) , 2021, pp. 1079–1088
2021
Later among the works it cites.
C. Bai and P. Wu, “Prrl: Path rotation based knowledge graph representation learning method,” in 2021 IEEE/ACM 8th International Conference on Big Data Computing, Applications and Technologies (BDCAT’21) , 2021, pp. 38–45
2021
Later among the works it cites.
X.-C. Lou and X. Feng, “Multimodal medical image fusion based on multiple latent low-rank representation,” Computational and Mathematical Methods in Medicine , vol. 2021, 2021
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “Bert: Pre-training of deep bidirectional transformers for language understanding,” in Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics , 2019
2019
Cited alongside, same era.
Y. Yu, X. Si, C. Hu, and J. Zhang, “A review of recurrent neural networks: Lstm cells and network architectures,” Neural computation , vol. 31, no. 7, pp. 1235–1270, 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
X. Li, W. Wang, X. Hu, and J. Yang, “Selective kernel networks,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 510–519
2019
Cited alongside, same era.
S. Anwar and N. Barnes, “Real image denoising with feature attention,” in International Conference on Computer Vision (ICCV) , 2019, pp. 3155–3164
2019
Cited alongside, same era.
A. Wu, L. Zhu, Y. Han, and Y. Yang, “Connective cognition network for directional visual commonsense reasoning,” Advances in neural information processing systems (NeurIPS) , vol. 32, 2019
2019
Cited alongside, same era.
Z. Sun, Z.-H. Deng, J.-Y. Nie, and J. Tang, “Rotate: Knowledge graph embedding by relational rotation in complex space,” in International Conference on Learning Representations (ICLR) , 2019
2019
Cited alongside, same era.
B. Li, X. Qi, T. Lukasiewicz, and P. Torr, “Controllable text-to-image generation,” Advances in neural information processing systems (NeurIPS) , vol. 32, 2019
2019
Cited alongside, same era.
X. Zheng, X. Wu, L. Huan, W. He, and H. Zhang, “A gather-to-guide network for remote sensing semantic segmentation of rgb and auxiliary image,” IEEE Transactions on Geoscience and Remote Sensing , vol. 60, pp. 1–15, 2021
2021
Later among the works it cites.
J. Jam, C. Kendrick, V. Drouard, K. Walker, G.-S. Hsu, and M. H. Yap, “R-mnet: A perceptual adversarial network for image inpainting,” in Winter Conference on Applications of Computer Vision (WACV) , 2021, pp. 2714–2723
2021
Later among the works it cites.
G. Wadhwa, A. Dhall, S. Murala, and U. Tariq, “Hyperrealistic image inpainting with hypergraphs,” in Winter Conference on Applications of Computer Vision (WACV) , 2021, pp. 3912–3921
2021
Later among the works it cites.
L. Hoyer, D. Dai, Y. Chen, A. Koring, S. Saha, and L. Van Gool, “Three ways to improve semantic segmentation with self-supervised depth estimation,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2021, pp. 11 130–11 140
2021
Later among the works it cites.
L. Wu, J. Li, Y. Wang, Q. Meng, T. Qin, W. Chen, M. Zhang, T.-Y. Liu et al. , “R-drop: Regularized dropout for neural networks,” Advances in neural information processing systems (NeurIPS) , vol. 34, pp. 10 890–10 905, 2021
2021
Later among the works it cites.
R. Arora, P. Bartlett, P. Mianjy, and N. Srebro, “Dropout: Explicit forms and capacity control,” in International Conference on Machine Learning (ICML) , 2021, pp. 351–361
2021
Later among the works it cites.
2021
Later among the works it cites.
X. Zhou, W. Zhang, H. Xu, and T. Zhang, “Effective sparsification of neural networks with global sparsity constraint,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2021, pp. 3599–3608
2021
Later among the works it cites.
X. Fu, S. Jia, L. Zhuang, M. Xu, J. Zhou, and Q. Li, “Hyperspectral anomaly detection via deep plug-and-play denoising cnn regularization,” IEEE Transactions on Geoscience and Remote Sensing , vol. 59, no. 11, pp. 9553–9568, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
D. So, W. Mańke, H. Liu, Z. Dai, N. Shazeer, and Q. V. Le, “Searching for efficient transformers for language modeling,” in Advances in neural information processing systems (NeurIPS) , vol. 34, 2021, pp. 6010–6022
2021
Later among the works it cites.
H. Zhu, H. Zeng, J. Liu, and X. Zhang, “Logish: A new nonlinear nonmonotonic activation function for convolutional neural network,” Neurocomputing , vol. 458, pp. 490–499, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
Y. Cao, Z. Fang, Y. Wu, D.-X. Zhou, and Q. Gu, “Towards understanding the spectral bias of deep learning,” in International Joint Conferences on Artificial Intelligence (IJCAI) , 2021
2021
Later among the works it cites.
Q. Nguyen, M. Mondelli, and G. F. Montufar, “Tight bounds on the smallest eigenvalue of the neural tangent kernel for deep relu networks,” in International Conference on Machine Learning (ICML) , 2021, pp. 8119–8129
2021
Later among the works it cites.
S. Wang, H. Zhang, K. Xu, X. Lin, S. Jana, C.-J. Hsieh, and J. Z. Kolter, “Beta-crown: Efficient bound propagation with per-neuron split constraints for neural network robustness verification,” in Advances in neural information processing systems (NeurIPS) , 2021
2021
Later among the works it cites.
K. Xu, M. Zhang, J. Li, S. S. Du, K.-I. Kawarabayashi, and S. Jegelka, “How neural networks extrapolate: From feedforward to graph neural networks,” in International Conference on Learning Representations (ICLR) , 2021
2021
Later among the works it cites.
W. Hua, Z. Dai, H. Liu, and Q. Le, “Transformer quality in linear time,” in International Conference on Machine Learning (ICML) , 2022, pp. 9099–9117
2022
Later among the works it cites.
N. Hyeon-Woo, M. Ye-Bin, and T.-H. Oh, “Fedpara: Low-rank hadamard product for communication-efficient federated learning,” in International Conference on Learning Representations (ICLR) , 2022
2022
Later among the works it cites.
Y. Wu, Z. Zhu, F. Liu, G. G. Chrysos, and V. Cevher, “Extrapolation and spectral bias of neural nets with hadamard product: a polynomial net study,” in Advances in neural information processing systems (NeurIPS) , 2022
2022
Later among the works it cites.
K. Han, Y. Wang, H. Chen, X. Chen, J. Guo, Z. Liu, Y. Tang, A. Xiao, C. Xu, Y. Xu et al. , “A survey on vision transformer,” IEEE Transactions on Pattern Analysis and Machine Intelligence (T-PAMI) , 2022
2022
Later among the works it cites.
Y. Peng, S. Dalmia, I. Lane, and S. Watanabe, “Branchformer: Parallel mlp-attention architectures to capture local and global context for speech recognition and understanding,” in International Conference on Machine Learning (ICML) , 2022, pp. 17 627–17 643
2022
Later among the works it cites.
2022
Later among the works it cites.
H. Pan, S. He, K. Zhang, B. Qu, C. Chen, and K. Shi, “Amam: An attention-based multimodal alignment model for medical visual question answering,” Knowledge-Based Systems , vol. 255, p. 109763, 2022
2022
Later among the works it cites.
M. Wang, X. He, L. Liu, L. Qing, H. Chen, Y. Liu, and C. Ren, “Medical visual question answering based on question-type reasoning and semantic space constraint,” Artificial Intelligence in Medicine , vol. 131, p. 102346, 2022
2022
Later among the works it cites.
G. G. Chrysos, M. Georgopoulos, J. Deng, J. Kossaifi, Y. Panagakis, and A. Anandkumar, “Augmenting deep classifiers with polynomial neural networks,” in European Conference on Computer Vision (ECCV) , 2022, pp. 692–716
2022
Later among the works it cites.
D. B. Lindell, D. Van Veen, J. J. Park, and G. Wetzstein, “Bacon: Band-limited coordinate networks for multiscale scene representation,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 16 252–16 262
2022
Later among the works it cites.
2022
Later among the works it cites.
Z.-L. Ni, G.-B. Bian, Z. Li, X.-H. Zhou, R.-Q. Li, and Z.-G. Hou, “Space squeeze reasoning and low-rank bilinear feature fusion for surgical image segmentation,” IEEE Journal of Biomedical and Health Informatics , 2022
2022
Later among the works it cites.
J. Wang, S. Tian, L. Yu, Y. Wang, F. Wang, and Z. Zhou, “Sbdf-net: A versatile dual-branch fusion network for medical image segmentation,” Biomedical Signal Processing and Control , vol. 78, p. 103928, 2022
2022
Later among the works it cites.
Z.-L. Ni, X.-H. Zhou, G.-A. Wang, W.-Q. Yue, Z. Li, G.-B. Bian, and Z.-G. Hou, “Surginet: Pyramid attention aggregation and class-wise self-distillation for surgical instrument segmentation,” Medical Image Analysis , vol. 76, p. 102310, 2022
2022
Later among the works it cites.
X. Wu, Y. Lao, L. Jiang, X. Liu, and H. Zhao, “Point transformer v2: Grouped vector attention and partition-based pooling,” in Advances in neural information processing systems (NeurIPS) , 2022
2022
Later among the works it cites.
S. Ploumpis, S. Moschoglou, V. Triantafyllou, and S. Zafeiriou, “3d human tongue reconstruction from single ”in-the-wild” images,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 2771–2780
2022
Later among the works it cites.
2022
Later among the works it cites.
J. Yang, C. Li, and J. Gao, “Focal modulation networks,” in Advances in neural information processing systems (NeurIPS) , 2022
2022
Later among the works it cites.
H. Pan and J. Huang, “Multimodal high-order relational network for vision-and-language tasks,” Neurocomputing , vol. 492, pp. 62–75, 2022
2022
Later among the works it cites.
Q. CAO, P. KHANNA, N. D. LANE, and A. BALASUBRAMANIAN, “Mobivqa: Efficient on-device visual question answering,” Proc. ACM Interact. Mob. Wearable Ubiquitous Technol. , vol. 6, no. 2, jul 2022
2022
Later among the works it cites.
B. X. Nguyen, T. Do, H. Tran, E. Tjiputra, Q. D. Tran, and A. Nguyen, “Coarse-to-fine reasoning for visual question answering,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 4558–4566
2022
Later among the works it cites.
Y. Xiang, C. Zhang, Z. Han, H. Yu, J. Li, and L. Zhu, “Path-wise attention memory network for visual question answering,” Mathematics , vol. 10, no. 18, p. 3244, 2022
2022
Later among the works it cites.
J. Lu, C. Shan, K. Jin, X. Deng, S. Wang, Y. Wu, J. Li, and Y. Guo, “Onavi: Data-driven based multi-sensor fusion positioning system in indoor environments,” in 2022 IEEE 12th International Conference on Indoor Positioning and Indoor Navigation (IPIN) , 2022, pp. 1–8
2022
Later among the works it cites.
F. Zhan, Y. Yu, R. Wu, J. Zhang, K. Cui, A. Xiao, S. Lu, and C. Miao, “Bi-level feature alignment for versatile image translation and manipulation,” in European Conference on Computer Vision (ECCV) , 2022, pp. 224–241
2022
Later among the works it cites.
W. Liao, K. Hu, M. Y. Yang, and B. Rosenhahn, “Text to image generation with semantic-spatial aware gan,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 18 187–18 196
2022
Later among the works it cites.
F. Wu, L. Liu, F. Hao, F. He, and J. Cheng, “Text-to-image synthesis based on object-guided joint-decoding transformer,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 18 113–18 122
2022
Later among the works it cites.
——, “Language-based image manipulation built on language-guided ranking,” IEEE Transactions on Multimedia , 2022
2022
Later among the works it cites.
Z. Lv, X. Li, Z. Niu, B. Cao, and W. Zuo, “Semantic-shape adaptive feature modulation for semantic image synthesis,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 11 214–11 223
2022
Later among the works it cites.
W. Jiang and H. Hu, “Hadamard product perceptron attention for image captioning,” Neural Processing Letters , pp. 1–18, 2022
2022
Later among the works it cites.
Y. Liang, P. Zhang, Y. Mei, and T. Wang, “Pmacnet: Parallel multiscale attention constraint network for pan-sharpening,” IEEE Geoscience and Remote Sensing Letters , vol. 19, pp. 1–5, 2022
2022
Later among the works it cites.
C. Jing, Y. Jia, Y. Wu, X. Liu, and Q. Wu, “Maintaining reasoning consistency in compositional visual question answering,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 5099–5108
2022
Later among the works it cites.
Y. Li, S. Long, Z. Yang, H. Weng, K. Zeng, Z. Huang, F. L. Wang, and T. Hao, “A bi-level representation learning model for medical visual question answering,” Journal of Biomedical Informatics , vol. 134, p. 104183, 2022
2022
Later among the works it cites.
J. Jin, W. Zhou, L. Ye, J. Lei, L. Yu, X. Qian, and T. Luo, “Dasfnet: Dense-attention–similarity-fusion network for scene classification of dual-modal remote-sensing images,” International Journal of Applied Earth Observation and Geoinformation , vol. 115, p. 103087, 2022
2022
Later among the works it cites.
Y. Feng, H. Xu, J. Jiang, H. Liu, and J. Zheng, “Icif-net: Intra-scale cross-interaction and inter-scale feature fusion network for bitemporal remote sensing images change detection,” IEEE Transactions on Geoscience and Remote Sensing , vol. 60, pp. 1–13, 2022
2022
Later among the works it cites.
H. Liu, D. Tam, M. Muqeeth, J. Mohta, T. Huang, M. Bansal, and C. Raffel, “Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning,” in Advances in neural information processing systems (NeurIPS) , 2022
2022
Later among the works it cites.
J. Bae, M. Kwon, and Y. Uh, “Furrygan: High quality foreground-aware image synthesis,” in European Conference on Computer Vision (ECCV) , 2022, pp. 696–712
2022
Later among the works it cites.
J. Huang, Y. Jin, K. M. Yi, and L. Sigal, “Layered controllable video generation,” in European Conference on Computer Vision (ECCV) , 2022, pp. 546–564
2022
Later among the works it cites.
J. T. Jewell, V. R. Khazaie, and Y. Mohsenzadeh, “One-class learned encoder-decoder network with adversarial context masking for novelty detection,” in Winter Conference on Applications of Computer Vision (WACV) , 2022, pp. 3591–3601
2022
Later among the works it cites.
F. Cong, S. Xu, L. Guo, and Y. Tian, “Anomaly matters: An anomaly-oriented model for medical visual question answering,” IEEE Transactions on Medical Imaging , vol. 41, no. 11, pp. 3385–3397, 2022
2022
Later among the works it cites.
D. Zhao, Y. Zeng, and Y. Li, “Backeisnn: A deep spiking neural network with adaptive self-feedback and balanced excitatory–inhibitory neurons,” Neural Networks , vol. 154, pp. 68–77, 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
S. N. Rai, R. Saluja, C. Arora, V. N. Balasubramanian, A. Subramanian, and C. Jawahar, “Fluid: Few-shot self-supervised image deraining,” in Winter Conference on Applications of Computer Vision (WACV) , 2022, pp. 3077–3086
2022
Later among the works it cites.
W. Chen, Y. Liu, J. Hu, and Y. Yuan, “Dynamic depth-aware network for endoscopy super-resolution,” IEEE Journal of Biomedical and Health Informatics , vol. 26, no. 10, pp. 5189–5200, 2022
2022
Later among the works it cites.
Y. Zhu, X. Wang, L. Chen, and R. Nie, “Cefusion: Multi-modal medical image fusion via cross encoder,” IET Image Processing , 2022
2022
Later among the works it cites.
J. Chen, A. Agarwal, S. Abdelkarim, D. Zhu, and M. Elhoseiny, “Reltransformer: A transformer-based long-tail visual relationship recognition,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 19 507–19 517
2022
Later among the works it cites.
T.-Y. Ji, D. Chu, X.-L. Zhao, and D. Hong, “A unified framework of cloud detection and removal based on low-rank and group sparse regularizations for multitemporal multispectral images,” IEEE Transactions on Geoscience and Remote Sensing , vol. 60, pp. 1–15, 2022
2022
Later among the works it cites.
R. Dai, L. Shen, F. He, X. Tian, and D. Tao, “Dispfl: Towards communication-efficient personalized federated learning via decentralized sparse training,” in International Conference on Machine Learning (ICML) , 2022
2022
Later among the works it cites.
K. Tan, W. Huang, X. Liu, J. Hu, and S. Dong, “A multi-modal fusion framework based on multi-task correlation learning for cancer prognosis prediction,” Artificial Intelligence in Medicine , vol. 126, p. 102260, 2022
2022
Later among the works it cites.
N. Hyeon-Woo, M. Ye-Bin, and T.-H. Oh, “Fedpara: Low-rank hadamard product for communication-efficient federated learning,” in International Conference on Learning Representations (ICLR) , 2022
2022
Later among the works it cites.
Y. Jiang, Y. Zhang, X. Lin, J. Dong, T. Cheng, and J. Liang, “Swinbts: A method for 3d multimodal brain tumor segmentation using swin transformer,” Brain Sciences , vol. 12, no. 6, p. 797, 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
M. Choraria, L. T. Dadi, G. G. Chrysos, J. Mairal, and V. Cevher, “The spectral bias of polynomial neural networks,” in International Conference on Learning Representations (ICLR) , 2022
2022
Later among the works it cites.
Z. Zhu, F. Liu, G. G. Chrysos, and V. Cevher, “Generalization properties of nas under activation and skip connection search,” in Advances in neural information processing systems (NeurIPS) , 2022
2022
Later among the works it cites.
Z. Zhenyu, F. Latorre, G. G. Chrysos, and V. Cevher, “Controlling the complexity and lipschitz constant improves polynomial nets,” in International Conference on Learning Representations (ICLR) , 2022
2022
Later among the works it cites.
E. A. Rocamora, M. F. Sahin, F. Liu, G. G. Chrysos, and V. Cevher, “Sound and complete verification of polynomial networks,” in Advances in neural information processing systems (NeurIPS) , 2022
2022
Later among the works it cites.
F. D. Keles, P. M. Wijewardena, and C. Hegde, “On the computational complexity of self-attention,” in International Conference on Algorithmic Learning Theory (ALT) , 2023
2023
Later among the works it cites.
S. T. Wasim, M. U. Khattak, M. Naseer, S. Khan, M. Shah, and F. S. Khan, “Video-focalnets: Spatio-temporal focal modulation for video action recognition,” in International Conference on Computer Vision (ICCV) , 2023, pp. 13 778–13 789
2023
Later among the works it cites.
D. Fu, S. Arora, J. Grogan, I. Johnson, E. S. Eyuboglu, A. Thomas, B. Spector, M. Poli, A. Rudra, and C. Ré, “Monarch mixer: A simple sub-quadratic gemm-based architecture,” in Advances in neural information processing systems (NeurIPS) , vol. 36, 2023
2023
Later among the works it cites.
O. Bar-Tal, L. Yariv, Y. Lipman, and T. Dekel, “Multidiffusion: Fusing diffusion paths for controlled image generation,” in International Conference on Machine Learning (ICML) , 2023
2023
Later among the works it cites.
Y. Kim, J. Lee, J.-H. Kim, J.-W. Ha, and J.-Y. Zhu, “Dense text-to-image generation with attention modulation,” in International Conference on Computer Vision (ICCV) , 2023, pp. 7701–7711
2023
Later among the works it cites.
T. Cao, K. Kreis, S. Fidler, N. Sharp, and K. Yin, “Texfusion: Synthesizing 3d textures with text-guided image diffusion models,” in International Conference on Computer Vision (ICCV) , 2023, pp. 4169–4181
2023
Later among the works it cites.
F. Babiloni, I. Marras, J. Deng, F. Kokkinos, M. Maggioni, G. Chrysos, P. Torr, and S. Zafeiriou, “Linear complexity self-attention with 3rd order polynomials,” IEEE Transactions on Pattern Analysis and Machine Intelligence (T-PAMI) , vol. 45, no. 11, pp. 12 726–12 737, 2023
2023
Later among the works it cites.
G. G. Chrysos, B. Wang, J. Deng, and V. Cevher, “Regularization of polynomial networks for image recognition,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2023, pp. 16 123–16 132
2023
Later among the works it cites.
X. Xu, J. Sun, Z. Cao, Y. Zhang, X. Zhu, and H. T. Shen, “Tfun: Trilinear fusion network for ternary image-text retrieval,” Information Fusion , vol. 91, pp. 327–337, 2023
2023
Later among the works it cites.
T. Le, N. Le, and B. Le, “Knowledge graph embedding by relational rotation and complex convolution for link prediction,” Expert Systems with Applications , vol. 214, p. 119122, 2023
2023
Later among the works it cites.
A. Nova, H. Dai, and D. Schuurmans, “Gradient-free structured pruning with unlabeled data,” in International Conference on Machine Learning (ICML) . PMLR, 2023, pp. 26 326–26 341
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Y. Chen, G. G. Chrysos, M. Georgopoulos, and V. Cevher, “Multilinear operator networks,” in International Conference on Learning Representations (ICLR) , 2024
2024
Later among the works it cites.
Q. Fan, H. Huang, M. Chen, H. Liu, and R. He, “Rmt: Retentive networks meet vision transformers,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2024, pp. 5641–5651
2024
Later among the works it cites.
B. Zou, C. Yang, Y. Qiao, C. Quan, and Y. Zhao, “Language-aware visual semantic distillation for video question answering,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2024, pp. 27 113–27 123
2024
Later among the works it cites.
L. Iurada, M. Ciccone, and T. Tommasi, “Finding lottery tickets in vision models via data-driven spectral foresight pruning,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2024, pp. 16 142–16 151
2024
Later among the works it cites.
J. J. A. Guerreiro, N. Inoue, K. Masui, M. Otani, and H. Nakayama, “Layoutflow: flow matching for layout generation,” in European Conference on Computer Vision (ECCV) . Springer, 2024, pp. 56–72
2024
Later among the works it cites.
S. Yang, B. Wang, Y. Zhang, Y. Shen, and Y. Kim, “Parallelizing linear transformers with the delta rule over sequence length,” in Advances in neural information processing systems (NeurIPS) , 2024
2024
Later among the works it cites.
A. Gu and T. Dao, “Mamba: Linear-time sequence modeling with selective state spaces,” in First Conference on Language Modeling , 2024
2024
Later among the works it cites.
T. Dao and A. Gu, “Transformers are ssms: Generalized models and efficient algorithms through structured state space duality,” in International Conference on Machine Learning (ICML) , 2024
2024
Later among the works it cites.
Z. Qin, S. Yang, W. Sun, X. Shen, D. Li, W. Sun, and Y. Zhong, “HGRN2: Gated linear RNNs with state expansion,” in First Conference on Language Modeling , 2024
2024
Later among the works it cites.
S. Yang, B. Wang, Y. Shen, R. Panda, and Y. Kim, “Gated linear attention transformers with hardware-efficient training,” in International Conference on Machine Learning (ICML) , 2024
2024
Later among the works it cites.
X. Ma, X. Dai, Y. Bai, Y. Wang, and Y. Fu, “Rewrite the stars,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2024, pp. 5694–5703
2024
Later among the works it cites.
2024
Later among the works it cites.
M. Beck, K. Pöppel, M. Spanring, A. Auer, O. Prudnikova, M. Kopp, G. Klambauer, J. Brandstetter, and S. Hochreiter, “xlstm: Extended long short-term memory,” in Advances in neural information processing systems (NeurIPS) , 2024
2024
Later among the works it cites.
Y. Zhang, S. Yang, R. Zhu, Y. Zhang, L. Cui, Y. Wang, B. Wang, F. Shi, B. Wang, W. Bi et al. , “Gated slot attention for efficient linear-time sequence modeling,” in Advances in neural information processing systems (NeurIPS) , 2024
2024
Later among the works it cites.
Y. Duan, W. Wang, Z. Chen, X. Zhu, L. Lu, T. Lu, Y. Qiao, H. Li, J. Dai, and W. Wang, “Vision-rwkv: Efficient and scalable visual perception with rwkv-like architectures,” in International Conference on Learning Representations (ICLR) , 2025
2025
Closest in time.
Q. Huang, T. Ko, Z. Zhuang, L. Tang, and Y. Zhang, “HiRA: Parameter-efficient hadamard high-rank adaptation for large language models,” in International Conference on Learning Representations (ICLR) , 2025
2025
Closest in time.
S. Yang, J. Kautz, and A. Hatamizadeh, “Gated delta networks: Improving mamba2 with delta rule,” in International Conference on Learning Representations (ICLR) , 2025
2025
Closest in time.