Saturated transformers are constant-depth threshold circuits
William Merrill, Ashish Sabharwal, and Noah A Smith · 2022
Later among the works it cites.
A theoretical comparison of graph neural network extensions
Pál András Papp and Roger Wattenhofer · 2022
Later among the works it cites.
Ordered subgraph aggregation networks
Chendi Qian, Gaurav Rattan, Floris Geerts, Mathias Niepert, and Christopher Morris · 2022
Later among the works it cites.
A new perspective on ”how graph neural networks go beyond weisfeiler-lehman?”
Asiri Wijesinghe and Qing Wang · 2022
Later among the works it cites.
Molecular contrastive learning of representations via graph neural networks
Yuyang Wang, Jianren Wang, Zhonglin Cao, and Amir Barati Farimani · 2022
Later among the works it cites.
The expressive power of pooling in graph neural networks
Filippo Maria Bianchi and Veronica Lachi · 2023
Later among the works it cites.
Mamba: Linear-time sequence modeling with selective state spaces
Original
Albert Gu and Tri Dao · 2023
Later among the works it cites.
Transformers learn shortcuts to automata
Bingbin Liu, Jordan T. Ash, Surbhi Goel, Akshay Krishnamurthy, and Cyril Zhang · 2023
Later among the works it cites.
The parallelism tradeoff: Limitations of log-precision transformers
William Merrill and Ashish Sabharwal · 2023
Later among the works it cites.
A complete expressiveness hierarchy for subgraph gnns via subgraph weisfeiler-lehman tests
Bohang Zhang, Guhao Feng, Yiheng Du, Di He, and Liwei Wang · 2023
Later among the works it cites.
Transformers in uniform 𝖳𝖢 0 \mathsf{TC}^{0}
Original
David Chiang · 2024
Later among the works it cites.
Circuit complexity bounds for rope-based transformer architecture
Original
Bo Chen, Xiaoyu Li, Yingyu Liang, Jiangxuan Long, Zhenmei Shi, and Zhao Song · 2024
Later among the works it cites.
The computational limits of state-space models and mamba via the lens of circuit complexity
Original
Yifang Chen, Xiaoyu Li, Yingyu Liang, Zhenmei Shi, and Zhao Song · 2024
Later among the works it cites.
Rethinking the expressiveness of gnns: A computational model perspective
Original
Guanyu Cui, Zhewei Wei, and Hsin-Hao Su · 2024
Later among the works it cites.
On the expressive power of modern hopfield networks
Original
Xiaoyu Li, Yuanpeng Li, Yingyu Liang, Zhenmei Shi, and Zhao Song · 2024
Later among the works it cites.
Theoretical constraints on the expressive power of rope-based tensor attention transformers
Original
Xiaoyu Li, Yingyu Liang, Zhenmei Shi, Zhao Song, and Mingda Wan · 2024
Later among the works it cites.
A logic for expressing log-precision transformers
William Merrill and Ashish Sabharwal · 2024
Later among the works it cites.
Roformer: Enhanced transformer with rotary position embedding
Jianlin Su, Murtadha Ahmed, Yu Lu, Shengfeng Pan, Wen Bo, and Yunfeng Liu · 2024
Later among the works it cites.
Beyond weisfeiler-lehman: A quantitative framework for GNN expressiveness
Bohang Zhang, Jingchu Gai, Yiheng Du, Qiwei Ye, Di He, and Liwei Wang · 2024
Later among the works it cites.