Fetching the paper…
Reading the bibliography…
We propose novel Stacked Spatio-Temporal Graph Convolutional Networks (Stacked-STGCN) for action segmentation, i.e., predicting and localizing a sequence of actions over long videos.
Graph neural networks for ranking web pages
F. Scarselli, S. L. Yong, M. Gori, M. Hagenbuchner, A. C. Tsoi, and M. Maggini · 2005
Earlier work this paper cites.
The graph neural network model
F. Scarselli, M. Gori, A. C. Tsoi, M. Hagenbuchner, and G. Monfardini · 2009
Earlier work this paper cites.
Wavelets on graphs via spectral graph theory
D. K. Hammond, P. Vandergheynst, and R. Gribonval · 2011
Earlier work this paper cites.
Learning human activities and object affordances from rgb-d videos
H. S. Koppula, R. Gupta, and A. Saxena · 2013
Earlier work this paper cites.
Parsing videos of actions with segmental grammars
H. Pirsiavash and D. Ramanan · 2014
Earlier work this paper cites.
Two-stream convolutional networks for action recognition in videos
K. Simonyan and A. Zisserman · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Earlier work this paper cites.
Convolutional networks on graphs for learning molecular fingerprints
D. K. Duvenaud, D. Maclaurin, J. Iparraguirre, R. Bombarell, T. Hirzel, A. Aspuru-Guzik, and R. P. Adams · 2015
Earlier work this paper cites.
Gated graph sequence neural networks
Y. Li, D. Tarlow, M. Brockschmidt, and R. Zemel · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Earlier work this paper cites.
Diffusion-convolutional neural networks
J. Atwood and D. Towsley · 2016
Earlier work this paper cites.
Convolutional neural networks on graphs with fast localized spectral filtering
M. Defferrard, X. Bresson, and P. Vandergheynst · 2016
Earlier work this paper cites.
Structural-rnn: Deep learning on spatio-temporal graphs
A. Jain, A. R. Zamir, S. Savarese, and A. Saxena · 2016
Earlier work this paper cites.
Semi-supervised classification with graph convolutional networks
T. N. Kipf and M. Welling · 2016
Earlier work this paper cites.
Anticipating human activities using object affordances for reactive robotic response
H. S. Koppula and A. Saxena · 2016
Earlier work this paper cites.
Stacked hourglass networks for human pose estimation
A. Newell, K. Yang, and J. Deng · 2016
Earlier work this paper cites.
Ntu rgb+ d: A large scale dataset for 3d human activity analysis
A. Shahroudy, J. Liu, T.-T. Ng, and G. Wang · 2016
Earlier work this paper cites.
Hollywood in homes: Crowdsourcing data collection for activity understanding
G. A. Sigurdsson, G. Varol, X. Wang, A. Farhadi, I. Laptev, and A. Gupta · 2016
Cited alongside, same era.
A multi-stream bi-directional recurrent neural network for fine-grained action detection
B. Singh, T. K. Marks, M. Jones, O. Tuzel, and M. Shao · 2016
Cited alongside, same era.
Temporal segment networks: Towards good practices for deep action recognition
L. Wang, Y. Xiong, Z. Wang, Y. Qiao, D. Lin, X. Tang, and L. Van Gool · 2016
Cited alongside, same era.
Situation recognition: Visual semantic role labeling for image understanding
M. Yatskar, L. Zettlemoyer, and A. Farhadi · 2016
Cited alongside, same era.
Community detection with graph neural networks
J. Bruna and X. Li · 2017
Cited alongside, same era.
The description length of deep learning models
L. Blier and Y. Ollivier · 2018
Closest in time.
Weakly-supervised action segmentation with iterative soft boundary assignment
L. Ding and C. Xu · 2018
Closest in time.
Neural graph matching networks for fewshot 3d action recognition
M. Guo, E. Chou, D.-A. Huang, S. Song, S. Yeung, and L. Fei-Fei · 2018
Closest in time.
Temporal deformable residual networks for action segmentation in videos
P. Lei and S. Todorovic · 2018
Closest in time.
Factorizable net: an efficient subgraph-based framework for scene graph generation
Y. Li, W. Ouyang, B. Zhou, J. Shi, C. Zhang, and X. Wang · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Quo vadis, action recognition? a new model and the kinetics dataset
J. Carreira and A. Zisserman · 2017
Cited alongside, same era.
Predictivecorrective networks for action detection
A. Dave, O. Russakovsky, and D. Ramanan · 2017
Cited alongside, same era.
Temporal 3d convnets: New architecture and transfer learning for video classification
A. Diba, M. Fayyaz, V. Sharma, A. H. Karami, M. M. Arzani, R. Yousefzadeh, and L. Van Gool · 2017
Cited alongside, same era.
Tricornet: A hybrid temporal convolutional and recurrent network for video action segmentation
L. Ding and C. Xu · 2017
Cited alongside, same era.
The kinetics human action video dataset
W. Kay, J. Carreira, K. Simonyan, B. Zhang, C. Hillier, S. Vijayanarasimhan, F. Viola, T. Green, T. Back, P. Natsev, et al · 2017
Cited alongside, same era.
Situation recognition with graph neural networks
R. Li, M. Tapaswi, R. Liao, J. Jia, R. Urtasun, and S. Fidler · 2017
Cited alongside, same era.
Learning spatio-temporal representation with pseudo-3d residual networks
Z. Qiu, T. Yao, and T. Mei · 2017
Cited alongside, same era.
E. Mavroudi, D. Bhaskara, S. Sefati, H. Ali, and R. Vidal · 2018
Closest in time.
Learning latent super-events to detect multiple activities in videos
A. Piergiovanni and M. S. Ryoo · 2018
Closest in time.
Modeling relational data with graph convolutional networks
M. Schlichtkrull, T. N. Kipf, P. Bloem, R. van den Berg, I. Titov, and M. Welling · 2018
Closest in time.
Actor and observer: Joint modeling of first and third-person videos
G. A. Sigurdsson, A. Gupta, C. Schmid, A. Farhadi, and K. Alahari · 2018
Closest in time.
Rgcnn: Regularized graph cnn for point cloud segmentation
G. Te, W. Hu, Z. Guo, and A. Zheng · 2018
Closest in time.
A closer look at spatiotemporal convolutions for action recognition
D. Tran, H. Wang, L. Torresani, J. Ray, Y. LeCun, and M. Paluri · 2018
Closest in time.
Videos as space-time region graphs
X. Wang and A. Gupta · 2018
Closest in time.
S3d: Stacking segmental p3d for action quality assessment
X. Xiang, Y. Tian, A. Reiter, G. D. Hager, and T. D. Tran · 2018
Closest in time.
Spatial temporal graph convolutional networks for skeleton-based action recognition
S. Yan, Y. Xiong, and D. Lin · 2018
Closest in time.
Every moment counts: Dense detailed labeling of actions in complex videos
S. Yeung, O. Russakovsky, N. Jin, M. Andriluka, G. Mori, and L. Fei-Fei · 2018
Closest in time.
Graph edge convolutional neural networks for skeleton based action recognition
X. Zhang, C. Xu, and D. Tao · 2018
Closest in time.