REVIEW 4 cited by
Spatial Temporal Graph Convolutional Networks for Skeleton-Based Action Recognition
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Dynamics of human body skeletons convey significant information for human action recognition. Conventional approaches for modeling skeletons usually rely on hand-crafted parts or traversal rules, thus resulting in limited expressive power and difficulties of generalization. In this work, we propose a novel model of dynamic skeletons called Spatial-Temporal Graph Convolutional Networks (ST-GCN), which moves beyond the limitations of previous methods by automatically learning both the spatial and temporal patterns from data. This formulation not only leads to greater expressive power but also stronger generalization capability. On two large datasets, Kinetics and NTU-RGBD, it achieves substantial improvements over mainstream methods.
Forward citations
Cited by 4 Pith papers
-
Physical Self-Supervised Learning: IMU Sensing without Manual Labels
A physics-based self-supervised autoencoder achieves label-free IMU tracking and motion capture that outperforms supervised baselines in generalization tests.
-
Enhancing Sports Strategy with Video Analytics and Data Mining: Automated Video-Based Analytics Framework for Tennis Doubles
A tennis doubles annotation framework is built and evaluated, showing transfer-learned CNNs outperform pose-only GCNs for automated shot and formation labeling.
-
CascadeFormer: A Family of Two-stage Cascading Transformers for Skeleton-based Human Action Recognition
A masked-pretrained skeleton transformer with a second fine-tuning transformer and cross-attention fusion reaches 94.66% on Penn Action, 91.16% on N-UCLA, and 81.01%/88.17% on NTU RGB+D 60 cross-subject/cross-view.
-
Variational Graph Convolutional Neural Networks
Variational graph convolutional networks that sample layer outputs from learned Gaussians provide uncertainty estimates and small accuracy improvements on social trading and skeleton action recognition benchmarks.
Discussion (0). Sign in to comment.