Deep Graph Pose: a semi-supervised deep graphical model for improved animal pose tracking
Wu, A.; Buchanan, E. K.; Whiteway, M. R.; Schartner, M.; Meijer, G. T.; Noel, J.-P.; Rodriguez, E.; Everett, C.; Norovich, A.; Schaffer, E. S.; Mishra, N.; Salzman, C. D.; Angelaki, D. E.; International Brain Laboratory, T.; Cunningham, J. P.; Paninski, L.
Show abstract
Noninvasive behavioral tracking of animals is crucial for many scientific investigations. Recent transfer learning approaches for behavioral tracking have considerably advanced the state of the art. Typically these methods treat each video frame and each object to be tracked independently. In this work, we improve on these methods (particularly in the regime of few training labels) by leveraging the rich spatiotemporal structures pervasive in behavioral video -- specifically, the spatial statistics imposed by physical constraints (e.g., paw to elbow distance), and the temporal statistics imposed by smoothness from frame to frame. We propose a probabilistic graphical model built on top of deep neural networks, Deep Graph Pose (DGP), to leverage these useful spatial and temporal constraints, and develop an efficient structured variational approach to perform inference in this model. The resulting semi-supervised model exploits both labeled and unlabeled frames to achieve significantly more accurate and robust tracking while requiring users to label fewer training frames. In turn, these tracking improvements enhance performance on downstream applications, including robust unsupervised segmentation of behavioral "syllables," and estimation of interpretable "disentangled" low-dimensional representations of the full behavioral video. Open source code is available at https://github.com/paninski-lab/deepgraphpose.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Automated construction of cognitive maps with predictive coding 94%
- Generalized Radiograph Representation Learning via Cross-supervision between Images and Free-text Radiology Reports 93%
- UDCT: Unsupervised data to content transformation with histogram-matching cycle-consistent generative adversarial networks 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.