PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
July 1, 201749 citations

Budget-Aware Deep Semantic Video Segmentation

View Full Paper
BMBehrooz MahasseniOregon State UniversitySTSiniša TodorovićOregon State UniversityAFAlan FernOregon State University

Key Points

Key points are not available for this paper at this time.

Abstract

In this work, we study a poorly understood trade-off between accuracy and runtime costs for deep semantic video segmentation. While recent work has demonstrated advantages of learning to speed-up deep activity detection, it is not clear if similar advantages will hold for our very different segmentation loss function, which is defined over individual pixels across the frames. In deep video segmentation, the most time consuming step represents the application of a CNN to every frame for assigning class labels to every pixel, typically taking 6-9 times of the video footage. This motivates our new budget-aware framework that learns to optimally select a small subset of frames for pixelwise labeling by a CNN, and then efficiently interpolates the obtained segmentations to yet unprocessed frames. This interpolation may use either a simple optical-flow guided mapping of pixel labels, or another significantly less complex and thus faster CNN. We formalize the frame selection as a Markov Decision Process, and specify a Long Short-Term Memory (LSTM) network to model a policy for selecting the frames. For training the LSTM, we develop a policy-gradient reinforcement-learning approach for approximating the gradient of our non-decomposable and non-differentiable objective. Evaluation on two benchmark video datasets show that our new framework is able to significantly reduce computation time, and maintain competitive video segmentation accuracy under varying budgets.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Mahasseni et al. (2017) studied this question.

synapsesocial.com/papers/6a197fdab1a1e919c389020dhttps://doi.org/10.1109/cvpr.2017.224
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Dynamic Structured Model Selection2013 · 17 citations
  2. 2How to Discount Deep Reinforcement Learning: Towards New Dynamic Strategies2015 · 80 citations
  3. 3Are we ready for autonomous driving? The KITTI vision benchmark suite2012 · 14,803 citations
  4. 4Bayesian SegNet: Model Uncertainty in Deep Convolutional Encoder-Decoder Architectures for Scene Understanding2015 · 301 citations
  5. 5Efficient Inference in Fully Connected CRFs with Gaussian Edge Potentials2012 · 2,977 citations