PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 27, 2024Sensors1 citationsOpen Access

Real-Time Multi-Person Video Synthesis with Controllable Prior-Guided Matting

View Full Paper
ACAoran ChenHHHai HuangYZYueyan Zhu

Key Points

Key points are not available for this paper at this time.

Abstract

In order to enhance the matting performance in multi-person dynamic scenarios, we introduce a robust, real-time, high-resolution, and controllable human video matting method that achieves state of the art on all metrics. Unlike most existing methods that perform video matting frame by frame as independent images, we design a unified architecture using a controllable generation model to solve the problem of the lack of overall semantic information in multi-person video. Our method, called ControlMatting, uses an independent recurrent architecture to exploit temporal information in videos and achieves significant improvements in temporal coherence and detailed matting quality. ControlMatting adopts a mixed training strategy comprised of matting and a semantic segmentation dataset, which effectively improves the semantic understanding ability of the model. Furthermore, we propose a novel deep learning-based image filter algorithm that enforces our detailed augmentation ability on both matting and segmentation objectives. Our experiments have proved that prior information about the human body from the image itself can effectively combat the defect masking problem caused by complex dynamic scenarios with multiple people.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Chen et al. (2024) studied this question.

synapsesocial.com/papers/68e6d431b6db643587651fffhttps://doi.org/10.3390/s24092795
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Multi-Temporal Ultra Dense Memory Network for Video Super-Resolution2019 · 141 citations
  2. 2Transformer-based Cross Reference Network for video salient object detection2022 · 40 citations
  3. 3DiffusionDet: Diffusion Model for Object Detection2023 · 575 citations
  4. 4Action Recognition With Spatio–Temporal Visual Attention on Skeleton Image Sequences2018 · 177 citations
  5. 5Progressively real-time video salient object detection via cascaded fully convolutional networks with motion attention2021 · 28 citations