Interactive Video Object Mask Annotation

Abstract

In this paper, we introduce a practical system for interactive video object mask annotation, which can support multiple back-end methods. To demonstrate the generalization of our system, we introduce a novel approach for video object annotation. Our proposed system takes scribbles at a chosen key-frame from the end-users via a user-friendly interface and produces masks of corresponding objects at the key-frame via the Control-Point-based Scribbles-to-Mask (CPSM) module. The object masks at the key-frame are then propagated to other frames and refined through the Multi-Referenced Guided Segmentation (MRGS) module. Last but not least, the user can correct wrong segmentation at some frames, and the corrected mask is continuously propagated to other frames in the video via the MRGS to produce the object masks at all video frames.

Cite

Text

Le et al. "Interactive Video Object Mask Annotation." AAAI Conference on Artificial Intelligence, 2021. doi:10.1609/AAAI.V35I18.18014

Markdown

[Le et al. "Interactive Video Object Mask Annotation." AAAI Conference on Artificial Intelligence, 2021.](https://mlanthology.org/aaai/2021/le2021aaai-interactive/) doi:10.1609/AAAI.V35I18.18014

BibTeX

@inproceedings{le2021aaai-interactive,
  title     = {{Interactive Video Object Mask Annotation}},
  author    = {Le, Trung-Nghia and Nguyen, Tam V. and Tran, Quoc-Cuong and Nguyen, Lam and Hoang, Trung-Hieu and Le, Minh-Quan and Tran, Minh-Triet},
  booktitle = {AAAI Conference on Artificial Intelligence},
  year      = {2021},
  pages     = {16067-16070},
  doi       = {10.1609/AAAI.V35I18.18014},
  url       = {https://mlanthology.org/aaai/2021/le2021aaai-interactive/}
}