Topology Dictionary with Markov Model for 3D Video Content-Based Skimming and Description

Tung, Tony; Matsuyama, Takashi

doi:10.1109/CVPR.2009.5206823

Topology Dictionary with Markov Model for 3D Video Content-Based Skimming and Description

Tony Tung, Takashi Matsuyama

CVPR 2009 pp. 469-476

doi:10.1109/CVPR.2009.5206823 /cvpr/2009/tung2009cvpr-topology/

Abstract

This paper presents a novel approach to skim and describe 3D videos. 3D video is an imaging technology which consists in a stream of 3D models in motion captured by a synchronized set of video cameras. Each frame is composed of one or several 3D models, and therefore the acquisition of long sequences at video rate requires massive storage devices. In order to reduce the storage cost while keeping relevant information, we propose to encode 3D video sequences using a topology-based shape descriptor dictionary. This dictionary is either generated from a set of extracted patterns or learned from training input sequences with semantic annotations. It relies on an unsupervised 3D shape-based clustering of the dataset by Reeb graphs, and features a Markov network to characterize topological changes. The approach allows content-based compression and skimming with accurate recovery of sequences and can handle complex topological changes. Redundancies are detected and skipped based on a probabilistic discrimination process. Semantic description of video sequences is then automatically performed. In addition, forthcoming frame encoding is achieved using a multiresolution matching scheme and allows action recognition in 3D. Our experiments were performed on complex 3D video sequences. We demonstrate the robustness and accuracy of the 3D video skimming with dramatic low bitrate coding and high compression ratio.

PDF CVPR Semantic Scholar

Cite

Text

Tung and Matsuyama. "Topology Dictionary with Markov Model for 3D Video Content-Based Skimming and Description." IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2009. doi:10.1109/CVPR.2009.5206823

Markdown

[Tung and Matsuyama. "Topology Dictionary with Markov Model for 3D Video Content-Based Skimming and Description." IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2009.](https://mlanthology.org/cvpr/2009/tung2009cvpr-topology/) doi:10.1109/CVPR.2009.5206823

BibTeX

@inproceedings{tung2009cvpr-topology,
  title     = {{Topology Dictionary with Markov Model for 3D Video Content-Based Skimming and Description}},
  author    = {Tung, Tony and Matsuyama, Takashi},
  booktitle = {IEEE/CVF Conference on Computer Vision and Pattern Recognition},
  year      = {2009},
  pages     = {469-476},
  doi       = {10.1109/CVPR.2009.5206823},
  url       = {https://mlanthology.org/cvpr/2009/tung2009cvpr-topology/}
}