OVSNet : Towards One-Pass Real-Time Video Object Segmentation

Sun, Peng; Lin, Peiwen; Cheng, Guangliang; Shi, Jianping; Zhang, Jiawan; Li, Xi

Computer Science > Computer Vision and Pattern Recognition

arXiv:1905.10064 (cs)

[Submitted on 24 May 2019 (v1), last revised 2 Jul 2019 (this version, v2)]

Title:OVSNet : Towards One-Pass Real-Time Video Object Segmentation

Authors:Peng Sun, Peiwen Lin, Guangliang Cheng, Jianping Shi, Jiawan Zhang, Xi Li

View PDF

Abstract:Video object segmentation aims at accurately segmenting the target object regions across consecutive frames. It is technically challenging for coping with complicated factors (e.g., shape deformations, occlusion and out of the lens). Recent approaches have largely solved them by using backforth re-identification and bi-directional mask propagation. However, their methods are extremely slow and only support offline inference, which in principle cannot be applied in real time. Motivated by this observation, we propose a efficient detection-based paradigm for video object segmentation. We propose an unified One-Pass Video Segmentation framework (OVS-Net) for modeling spatial-temporal representation in a unified pipeline, which seamlessly integrates object detection, object segmentation, and object re-identification. The proposed framework lends itself to one-pass inference that effectively and efficiently performs video object segmentation. Moreover, we propose a maskguided attention module for modeling the multi-scale object boundary and multi-level feature fusion. Experiments on the challenging DAVIS 2017 demonstrate the effectiveness of the proposed framework with comparable performance to the state-of-the-art, and the great efficiency about 11.5 FPS towards pioneering real-time work to our knowledge, more than 5 times faster than other state-of-the-art methods.

Comments:	10 pages, 6 figures
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1905.10064 [cs.CV]
	(or arXiv:1905.10064v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1905.10064

Submission history

From: Peng Sun [view email]
[v1] Fri, 24 May 2019 07:12:48 UTC (7,809 KB)
[v2] Tue, 2 Jul 2019 11:37:33 UTC (7,810 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:OVSNet : Towards One-Pass Real-Time Video Object Segmentation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:OVSNet : Towards One-Pass Real-Time Video Object Segmentation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators