TASK	DATASET	MODEL	METRIC NAME	METRIC VALUE	GLOBAL RANK
Unsupervised Video Object Segmentation	DAVIS 2016 val	AMP	G	87.3	# 4
Unsupervised Video Object Segmentation	DAVIS 2016 val	AMP	J	87.1	# 3
Unsupervised Video Object Segmentation	DAVIS 2016 val	AMP	F	87.5	# 4
Unsupervised Video Object Segmentation	YouTube-Objects	AMP	J	75.0	# 1

Badge	Markdown
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/adaptive-multi-source-predictor-for-zero-shot/unsupervised-video-object-segmentation-on-12)](https://paperswithcode.com/sota/unsupervised-video-object-segmentation-on-12?p=adaptive-multi-source-predictor-for-zero-shot)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/adaptive-multi-source-predictor-for-zero-shot/unsupervised-video-object-segmentation-on-10)](https://paperswithcode.com/sota/unsupervised-video-object-segmentation-on-10?p=adaptive-multi-source-predictor-for-zero-shot)`

Adaptive Multi-source Predictor for Zero-shot Video Object Segmentation

18 Mar 2023 · Xiaoqi Zhao, Shijie Chang, Youwei Pang, Jiaxing Yang, Lihe Zhang, Huchuan Lu ·

Static and moving objects often occur in real-life videos. Most video object segmentation methods only focus on extracting and exploiting motion cues to perceive moving objects. Once faced with the frames of static objects, the moving object predictors may predict failed results caused by uncertain motion information, such as low-quality optical flow maps. Besides, different sources such as RGB, depth, optical flow and static saliency can provide useful information about the objects. However, existing approaches only consider either the RGB or RGB and optical flow. In this paper, we propose a novel adaptive multi-source predictor for zero-shot video object segmentation (ZVOS). In the static object predictor, the RGB source is converted to depth and static saliency sources, simultaneously. In the moving object predictor, we propose the multi-source fusion structure. First, the spatial importance of each source is highlighted with the help of the interoceptive spatial attention module (ISAM). Second, the motion-enhanced module (MEM) is designed to generate pure foreground motion attention for improving the representation of static and moving features in the decoder. Furthermore, we design a feature purification module (FPM) to filter the inter-source incompatible features. By using the ISAM, MEM and FPM, the multi-source features are effectively fused. In addition, we put forward an adaptive predictor fusion network (APF) to evaluate the quality of the optical flow map and fuse the predictions from the static object predictor and the moving object predictor in order to prevent over-reliance on the failed results caused by low-quality optical flow maps. Experiments show that the proposed model outperforms the state-of-the-art methods on three challenging ZVOS benchmarks. And, the static object predictor precisely predicts a high-quality depth map and static saliency map at the same time.

PDF Abstract

Code

Add Remove Mark official

xiaoqi-zhao-dlut/multi-source-aps-z… official

Tasks

Add Remove

Object

Optical Flow Estimation

Semantic Segmentation

Unsupervised Video Object Segmentation

Video Object Segmentation

Video Semantic Segmentation

Zero-Shot Video Object Segmentation

Datasets

DAVIS

DAVIS 2016

FBMS

NLPR

SIP

Results from the Paper

Edit

Ranked #1 on Unsupervised Video Object Segmentation on YouTube-Objects

Get a GitHub badge

Task	Dataset	Model	Metric Name	Metric Value	Global Rank	Benchmark
Unsupervised Video Object Segmentation	DAVIS 2016 val	AMP	G	87.3	# 4	Compare
			J	87.1	# 3	Compare
			F	87.5	# 4	Compare
Unsupervised Video Object Segmentation	YouTube-Objects	AMP	J	75.0	# 1	Compare

Methods

Add Remove

Average Pooling • Convolution • Max Pooling • Sigmoid Activation • Spatial Attention Module

Edit Social Preview

Adaptive Multi-source Predictor for Zero-shot Video Object Segmentation

Code Edit Add Remove Mark official

Tasks Edit Add Remove

Datasets Edit

Results from the Paper Edit

Methods Edit Add Remove

Code

Add Remove Mark official

Tasks

Add Remove

Datasets

Results from the Paper

Edit

Methods

Add Remove