TASK	DATASET	MODEL	METRIC NAME	METRIC VALUE	GLOBAL RANK
Video Object Segmentation	DAVIS 2016	AOC-MF (val)	Jaccard (Mean)	88.5	# 15
Video Object Segmentation	DAVIS 2016	AOC-MF (val)	F-Score	94.7	# 1
Video Object Segmentation	DAVIS 2017	AOC-MF (val)	Jaccard (Mean)	81.7	# 1
Video Object Segmentation	DAVIS 2017	AOC-MF (val)	F-Score	85.9	# 1
Visual Object Tracking	YouTube-VOS	AOC-MF	O (Average of Measures)	84	# 1
Visual Object Tracking	YouTube-VOS	AOC-MF	Jaccard (Seen)	82.7	# 1
Visual Object Tracking	YouTube-VOS	AOC-MF	Jaccard (Unseen)	78.8	# 1
Visual Object Tracking	YouTube-VOS	AOC-MF	F-Measure (Seen)	87.4	# 1
Visual Object Tracking	YouTube-VOS	AOC-MF	F-Measure (Unseen)	87.1	# 1
Visual Object Tracking	YouTube-VOS	AOC-Base	O (Average of Measures)	83.6	# 2
Visual Object Tracking	YouTube-VOS	AOC-Base	Jaccard (Seen)	82.6	# 2
Visual Object Tracking	YouTube-VOS	AOC-Base	Jaccard (Unseen)	78.3	# 2
Visual Object Tracking	YouTube-VOS	AOC-Base	F-Measure (Seen)	87.2	# 2
Visual Object Tracking	YouTube-VOS	AOC-Base	F-Measure (Unseen)	86.3	# 2

Badge	Markdown
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/towards-robust-video-object-segmentation-with/video-object-segmentation-on-davis-2016)](https://paperswithcode.com/sota/video-object-segmentation-on-davis-2016?p=towards-robust-video-object-segmentation-with)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/towards-robust-video-object-segmentation-with/video-object-segmentation-on-davis-2017)](https://paperswithcode.com/sota/video-object-segmentation-on-davis-2017?p=towards-robust-video-object-segmentation-with)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/towards-robust-video-object-segmentation-with/visual-object-tracking-on-youtube-vos-1)](https://paperswithcode.com/sota/visual-object-tracking-on-youtube-vos-1?p=towards-robust-video-object-segmentation-with)`

Towards Robust Video Object Segmentation with Adaptive Object Calibration

2 Jul 2022 · Xiaohao Xu, Jinglu Wang, Xiang Ming, Yan Lu ·

In the booming video era, video segmentation attracts increasing research attention in the multimedia community. Semi-supervised video object segmentation (VOS) aims at segmenting objects in all target frames of a video, given annotated object masks of reference frames. Most existing methods build pixel-wise reference-target correlations and then perform pixel-wise tracking to obtain target masks. Due to neglecting object-level cues, pixel-level approaches make the tracking vulnerable to perturbations, and even indiscriminate among similar objects. Towards robust VOS, the key insight is to calibrate the representation and mask of each specific object to be expressive and discriminative. Accordingly, we propose a new deep network, which can adaptively construct object representations and calibrate object masks to achieve stronger robustness. First, we construct the object representations by applying an adaptive object proxy (AOP) aggregation method, where the proxies represent arbitrary-shaped segments at multi-levels for reference. Then, prototype masks are initially generated from the reference-target correlations based on AOP. Afterwards, such proto-masks are further calibrated through network modulation, conditioning on the object proxy representations. We consolidate this conditional mask calibration process in a progressive manner, where the object representations and proto-masks evolve to be discriminative iteratively. Extensive experiments are conducted on the standard VOS benchmarks, YouTube-VOS-18/19 and DAVIS-17. Our model achieves the state-of-the-art performance among existing published works, and also exhibits superior robustness against perturbations. Our project repo is at https://github.com/JerryX1110/Robust-Video-Object-Segmentation

PDF Abstract

Code

Add Remove Mark official

jerryx1110/robust-video-object-segm… official

Tasks

Add Remove

Object

Segmentation

Semantic Segmentation

Semi-Supervised Video Object Segmentation

Video Object Segmentation

Video Segmentation

Video Semantic Segmentation

Visual Object Tracking

Datasets

DAVIS

DAVIS 2017

DAVIS 2016

YouTube-VOS 2018

Results from the Paper

Edit

Ranked #1 on Visual Object Tracking on YouTube-VOS

Get a GitHub badge

Task	Dataset	Model	Metric Name	Metric Value	Global Rank	Benchmark
Video Object Segmentation	DAVIS 2016	AOC-MF (val)	Jaccard (Mean)	88.5	# 15	Compare
Video Object Segmentation	DAVIS 2016	AOC-MF (val)	F-Score	94.7	# 1	Compare
Video Object Segmentation	DAVIS 2017	AOC-MF (val)	Jaccard (Mean)	81.7	# 1	Compare
Video Object Segmentation	DAVIS 2017	AOC-MF (val)	F-Score	85.9	# 1	Compare
Visual Object Tracking	YouTube-VOS	AOC-MF	O (Average of Measures)	84	# 1	Compare
			Jaccard (Seen)	82.7	# 1	Compare
			Jaccard (Unseen)	78.8	# 1	Compare
			F-Measure (Seen)	87.4	# 1	Compare
			F-Measure (Unseen)	87.1	# 1	Compare
Visual Object Tracking	YouTube-VOS	AOC-Base	O (Average of Measures)	83.6	# 2	Compare
			Jaccard (Seen)	82.6	# 2	Compare
			Jaccard (Unseen)	78.3	# 2	Compare
			F-Measure (Seen)	87.2	# 2	Compare
			F-Measure (Unseen)	86.3	# 2	Compare

Methods

Add Remove

VOS

Edit Social Preview

Towards Robust Video Object Segmentation with Adaptive Object Calibration

Code Edit Add Remove Mark official

Tasks Edit Add Remove

Datasets Edit

Results from the Paper Edit

Methods Edit Add Remove

Code

Add Remove Mark official

Tasks

Add Remove

Datasets

Results from the Paper

Edit

Methods

Add Remove