TASK	DATASET	MODEL	METRIC NAME	METRIC VALUE	GLOBAL RANK
Monocular Depth Estimation	KITTI Eigen split unsupervised	X-Distill (M+1024x320)	absolute relative error	0.102	# 17
Monocular Depth Estimation	KITTI Eigen split unsupervised	X-Distill (M+1024x320)	RMSE	4.439	# 14
Monocular Depth Estimation	KITTI Eigen split unsupervised	X-Distill (M+1024x320)	Sq Rel	0.698	# 12
Monocular Depth Estimation	KITTI Eigen split unsupervised	X-Distill (M+1024x320)	RMSE log	0.180	# 13
Monocular Depth Estimation	KITTI Eigen split unsupervised	X-Distill (M+1024x320)	Delta < 1.25	0.895	# 12
Monocular Depth Estimation	KITTI Eigen split unsupervised	X-Distill (M+1024x320)	Delta < 1.25^2	0.965	# 10
Monocular Depth Estimation	KITTI Eigen split unsupervised	X-Distill (M+1024x320)	Delta < 1.25^3	0.983	# 10

Badge	Markdown
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/x-distill-improving-self-supervised-monocular/monocular-depth-estimation-on-kitti-eigen-1)](https://paperswithcode.com/sota/monocular-depth-estimation-on-kitti-eigen-1?p=x-distill-improving-self-supervised-monocular)`

X-Distill: Improving Self-Supervised Monocular Depth via Cross-Task Distillation

24 Oct 2021 · Hong Cai, Janarbek Matai, Shubhankar Borse, Yizhe Zhang, Amin Ansari, Fatih Porikli ·

In this paper, we propose a novel method, X-Distill, to improve the self-supervised training of monocular depth via cross-task knowledge distillation from semantic segmentation to depth estimation. More specifically, during training, we utilize a pretrained semantic segmentation teacher network and transfer its semantic knowledge to the depth network. In order to enable such knowledge distillation across two different visual tasks, we introduce a small, trainable network that translates the predicted depth map to a semantic segmentation map, which can then be supervised by the teacher network. In this way, this small network enables the backpropagation from the semantic segmentation teacher's supervision to the depth network during training. In addition, since the commonly used object classes in semantic segmentation are not directly transferable to depth, we study the visual and geometric characteristics of the objects and design a new way of grouping them that can be shared by both tasks. It is noteworthy that our approach only modifies the training process and does not incur additional computation during inference. We extensively evaluate the efficacy of our proposed approach on the standard KITTI benchmark and compare it with the latest state of the art. We further test the generalizability of our approach on Make3D. Overall, the results show that our approach significantly improves the depth estimation accuracy and outperforms the state of the art.

PDF Abstract

Code

Add Remove Mark official

No code implementations yet. Submit your code now

Tasks

Add Remove

Depth Estimation

Knowledge Distillation

Monocular Depth Estimation

Segmentation

Semantic Segmentation

Datasets

Cityscapes

KITTI

Make3D

Results from the Paper

Edit

Ranked #17 on Monocular Depth Estimation on KITTI Eigen split unsupervised

Get a GitHub badge

Task	Dataset	Model	Metric Name	Metric Value	Global Rank	Benchmark
Monocular Depth Estimation	KITTI Eigen split unsupervised	X-Distill (M+1024x320)	absolute relative error	0.102	# 17	Compare
			RMSE	4.439	# 14	Compare
			Sq Rel	0.698	# 12	Compare
			RMSE log	0.180	# 13	Compare
			Delta < 1.25	0.895	# 12	Compare
			Delta < 1.25^2	0.965	# 10	Compare
			Delta < 1.25^3	0.983	# 10	Compare

Methods

Add Remove

Knowledge Distillation • Test

Edit Social Preview

X-Distill: Improving Self-Supervised Monocular Depth via Cross-Task Distillation

Code Edit Add Remove Mark official

Tasks Edit Add Remove

Datasets Edit

Results from the Paper Edit

Methods Edit Add Remove

Code

Add Remove Mark official

Tasks

Add Remove

Datasets

Results from the Paper

Edit

Methods

Add Remove