TASK	DATASET	MODEL	METRIC NAME	METRIC VALUE	GLOBAL RANK
Semantic Segmentation	Cityscapes val	Trans4PASS (Tiny)	mIoU	79.1%	# 50
Semantic Segmentation	Cityscapes val	Trans4PASS (Small)	mIoU	81.1%	# 39
Semantic Segmentation	DensePASS	Trans4PASS (multi-scale)	mIoU	56.38%	# 3
Semantic Segmentation	DensePASS	Trans4PASS (single-scale)	mIoU	55.25%	# 4
Semantic Segmentation	Stanford2D3D Panoramic	Trans4PASS (Supervised + Small)	mIoU	52.1%	# 12
Semantic Segmentation	Stanford2D3D Panoramic	Trans4PASS (Supervised + Small + MS)	mIoU	53.0%	# 8
Semantic Segmentation	Stanford2D3D Panoramic	Trans4PASS (UDA + Source Only)	mIoU	48.1%	# 17
Semantic Segmentation	Stanford2D3D Panoramic	Trans4PASS (UDA + MPA)	mIoU	50.8%	# 15
Semantic Segmentation	Stanford2D3D Panoramic	Trans4PASS (UDA + MPA + MS)	mIoU	51.2%	# 14
Semantic Segmentation	SynPASS	Trans4PASS	mIoU	38.57%	# 2

Badge	Markdown
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/bending-reality-distortion-aware-transformers/semantic-segmentation-on-synpass)](https://paperswithcode.com/sota/semantic-segmentation-on-synpass?p=bending-reality-distortion-aware-transformers)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/bending-reality-distortion-aware-transformers/semantic-segmentation-on-densepass)](https://paperswithcode.com/sota/semantic-segmentation-on-densepass?p=bending-reality-distortion-aware-transformers)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/bending-reality-distortion-aware-transformers/semantic-segmentation-on-stanford2d3d-1)](https://paperswithcode.com/sota/semantic-segmentation-on-stanford2d3d-1?p=bending-reality-distortion-aware-transformers)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/bending-reality-distortion-aware-transformers/semantic-segmentation-on-cityscapes-val)](https://paperswithcode.com/sota/semantic-segmentation-on-cityscapes-val?p=bending-reality-distortion-aware-transformers)`

Bending Reality: Distortion-aware Transformers for Adapting to Panoramic Semantic Segmentation

CVPR 2022 · Jiaming Zhang, Kailun Yang, Chaoxiang Ma, Simon Reiß, Kunyu Peng, Rainer Stiefelhagen ·

Panoramic images with their 360-degree directional view encompass exhaustive information about the surrounding space, providing a rich foundation for scene understanding. To unfold this potential in the form of robust panoramic segmentation models, large quantities of expensive, pixel-wise annotations are crucial for success. Such annotations are available, but predominantly for narrow-angle, pinhole-camera images which, off the shelf, serve as sub-optimal resources for training panoramic models. Distortions and the distinct image-feature distribution in 360-degree panoramas impede the transfer from the annotation-rich pinhole domain and therefore come with a big dent in performance. To get around this domain difference and bring together semantic annotations from pinhole- and 360-degree surround-visuals, we propose to learn object deformations and panoramic image distortions in the Deformable Patch Embedding (DPE) and Deformable MLP (DMLP) components which blend into our Transformer for PAnoramic Semantic Segmentation (Trans4PASS) model. Finally, we tie together shared semantics in pinhole- and panoramic feature embeddings by generating multi-scale prototype features and aligning them in our Mutual Prototypical Adaptation (MPA) for unsupervised domain adaptation. On the indoor Stanford2D3D dataset, our Trans4PASS with MPA maintains comparable performance to fully-supervised state-of-the-arts, cutting the need for over 1,400 labeled panoramas. On the outdoor DensePASS dataset, we break state-of-the-art by 14.39% mIoU and set the new bar at 56.38%. Code will be made publicly available at https://github.com/jamycheung/Trans4PASS.

PDF Abstract CVPR 2022 PDF CVPR 2022 Abstract

Code

Add Remove Mark official

jamycheung/trans4pass official

Tasks

Add Remove

Domain Adaptation

Scene Understanding

Semantic Segmentation

Unsupervised Domain Adaptation

Datasets

Cityscapes

2D-3D-S

DensePASS

Results from the Paper

Add Remove

Ranked #2 on Semantic Segmentation on SynPASS

Get a GitHub badge

Task	Dataset	Model	Metric Name	Metric Value	Global Rank	Benchmark
Semantic Segmentation	Cityscapes val	Trans4PASS (Tiny)	mIoU	79.1%	# 50	Compare
Semantic Segmentation	Cityscapes val	Trans4PASS (Small)	mIoU	81.1%	# 39	Compare
Semantic Segmentation	DensePASS	Trans4PASS (multi-scale)	mIoU	56.38%	# 3	Compare
Semantic Segmentation	DensePASS	Trans4PASS (single-scale)	mIoU	55.25%	# 4	Compare
Semantic Segmentation	Stanford2D3D Panoramic	Trans4PASS (Supervised + Small)	mIoU	52.1%	# 12	Compare
Semantic Segmentation	Stanford2D3D Panoramic	Trans4PASS (Supervised + Small + MS)	mIoU	53.0%	# 8	Compare
Semantic Segmentation	Stanford2D3D Panoramic	Trans4PASS (UDA + Source Only)	mIoU	48.1%	# 17	Compare
Semantic Segmentation	Stanford2D3D Panoramic	Trans4PASS (UDA + MPA)	mIoU	50.8%	# 15	Compare
Semantic Segmentation	Stanford2D3D Panoramic	Trans4PASS (UDA + MPA + MS)	mIoU	51.2%	# 14	Compare
Semantic Segmentation	SynPASS	Trans4PASS	mIoU	38.57%	# 2	Compare

Methods

Add Remove

Absolute Position Encodings • Adam • BPE • Dense Connections • Dropout • Label Smoothing • Layer Normalization • Linear Layer • Multi-Head Attention • Position-Wise Feed-Forward Layer • Residual Connection • Scaled Dot-Product Attention • Softmax • Transformer

Edit Social Preview

Bending Reality: Distortion-aware Transformers for Adapting to Panoramic Semantic Segmentation

Code Edit Add Remove Mark official

Tasks Edit Add Remove

Datasets Edit

Results from the Paper Edit Add Remove

Methods Edit Add Remove

Code

Add Remove Mark official

Tasks

Add Remove

Datasets

Results from the Paper

Add Remove

Methods

Add Remove