Search Results for author: Chen Change Loy

Found 269 papers, 180 papers with code

Region Proposal by Guided Anchoring

2 code implementations • CVPR 2019 • Jiaqi Wang, Kai Chen, Shuo Yang, Chen Change Loy, Dahua Lin

State-of-the-art detectors mostly rely on a dense anchoring scheme, where anchors are sampled uniformly over the spatial domain with a predefined set of scales and aspect ratios.

Ranked #1 on Region Proposal on COCO test-dev

object-detection Object Detection +1

27,806

Paper
Code

Hybrid Task Cascade for Instance Segmentation

5 code implementations • CVPR 2019 • Kai Chen, Jiangmiao Pang, Jiaqi Wang, Yu Xiong, Xiaoxiao Li, Shuyang Sun, Wansen Feng, Ziwei Liu, Jianping Shi, Wanli Ouyang, Chen Change Loy, Dahua Lin

In exploring a more effective approach, we find that the key to a successful instance segmentation cascade is to fully leverage the reciprocal relationship between detection and segmentation.

Ranked #32 on Object Detection on COCO-O

Instance Segmentation object-detection +4

27,806

Paper
Code

Prime Sample Attention in Object Detection

1 code implementation • CVPR 2020 • Yuhang Cao, Kai Chen, Chen Change Loy, Dahua Lin

Our experiments demonstrate that it is often more effective to focus on prime samples than hard samples when training a detector.

Object object-detection +1

27,806

Paper
Code

CARAFE: Content-Aware ReAssembly of FEatures

3 code implementations • ICCV 2019 • Jiaqi Wang, Kai Chen, Rui Xu, Ziwei Liu, Chen Change Loy, Dahua Lin

CARAFE introduces little computational overhead and can be readily integrated into modern network architectures.

Ranked #3 on Feature Upsampling on ImageNet

Feature Upsampling Instance Segmentation +3

27,806

Paper
Code

MMDetection: Open MMLab Detection Toolbox and Benchmark

144 code implementations • 17 Jun 2019 • Kai Chen, Jiaqi Wang, Jiangmiao Pang, Yuhang Cao, Yu Xiong, Xiaoxiao Li, Shuyang Sun, Wansen Feng, Ziwei Liu, Jiarui Xu, Zheng Zhang, Dazhi Cheng, Chenchen Zhu, Tianheng Cheng, Qijie Zhao, Buyu Li, Xin Lu, Rui Zhu, Yue Wu, Jifeng Dai, Jingdong Wang, Jianping Shi, Wanli Ouyang, Chen Change Loy, Dahua Lin

In this paper, we introduce the various features of this toolbox.

Benchmarking Instance Segmentation +2

27,806

Paper
Code

Side-Aware Boundary Localization for More Precise Object Detection

3 code implementations • ECCV 2020 • Jiaqi Wang, Wenwei Zhang, Yuhang Cao, Kai Chen, Jiangmiao Pang, Tao Gong, Jianping Shi, Chen Change Loy, Dahua Lin

To tackle the difficulty of precise localization in the presence of displacements with large variance, we further propose a two-step localization scheme, which first predicts a range of movement through bucket prediction and then pinpoints the precise position within the predicted bucket.

Object object-detection +2

27,806

Paper
Code

Seesaw Loss for Long-Tailed Instance Segmentation

2 code implementations • CVPR 2021 • Jiaqi Wang, Wenwei Zhang, Yuhang Zang, Yuhang Cao, Jiangmiao Pang, Tao Gong, Kai Chen, Ziwei Liu, Chen Change Loy, Dahua Lin

Instances of head classes dominate a long-tailed dataset and they serve as negative samples of tail categories.

Instance Segmentation Semantic Segmentation

27,806

Paper
Code

Feature Pyramid Grids

1 code implementation • 7 Apr 2020 • Kai Chen, Yuhang Cao, Chen Change Loy, Dahua Lin, Christoph Feichtenhofer

Feature pyramid networks have been widely adopted in the object detection literature to improve feature representations for better handling of variations in scale.

Neural Architecture Search object-detection +2

27,799

Paper
Code

Image Super-Resolution Using Deep Convolutional Networks

60 code implementations • 31 Dec 2014 • Chao Dong, Chen Change Loy, Kaiming He, Xiaoou Tang

We further show that traditional sparse-coding-based SR methods can also be viewed as a deep convolutional network.

Ranked #2 on Video Super-Resolution on Xiph HD - 4x upscaling

Image Super-Resolution Video Super-Resolution

27,167

Paper
Code

ESRGAN: Enhanced Super-Resolution Generative Adversarial Networks

45 code implementations • 1 Sep 2018 • Xintao Wang, Ke Yu, Shixiang Wu, Jinjin Gu, Yihao Liu, Chao Dong, Chen Change Loy, Yu Qiao, Xiaoou Tang

To further enhance the visual quality, we thoroughly study three key components of SRGAN - network architecture, adversarial loss and perceptual loss, and improve each of them to derive an Enhanced SRGAN (ESRGAN).

Ranked #2 on Face Hallucination on FFHQ 512 x 512 - 16x upscaling

Face Hallucination Generative Adversarial Network +2

15,717

Paper
Code

Towards Robust Blind Face Restoration with Codebook Lookup Transformer

1 code implementation • 22 Jun 2022 • Shangchen Zhou, Kelvin C. K. Chan, Chongyi Li, Chen Change Loy

In this paper, we demonstrate that a learned discrete codebook prior in a small proxy space largely reduces the uncertainty and ambiguity of restoration mapping by casting blind face restoration as a code prediction task, while providing rich visual atoms for generating high-quality faces.

Ranked #1 on Blind Face Restoration on CelebA-Test

Blind Face Restoration

13,381

Paper
Code

PSANet: Point-wise Spatial Attention Network for Scene Parsing

4 code implementations • ECCV 2018 • Hengshuang Zhao, Yi Zhang, Shu Liu, Jianping Shi, Chen Change Loy, Dahua Lin, Jiaya Jia

We notice information flow in convolutional neural networks is restricted inside local neighborhood regions due to the physical design of convolutional filters, which limits the overall understanding of complex scenes.

Ranked #51 on Semantic Segmentation on Cityscapes test

Position Scene Parsing +1

7,408

Paper
Code

EDVR: Video Restoration with Enhanced Deformable Convolutional Networks

11 code implementations • 7 May 2019 • Xintao Wang, Kelvin C. K. Chan, Ke Yu, Chao Dong, Chen Change Loy

In this work, we propose a novel Video Restoration framework with Enhanced Deformable networks, termed EDVR, to address these challenges.

Ranked #2 on Deblurring on REDS

Deblurring Video Enhancement +2

6,575

Paper
Code

Texture Memory-Augmented Deep Patch-Based Image Inpainting

1 code implementation • 28 Sep 2020 • Rui Xu, Minghao Guo, Jiaqi Wang, Xiaoxiao Li, Bolei Zhou, Chen Change Loy

By bringing together the best of both paradigms, we propose a new deep inpainting framework where texture generation is guided by a texture memory of patch samples extracted from unmasked regions.

Image Inpainting Retrieval +1

6,574

Paper
Code

BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and Alignment

3 code implementations • CVPR 2022 • Kelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change Loy

We show that by empowering the recurrent framework with the enhanced propagation and alignment, one can exploit spatiotemporal information across misaligned video frames more effectively.

Ranked #1 on Video Enhancement on MFQE v2

Analog Video Restoration Video Enhancement +1

6,574

Paper
Code

GLEAN: Generative Latent Bank for Image Super-Resolution and Beyond

1 code implementation • 29 Jul 2022 • Kelvin C. K. Chan, Xiangyu Xu, Xintao Wang, Jinwei Gu, Chen Change Loy

While most existing perceptual-oriented approaches attempt to generate realistic outputs through learning with adversarial loss, our method, Generative LatEnt bANk (GLEAN), goes beyond existing practices by directly leveraging rich and diverse priors encapsulated in a pre-trained GAN.

Colorization Image Colorization +2

6,574

Paper
Code

BasicVSR: The Search for Essential Components in Video Super-Resolution and Beyond

6 code implementations • CVPR 2021 • Kelvin C. K. Chan, Xintao Wang, Ke Yu, Chao Dong, Chen Change Loy

Video super-resolution (VSR) approaches tend to have more components than the image counterparts as they need to exploit the additional temporal dimension.

Ranked #2 on Video Super-Resolution on MSU Video Super Resolution Benchmark: Detail Restoration

Video Super-Resolution

6,573

Paper
Code

Recovering Realistic Texture in Image Super-resolution by Deep Spatial Feature Transform

4 code implementations • CVPR 2018 • Xintao Wang, Ke Yu, Chao Dong, Chen Change Loy

In this paper, we show that it is possible to recover textures faithful to semantic classes.

Ranked #55 on Image Super-Resolution on BSD100 - 4x upscaling

Image Super-Resolution Semantic Segmentation

6,187

Paper
Code

Exploring Data Augmentation for Multi-Modality 3D Object Detection

8 code implementations • 23 Dec 2020 • Wenwei Zhang, Zhe Wang, Chen Change Loy

Due to the fact that multi-modality data augmentation must maintain consistency between point cloud and images, recent methods in this field typically use relatively insufficient data augmentation.

3D Object Detection Autonomous Driving +3

4,810

Paper
Code

ProPainter: Improving Propagation and Transformer for Video Inpainting

1 code implementation • ICCV 2023 • Shangchen Zhou, Chongyi Li, Kelvin C. K. Chan, Chen Change Loy

We also propose a mask-guided sparse video Transformer, which achieves high efficiency by discarding unnecessary and redundant tokens.

Ranked #1 on Video Inpainting on YouTube-VOS 2018

Optical Flow Estimation Video Inpainting

4,627

Paper
Code

VToonify: Controllable High-Resolution Portrait Video Style Transfer

1 code implementation • 22 Sep 2022 • Shuai Yang, Liming Jiang, Ziwei Liu, Chen Change Loy

Although a series of successful portrait image toonification models built upon the powerful StyleGAN have been proposed, these image-oriented methods have obvious limitations when applied to videos, such as the fixed frame size, the requirement of face alignment, missing non-facial details and temporal inconsistency.

Face Alignment Style Transfer +2

3,463

Paper
Code

Online Deep Clustering for Unsupervised Representation Learning

1 code implementation • CVPR 2020 • Xiaohang Zhan, Jiahao Xie, Ziwei Liu, Yew Soon Ong, Chen Change Loy

In this way, labels and the network evolve shoulder-to-shoulder rather than alternatingly.

Clustering Deep Clustering +1

3,084

Paper
Code

Delving into Inter-Image Invariance for Unsupervised Visual Representations

2 code implementations • 26 Aug 2020 • Jiahao Xie, Xiaohang Zhan, Ziwei Liu, Yew Soon Ong, Chen Change Loy

In this work, we present a comprehensive empirical study to better understand the role of inter-image invariance learning from three main constituting components: pseudo-label maintenance, sampling strategy, and decision boundary design.

Contrastive Learning Pseudo Label +1

3,084

Paper
Code

PolyNet: A Pursuit of Structural Diversity in Very Deep Networks

3 code implementations • CVPR 2017 • Xingcheng Zhang, Zhizhong Li, Chen Change Loy, Dahua Lin

A number of studies have shown that increasing the depth or width of convolutional networks is a rewarding approach to improve the performance of image recognition.

Image Classification

2,917

Paper
Code

Deep Flow-Guided Video Inpainting

2 code implementations • CVPR 2019 • Rui Xu, Xiaoxiao Li, Bolei Zhou, Chen Change Loy

Then the synthesized flow field is used to guide the propagation of pixels to fill up the missing regions in the video.

Ranked #8 on Video Inpainting on DAVIS

One-shot visual object segmentation Optical Flow Estimation +2

2,316

Paper
Code

Exploiting Diffusion Prior for Real-World Image Super-Resolution

3 code implementations • 11 May 2023 • Jianyi Wang, Zongsheng Yue, Shangchen Zhou, Kelvin C. K. Chan, Chen Change Loy

We present a novel approach to leverage prior knowledge encapsulated in pre-trained text-to-image diffusion models for blind super-resolution (SR).

Blind Super-Resolution Image Super-Resolution

1,823

Paper
Code

Pastiche Master: Exemplar-Based High-Resolution Portrait Style Transfer

1 code implementation • CVPR 2022 • Shuai Yang, Liming Jiang, Ziwei Liu, Chen Change Loy

Recent studies on StyleGAN show high performance on artistic portrait generation by transfer learning with limited data.

Style Transfer Transfer Learning +1

1,573

Paper
Code

Learning to Prompt for Vision-Language Models

13 code implementations • 2 Sep 2021 • Kaiyang Zhou, Jingkang Yang, Chen Change Loy, Ziwei Liu

Large pre-trained vision-language models like CLIP have shown great potential in learning representations that are transferable across a wide range of downstream tasks.

Ranked #2 on Few-shot Age Estimation on MORPH Album2

Domain Generalization Few-shot Age Estimation +2

1,481

Paper
Code

Conditional Prompt Learning for Vision-Language Models

9 code implementations • CVPR 2022 • Kaiyang Zhou, Jingkang Yang, Chen Change Loy, Ziwei Liu

With the rise of powerful pre-trained vision-language models like CLIP, it becomes essential to investigate ways to adapt these models to downstream datasets.

Ranked #3 on Prompt Engineering on ImageNet V2

Domain Generalization Prompt Engineering

1,481

Paper
Code

Knowledge Distillation Meets Self-Supervision

2 code implementations • ECCV 2020 • Guodong Xu, Ziwei Liu, Xiaoxiao Li, Chen Change Loy

Knowledge distillation, which involves extracting the "dark knowledge" from a teacher network to guide the learning of a student network, has emerged as an important technique for model compression and transfer learning.

Ranked #32 on Knowledge Distillation on ImageNet

Contrastive Learning Knowledge Distillation +2

1,272

Paper
Code

StyleGAN-Human: A Data-Centric Odyssey of Human Generation

4 code implementations • 25 Apr 2022 • Jianglin Fu, Shikai Li, Yuming Jiang, Kwan-Yee Lin, Chen Qian, Chen Change Loy, Wayne Wu, Ziwei Liu

In addition, a model zoo and human editing applications are demonstrated to facilitate future research in the community.

Image Generation

1,096

Paper
Code

Domain Generalization: A Survey

2 code implementations • 3 Mar 2021 • Kaiyang Zhou, Ziwei Liu, Yu Qiao, Tao Xiang, Chen Change Loy

Generalization to out-of-distribution (OOD) data is a capability natural to humans yet challenging for machines to reproduce.

Action Recognition Data Augmentation +8

1,083

Paper
Code

Semi-Supervised Domain Generalization with Stochastic StyleMatch

2 code implementations • 1 Jun 2021 • Kaiyang Zhou, Chen Change Loy, Ziwei Liu

We find that the DG methods, which by design are unable to handle unlabeled data, perform poorly with limited labels in SSDG; the SSL methods, especially FixMatch, obtain much better results but are still far away from the basic vanilla model trained using full labels.

Domain Generalization Semi-Supervised Domain Generalization

1,083

Paper
Code

Learning Lightweight Lane Detection CNNs by Self Attention Distillation

2 code implementations • ICCV 2019 • Yuenan Hou, Zheng Ma, Chunxiao Liu, Chen Change Loy

Training deep models for lane detection is challenging due to the very subtle and sparse supervisory signals inherent in lane annotations.

Ranked #5 on Lane Detection on BDD100K val

Knowledge Distillation Lane Detection +1

1,023

Paper
Code

Pose-Controllable Talking Face Generation by Implicitly Modularized Audio-Visual Representation

1 code implementation • CVPR 2021 • Hang Zhou, Yasheng Sun, Wayne Wu, Chen Change Loy, Xiaogang Wang, Ziwei Liu

While speech content information can be defined by learning the intrinsic synchronization between audio-visual modalities, we identify that a pose code will be complementarily learned in a modulated convolution-based reconstruction framework.

Talking Face Generation

904

Paper
Code

LiteFlowNet: A Lightweight Convolutional Neural Network for Optical Flow Estimation

4 code implementations • CVPR 2018 • Tak-Wai Hui, Xiaoou Tang, Chen Change Loy

FlowNet2, the state-of-the-art convolutional neural network (CNN) for optical flow estimation, requires over 160M parameters to achieve accurate flow estimation.

Ranked #10 on Optical Flow Estimation on KITTI 2012

Optical Flow Estimation

892

Paper
Code

A Lightweight Optical Flow CNN -- Revisiting Data Fidelity and Regularization

3 code implementations • 15 Mar 2019 • Tak-Wai Hui, Xiaoou Tang, Chen Change Loy

Over four decades, the majority addresses the problem of optical flow estimation using variational methods.

Ranked #6 on Optical Flow Estimation on KITTI 2012

Optical Flow Estimation

892

Paper
Code

Investigating Tradeoffs in Real-World Video Super-Resolution

1 code implementation • CVPR 2022 • Kelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change Loy

The diversity and complexity of degradations in real-world video super-resolution (VSR) pose non-trivial challenges in inference and training.

Ranked #9 on Video Super-Resolution on MSU Video Upscalers: Quality Enhancement

Benchmarking Video Super-Resolution

841

Paper
Code

Text2Human: Text-Driven Controllable Human Image Generation

2 code implementations • 31 May 2022 • Yuming Jiang, Shuai Yang, Haonan Qiu, Wayne Wu, Chen Change Loy, Ziwei Liu

In this work, we present a text-driven controllable framework, Text2Human, for a high-quality and diverse human generation.

Human Parsing Image Generation

804

Paper
Code

Self-Supervised Scene De-occlusion

2 code implementations • CVPR 2020 • Xiaohang Zhan, Xingang Pan, Bo Dai, Ziwei Liu, Dahua Lin, Chen Change Loy

This is achieved via Partial Completion Network (PCNet)-mask (M) and -content (C), that learn to recover fractions of object masks and contents, respectively, in a self-supervised manner.

Image Manipulation Scene Understanding

770

Paper
Code

Zero-Reference Deep Curve Estimation for Low-Light Image Enhancement

9 code implementations • CVPR 2020 • Chunle Guo, Chongyi Li, Jichang Guo, Chen Change Loy, Junhui Hou, Sam Kwong, Runmin Cong

The paper presents a novel method, Zero-Reference Deep Curve Estimation (Zero-DCE), which formulates light enhancement as a task of image-specific curve estimation with a deep network.

Ranked #1 on Color Constancy on INTEL-TUT2

Color Constancy Face Detection +1

732

Paper
Code

LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models

2 code implementations • 26 Sep 2023 • Yaohui Wang, Xinyuan Chen, Xin Ma, Shangchen Zhou, Ziqi Huang, Yi Wang, Ceyuan Yang, Yinan He, Jiashuo Yu, Peiqing Yang, Yuwei Guo, Tianxing Wu, Chenyang Si, Yuming Jiang, Cunjian Chen, Chen Change Loy, Bo Dai, Dahua Lin, Yu Qiao, Ziwei Liu

To this end, we propose LaVie, an integrated video generation framework that operates on cascaded video latent diffusion models, comprising a base T2V model, a temporal interpolation model, and a video super-resolution model.

Ranked #4 on Text-to-Video Generation on EvalCrafter Text-to-Video (ECTV) Dataset (using extra training data)

Text-to-Video Generation Video Generation +1

727

Paper
Code

Learning to Cluster Faces on an Affinity Graph

3 code implementations • CVPR 2019 • Lei Yang, Xiaohang Zhan, Dapeng Chen, Junjie Yan, Chen Change Loy, Dahua Lin

Face recognition sees remarkable progress in recent years, and its performance has reached a very high level.

Clustering Face Recognition +1

698

Paper
Code

Learning to Cluster Faces via Confidence and Connectivity Estimation

3 code implementations • CVPR 2020 • Lei Yang, Dapeng Chen, Xiaohang Zhan, Rui Zhao, Chen Change Loy, Dahua Lin

With the vertex confidence and edge connectivity, we can naturally organize more relevant vertices on the affinity graph and group them into clusters.

Clustering Connectivity Estimation +2

698

Paper
Code

Low-Light Image and Video Enhancement Using Deep Learning: A Survey

3 code implementations • 21 Apr 2021 • Chongyi Li, Chunle Guo, Linghao Han, Jun Jiang, Ming-Ming Cheng, Jinwei Gu, Chen Change Loy

Low-light image enhancement (LLIE) aims at improving the perception or interpretability of an image captured in an environment with poor illumination.

Face Detection Low-Light Image Enhancement +1

685

Paper
Code

EdgeSAM: Prompt-In-the-Loop Distillation for On-Device Deployment of SAM

1 code implementation • 11 Dec 2023 • Chong Zhou, Xiangtai Li, Chen Change Loy, Bo Dai

It is also the first SAM variant that can run at over 30 FPS on an iPhone 14.

681

Paper
Code

OMG-Seg: Is One Model Good Enough For All Segmentation?

1 code implementation • 18 Jan 2024 • Xiangtai Li, Haobo Yuan, Wei Li, Henghui Ding, Size Wu, Wenwei Zhang, Yining Li, Kai Chen, Chen Change Loy

In this work, we address various segmentation tasks, each traditionally tackled by distinct or partially unified models.

Interactive Segmentation Panoptic Segmentation +3

681

Paper
Code

FRESCO: Spatial-Temporal Correspondence for Zero-Shot Video Translation

2 code implementations • 19 Mar 2024 • Shuai Yang, Yifan Zhou, Ziwei Liu, Chen Change Loy

In this paper, we introduce FRESCO, intra-frame correspondence alongside inter-frame correspondence to establish a more robust spatial-temporal constraint.

Translation valid

603

Paper
Code

Focal Frequency Loss for Image Reconstruction and Synthesis

1 code implementation • ICCV 2021 • Liming Jiang, Bo Dai, Wayne Wu, Chen Change Loy

In this study, we show that narrowing gaps in the frequency domain can ameliorate image reconstruction and synthesis quality further.

Ranked #6 on Image-to-Image Translation on Cityscapes Labels-to-Photo

Image Reconstruction Image-to-Image Translation

591

Paper
Code

DifFace: Blind Face Restoration with Diffused Error Contraction

2 code implementations • 13 Dec 2022 • Zongsheng Yue, Chen Change Loy

Moreover, the transition distribution can contract the error of the restoration backbone and thus makes our method more robust to unknown degradations.

Ranked #5 on Blind Face Restoration on CelebA-Test

Blind Face Restoration

586

Paper
Code

Open-Vocabulary SAM: Segment and Recognize Twenty-thousand Classes Interactively

1 code implementation • 5 Jan 2024 • Haobo Yuan, Xiangtai Li, Chong Zhou, Yining Li, Kai Chen, Chen Change Loy

The CLIP and Segment Anything Model (SAM) are remarkable vision foundation models (VFMs).

Image Classification Interactive Segmentation +3

586

Paper
Code

Transformer-Based Visual Segmentation: A Survey

2 code implementations • 19 Apr 2023 • Xiangtai Li, Henghui Ding, Haobo Yuan, Wenwei Zhang, Jiangmiao Pang, Guangliang Cheng, Kai Chen, Ziwei Liu, Chen Change Loy

Recently, transformers, a type of neural network based on self-attention originally designed for natural language processing, have considerably surpassed previous convolutional or recurrent approaches in various vision processing tasks.

Autonomous Driving Point Cloud Segmentation +1

574

Paper
Code

Do 2D GANs Know 3D Shape? Unsupervised 3D shape reconstruction from 2D Image GANs

1 code implementation • ICLR 2021 • Xingang Pan, Bo Dai, Ziwei Liu, Chen Change Loy, Ping Luo

Through our investigation, we found that such a pre-trained GAN indeed contains rich 3D knowledge and thus can be used to recover 3D shape from a single 2D image in an unsupervised manner.

3D Shape Reconstruction Object

570

Paper
Code

On the Generalization of BasicVSR++ to Video Deblurring and Denoising

1 code implementation • 11 Apr 2022 • Kelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change Loy

The exploitation of long-term information has been a long-standing problem in video restoration.

Deblurring Denoising +2

531

Paper
Code

ResShift: Efficient Diffusion Model for Image Super-resolution by Residual Shifting

1 code implementation • NeurIPS 2023 • Zongsheng Yue, Jianyi Wang, Chen Change Loy

Diffusion-based image super-resolution (SR) methods are mainly limited by the low inference speed due to the requirements of hundreds or even thousands of sampling steps.

Image Super-Resolution

531

Paper
Code

Efficient Diffusion Model for Image Restoration by Residual Shifting

1 code implementation • 12 Mar 2024 • Zongsheng Yue, Jianyi Wang, Chen Change Loy

While diffusion-based image restoration (IR) methods have achieved remarkable success, they are still limited by the low inference speed attributed to the necessity of executing hundreds or even thousands of sampling steps.

Blind Face Restoration Image Inpainting +2

531

Paper
Code

DeeperForensics-1.0: A Large-Scale Dataset for Real-World Face Forgery Detection

1 code implementation • CVPR 2020 • Liming Jiang, Ren Li, Wayne Wu, Chen Qian, Chen Change Loy

The quality of generated videos outperforms those in existing datasets, validated by user studies.

Face Swapping Video Forensics

524

Paper
Code

DeeperForensics Challenge 2020 on Real-World Face Forgery Detection: Methods and Results

2 code implementations • 18 Feb 2021 • Liming Jiang, Zhengkui Guo, Wayne Wu, Zhaoyang Liu, Ziwei Liu, Chen Change Loy, Shuo Yang, Yuanjun Xiong, Wei Xia, Baoying Chen, Peiyu Zhuang, Sili Li, Shen Chen, Taiping Yao, Shouhong Ding, Jilin Li, Feiyue Huang, Liujuan Cao, Rongrong Ji, Changlei Lu, Ganchao Tan

This paper reports methods and results in the DeeperForensics Challenge 2020 on real-world face forgery detection.

valid

524

Paper
Code

Network Pruning via Resource Reallocation

1 code implementation • 2 Mar 2021 • Yuenan Hou, Zheng Ma, Chunxiao Liu, Zhe Wang, Chen Change Loy

Channel pruning is broadly recognized as an effective approach to obtain a small compact model through eliminating unimportant channels from a large cumbersome network.

Network Pruning

478

Paper
Code

Exploiting Deep Generative Prior for Versatile Image Restoration and Manipulation

1 code implementation • ECCV 2020 • Xingang Pan, Xiaohang Zhan, Bo Dai, Dahua Lin, Chen Change Loy, Ping Luo

Learning a good image prior is a long-term goal for image restoration and manipulation.

Generative Adversarial Network Image Manipulation +2

474

Paper
Code

StyleGANEX: StyleGAN-Based Manipulation Beyond Cropped Aligned Faces

1 code implementation • ICCV 2023 • Shuai Yang, Liming Jiang, Ziwei Liu, Chen Change Loy

Recent advances in face manipulation using StyleGAN have produced impressive results.

Attribute Super-Resolution

468

Paper
Code

MeViS: A Large-scale Benchmark for Video Segmentation with Motion Expressions

1 code implementation • ICCV 2023 • Henghui Ding, Chang Liu, Shuting He, Xudong Jiang, Chen Change Loy

To investigate the feasibility of using motion expressions to ground and segment objects in videos, we propose a large-scale dataset called MeViS, which contains numerous motion expressions to indicate target objects in complex environments.

Ranked #2 on Referring Video Object Segmentation on MeViS

Motion Expressions Guided Video Segmentation Object +6

459

Paper
Code

K-Net: Towards Unified Image Segmentation

1 code implementation • NeurIPS 2021 • Wenwei Zhang, Jiangmiao Pang, Kai Chen, Chen Change Loy

The framework, named K-Net, segments both instances and semantic categories consistently by a group of learnable kernels, where each kernel is responsible for generating a mask for either a potential instance or a stuff class.

Ranked #7 on Panoptic Segmentation on COCO test-dev

Image Segmentation Instance Segmentation +2

457

Paper
Code

Consensus-Driven Propagation in Massive Unlabeled Data for Face Recognition

6 code implementations • ECCV 2018 • Xiaohang Zhan, Ziwei Liu, Junjie Yan, Dahua Lin, Chen Change Loy

Face recognition has witnessed great progress in recent years, mainly attributed to the high-capacity model designed and the abundant labeled data collected.

Face Recognition

453

Paper
Code

1st Place Solutions for OpenImage2019 -- Object Detection and Instance Segmentation

2 code implementations • 17 Mar 2020 • Yu Liu, Guanglu Song, Yuhang Zang, Yan Gao, Enze Xie, Junjie Yan, Chen Change Loy, Xiaogang Wang

Given such good instance bounding box, we further design a simple instance-level semantic segmentation pipeline and achieve the 1st place on the segmentation challenge.

General Classification Instance Segmentation +6

453

Paper
Code

Optimizing Video Object Detection via a Scale-Time Lattice

1 code implementation • CVPR 2018 • Kai Chen, Jiaqi Wang, Shuo Yang, Xingcheng Zhang, Yuanjun Xiong, Chen Change Loy, Dahua Lin

High-performance object detection relies on expensive convolutional networks to compute features, often leading to significant challenges in applications, e. g. those that require detecting objects from video streams in real time.

Object object-detection +1

451

Paper
Code

The Devil of Face Recognition is in the Noise

2 code implementations • ECCV 2018 • Fei Wang, Liren Chen, Cheng Li, Shiyao Huang, Yanjie Chen, Chen Qian, Chen Change Loy

2) With the original datasets and cleaned subsets, we profile and analyze label noise properties of MegaFace and MS-Celeb-1M.

Face Recognition

431

Paper
Code

Pose-Robust Face Recognition via Deep Residual Equivariant Mapping

1 code implementation • CVPR 2018 • Kaidi Cao, Yu Rong, Cheng Li, Xiaoou Tang, Chen Change Loy

However, many contemporary face recognition models still perform relatively poor in processing profile faces compared to frontal faces.

Ranked #1 on Face Identification on IJB-A

Face Identification Face Recognition +2

388

Paper
Code

Deep Animation Video Interpolation in the Wild

1 code implementation • CVPR 2021 • Li SiYao, Shiyu Zhao, Weijiang Yu, Wenxiu Sun, Dimitris N. Metaxas, Chen Change Loy, Ziwei Liu

In the animation industry, cartoon videos are usually produced at low frame rate since hand drawing of such frames is costly and time-consuming.

Optical Flow Estimation Video Frame Interpolation

385

Paper
Code

Bailando: 3D Dance Generation by Actor-Critic GPT with Choreographic Memory

1 code implementation • CVPR 2022 • Li SiYao, Weijiang Yu, Tianpei Gu, Chunze Lin, Quan Wang, Chen Qian, Chen Change Loy, Ziwei Liu

With the learned choreographic memory, dance generation is realized on the quantized units that meet high choreography standards, such that the generated dancing sequences are confined within the spatial constraints.

Ranked #1 on Motion Synthesis on AIST++

Motion Synthesis

366

Paper
Code

CelebV-Text: A Large-Scale Facial Text-Video Dataset

1 code implementation • CVPR 2023 • Jianhui Yu, Hao Zhu, Liming Jiang, Chen Change Loy, Weidong Cai, Wayne Wu

This paper presents CelebV-Text, a large-scale, diverse, and high-quality dataset of facial text-video pairs, to facilitate research on facial text-to-video generation tasks.

Text Generation Text-to-Video Generation +1

362

Paper
Code

Extract Free Dense Labels from CLIP

1 code implementation • 2 Dec 2021 • Chong Zhou, Chen Change Loy, Bo Dai

Contrastive Language-Image Pre-training (CLIP) has made a remarkable breakthrough in open-vocabulary zero-shot image recognition.

Ranked #3 on Unsupervised Semantic Segmentation with Language-image Pre-training on KITTI-STEP

Novel Concepts Open Vocabulary Panoptic Segmentation +5

359

Paper
Code

CelebV-HQ: A Large-Scale Video Facial Attributes Dataset

1 code implementation • 25 Jul 2022 • Hao Zhu, Wayne Wu, Wentao Zhu, Liming Jiang, Siwei Tang, Li Zhang, Ziwei Liu, Chen Change Loy

Large-scale datasets have played indispensable roles in the recent success of face generation/editing and significantly facilitated the advances of emerging research fields.

Ranked #1 on Unconditional Video Generation on CelebV-HQ

Attribute Face Generation +1

352

Paper
Code

Text2Performer: Text-Driven Human Video Generation

1 code implementation • ICCV 2023 • Yuming Jiang, Shuai Yang, Tong Liang Koh, Wayne Wu, Chen Change Loy, Ziwei Liu

In this work, we present Text2Performer to generate vivid human videos with articulated motions from texts.

Video Generation

308

Paper
Code

Learning to Enhance Low-Light Image via Zero-Reference Deep Curve Estimation

4 code implementations • 1 Mar 2021 • Chongyi Li, Chunle Guo, Chen Change Loy

This paper presents a novel method, Zero-Reference Deep Curve Estimation (Zero-DCE), which formulates light enhancement as a task of image-specific curve estimation with a deep network.

Face Detection Image Enhancement

307

Paper
Code

Cross-Scale Internal Graph Neural Network for Image Super-Resolution

1 code implementation • NeurIPS 2020 • Shangchen Zhou, Jiawei Zhang, WangMeng Zuo, Chen Change Loy

Specifically, we dynamically construct a cross-scale graph by searching k-nearest neighboring patches in the downsampled LR image for each query patch in the LR image.

Image Restoration Image Super-Resolution

304

Paper
Code

Talk-to-Edit: Fine-Grained Facial Editing via Dialog

1 code implementation • ICCV 2021 • Yuming Jiang, Ziqi Huang, Xingang Pan, Chen Change Loy, Ziwei Liu

In this work, we propose Talk-to-Edit, an interactive facial editing framework that performs fine-grained attribute manipulation through dialog between the user and the system.

Ranked #1 on Fine-Grained Facial Editing on CelebA-Dialog

Attribute Facial Editing +1

302

Paper
Code

Video Object Segmentation with Re-identification

3 code implementations • 1 Aug 2017 • Xiaoxiao Li, Yuankai Qi, Zhe Wang, Kai Chen, Ziwei Liu, Jianping Shi, Ping Luo, Xiaoou Tang, Chen Change Loy

Specifically, our Video Object Segmentation with Re-identification (VS-ReID) model includes a mask propagation module and a ReID module.

Object Segmentation +4

289

Paper
Code

Real or Not Real, that is the Question

2 code implementations • ICLR 2020 • Yuanbo Xiangli, Yubin Deng, Bo Dai, Chen Change Loy, Dahua Lin

While generative adversarial networks (GAN) have been widely adopted in various topics, in this paper we generalize the standard GAN to a new perspective by treating realness as a random variable that can be estimated from multiple angles.

286

Paper
Code

Deep Geometrized Cartoon Line Inbetweening

1 code implementation • ICCV 2023 • Li SiYao, Tianpei Gu, Weiye Xiao, Henghui Ding, Ziwei Liu, Chen Change Loy

To preserve the precision and detail of the line drawings, we propose a new approach, AnimeInbet, which geometrizes raster line drawings into graphs of endpoints and reframes the inbetweening task as a graph fusion problem with vertex repositioning.

285

Paper
Code

Audio-Driven Emotional Video Portraits

1 code implementation • CVPR 2021 • Xinya Ji, Hang Zhou, Kaisiyuan Wang, Wayne Wu, Chen Change Loy, Xun Cao, Feng Xu

In this work, we present Emotional Video Portraits (EVP), a system for synthesizing high-quality video portraits with vivid emotional dynamics driven by audios.

Disentanglement Face Generation

284

Paper
Code

TSIT: A Simple and Versatile Framework for Image-to-Image Translation

1 code implementation • ECCV 2020 • Liming Jiang, Changxu Zhang, Mingyang Huang, Chunxiao Liu, Jianping Shi, Chen Change Loy

We introduce a simple and versatile framework for image-to-image translation.

Image-to-Image Translation Translation

271

Paper
Code

Exploring CLIP for Assessing the Look and Feel of Images

1 code implementation • 25 Jul 2022 • Jianyi Wang, Kelvin C. K. Chan, Chen Change Loy

Measuring the perception of visual content is a long-standing problem in computer vision.

Ranked #9 on Video Quality Assessment on MSU SR-QA Dataset

Image Quality Assessment Video Quality Assessment

260

Paper
Code

On-Device Domain Generalization

2 code implementations • 15 Sep 2022 • Kaiyang Zhou, Yuanhan Zhang, Yuhang Zang, Jingkang Yang, Chen Change Loy, Ziwei Liu

Another interesting observation is that the teacher-student gap on out-of-distribution data is bigger than that on in-distribution data, which highlights the capacity mismatch issue as well as the shortcoming of KD.

Data Augmentation Domain Generalization +2

256

Paper
Code

Robust Multi-Modality Multi-Object Tracking

1 code implementation • ICCV 2019 • Wenwei Zhang, Hui Zhou, Shuyang Sun, Zhe Wang, Jianping Shi, Chen Change Loy

Multi-sensor perception is crucial to ensure the reliability and accuracy in autonomous driving system, while multi-object tracking (MOT) improves that by tracing sequential movement of dynamic objects.

Ranked #10 on Multiple Object Tracking on KITTI Tracking test

Autonomous Driving Multi-Object Tracking +2

252

Paper
Code

Deceive D: Adaptive Pseudo Augmentation for GAN Training with Limited Data

2 code implementations • NeurIPS 2021 • Liming Jiang, Bo Dai, Wayne Wu, Chen Change Loy

Generative adversarial networks (GANs) typically require ample data for training in order to synthesize high-fidelity images.

251

Paper
Code

Scenimefy: Learning to Craft Anime Scene via Semi-Supervised Image-to-Image Translation

1 code implementation • ICCV 2023 • Yuxin Jiang, Liming Jiang, Shuai Yang, Chen Change Loy

The challenges of this task lie in the complexity of the scenes, the unique features of anime style, and the lack of high-quality datasets to bridge the domain gap.

Image-to-Image Translation

251

Paper
Code

LiteFlowNet3: Resolving Correspondence Ambiguity for More Accurate Optical Flow Estimation

1 code implementation • ECCV 2020 • Tak-Wai Hui, Chen Change Loy

The keys to success lie in the use of cost volume and coarse-to-fine flow inference.

Ranked #4 on Optical Flow Estimation on KITTI 2012

Optical Flow Estimation Scene Flow Estimation

239

Paper
Code

RenderMe-360: A Large Digital Asset Library and Benchmarks Towards High-fidelity Head Avatars

1 code implementation • NeurIPS 2023 • Dongwei Pan, Long Zhuo, Jingtan Piao, Huiwen Luo, Wei Cheng, Yuxin Wang, Siming Fan, Shengqi Liu, Lei Yang, Bo Dai, Ziwei Liu, Chen Change Loy, Chen Qian, Wayne Wu, Dahua Lin, Kwan-Yee Lin

It is a large-scale digital library for head avatars with three key attributes: 1) High Fidelity: all subjects are captured by 60 synchronized, high-resolution 2K cameras in 360 degrees.

2k Image Matting +2

217

Paper
Code

3D Human Texture Estimation from a Single Image with Transformers

1 code implementation • ICCV 2021 • Xiangyu Xu, Chen Change Loy

We propose a Transformer-based framework for 3D human texture estimation from a single image.

Garment Reconstruction

216

Paper
Code

Accelerating the Super-Resolution Convolutional Neural Network

14 code implementations • 1 Aug 2016 • Chao Dong, Chen Change Loy, Xiaoou Tang

As a successful deep model applied in image super-resolution (SR), the Super-Resolution Convolutional Neural Network (SRCNN) has demonstrated superior performance to the previous hand-crafted models either in speed and restoration quality.

Ranked #6 on Image Super-Resolution on FFHQ 256 x 256 - 4x upscaling

Image Super-Resolution

213

Paper
Code

Crafting a Toolchain for Image Restoration by Deep Reinforcement Learning

2 code implementations • CVPR 2018 • Ke Yu, Chao Dong, Liang Lin, Chen Change Loy

We investigate a novel approach for image restoration by reinforcement learning.

Image Restoration reinforcement-learning +1

208

Paper
Code

DNA-Rendering: A Diverse Neural Actor Repository for High-Fidelity Human-centric Rendering

1 code implementation • ICCV 2023 • Wei Cheng, Ruixiang Chen, Wanqi Yin, Siming Fan, Keyu Chen, Honglin He, Huiwen Luo, Zhongang Cai, Jingbo Wang, Yang Gao, Zhengming Yu, Zhengyu Lin, Daxuan Ren, Lei Yang, Ziwei Liu, Chen Change Loy, Chen Qian, Wayne Wu, Dahua Lin, Bo Dai, Kwan-Yee Lin

Realistic human-centric rendering plays a key role in both computer vision and computer graphics.

Camera Calibration Novel View Synthesis

199

Paper
Code

ReenactGAN: Learning to Reenact Faces via Boundary Transfer

1 code implementation • ECCV 2018 • Wayne Wu, Yunxuan Zhang, Cheng Li, Chen Qian, Chen Change Loy

A transformer is subsequently used to adapt the boundary of source face to the boundary of target face.

Face Reenactment Talking Face Generation +1

195

Paper
Code

Open-Vocabulary DETR with Conditional Matching

1 code implementation • 22 Mar 2022 • Yuhang Zang, Wei Li, Kaiyang Zhou, Chen Huang, Chen Change Loy

To this end, we propose a novel open-vocabulary detector based on DETR -- hence the name OV-DETR -- which, once trained, can detect any object given its class name or an exemplar image.

Ranked #21 on Open Vocabulary Object Detection on MSCOCO

Language Modelling object-detection +1

194

Paper
Code

Robust Reference-based Super-Resolution via C2-Matching

1 code implementation • CVPR 2021 • Yuming Jiang, Kelvin C. K. Chan, Xintao Wang, Chen Change Loy, Ziwei Liu

However, performing local transfer is difficult because of two gaps between input and reference images: the transformation gap (e. g. scale and rotation) and the resolution gap (e. g. HR and LR).

Reference-based Super-Resolution

193

Paper
Code

Reference-based Image and Video Super-Resolution via C2-Matching

1 code implementation • 19 Dec 2022 • Yuming Jiang, Kelvin C. K. Chan, Xintao Wang, Chen Change Loy, Ziwei Liu

To tackle these challenges, we propose C2-Matching in this work, which performs explicit robust matching crossing transformation and resolution.

Image Super-Resolution Reference-based Super-Resolution +2

193

Paper
Code

One-shot Face Reenactment

2 code implementations • 5 Aug 2019 • Yunxuan Zhang, Siwei Zhang, Yue He, Cheng Li, Chen Change Loy, Ziwei Liu

However, in real-world scenario end-users often only have one target face at hand, rendering existing methods inapplicable.

Face Reconstruction Face Reenactment

190

Paper
Code

LEDNet: Joint Low-light Enhancement and Deblurring in the Dark

1 code implementation • 7 Feb 2022 • Shangchen Zhou, Chongyi Li, Chen Change Loy

With the pipeline, we present the first large-scale dataset for joint low-light enhancement and deblurring.

Ranked #2 on Low-Light Image Enhancement on Sony-Total-Dark

Deblurring Low-light Image Deblurring and Enhancement +1

187

Paper
Code

Unsupervised Image-to-Image Translation with Generative Prior

1 code implementation • CVPR 2022 • Shuai Yang, Liming Jiang, Ziwei Liu, Chen Change Loy

In this work, we present a novel framework, Generative Prior-guided UNsupervised Image-to-image Translation (GP-UNIT), to improve the overall quality and applicability of the translation algorithm.

Translation Unsupervised Image-To-Image Translation

181

Paper
Code

GP-UNIT: Generative Prior for Versatile Unsupervised Image-to-Image Translation

1 code implementation • 7 Jun 2023 • Shuai Yang, Liming Jiang, Ziwei Liu, Chen Change Loy

In this paper, we introduce a novel versatile framework, Generative Prior-guided UNsupervised Image-to-image Translation (GP-UNIT), that improves the quality, applicability and controllability of the existing translation models.

Translation Unsupervised Image-To-Image Translation +1

181

Paper
Code

Learning Inclusion Matching for Animation Paint Bucket Colorization

1 code implementation • 27 Mar 2024 • Yuekun Dai, Shangchen Zhou, Qinyue Li, Chongyi Li, Chen Change Loy

In this work, we introduce a new learning-based inclusion matching pipeline, which directs the network to comprehend the inclusion relationships between segments rather than relying solely on direct visual correspondences.

Colorization

175

Paper
Code

TransEditor: Transformer-Based Dual-Space GAN for Highly Controllable Facial Editing

1 code implementation • CVPR 2022 • Yanbo Xu, Yueqin Yin, Liming Jiang, Qianyi Wu, Chengyao Zheng, Chen Change Loy, Bo Dai, Wayne Wu

In this study, we highlight the importance of interaction in a dual-space GAN for more controllable editing.

Attribute Disentanglement +1

173

Paper
Code

Learning Generative Structure Prior for Blind Text Image Super-resolution

1 code implementation • CVPR 2023 • Xiaoming Li, WangMeng Zuo, Chen Change Loy

To restrict the generative space of StyleGAN so that it obeys the structure of characters yet remains flexible in handling different font styles, we store the discrete features for each character in a codebook.

Image Super-Resolution

168

Paper
Code

Non-Local Recurrent Network for Image Restoration

1 code implementation • NeurIPS 2018 • Ding Liu, Bihan Wen, Yuchen Fan, Chen Change Loy, Thomas S. Huang

The main contributions of this work are: (1) Unlike existing methods that measure self-similarity in an isolated manner, the proposed non-local module can be flexibly integrated into existing deep networks for end-to-end training to capture deep feature correlation between each location and its neighborhood.

Ranked #1 on Grayscale Image Denoising on Set12 sigma30

Feature Correlation Image Denoising +2

167

Paper
Code

Aligning Bag of Regions for Open-Vocabulary Object Detection

1 code implementation • CVPR 2023 • Size Wu, Wenwei Zhang, Sheng Jin, Wentao Liu, Chen Change Loy

The embeddings of regions in a bag are treated as embeddings of words in a sentence, and they are sent to the text encoder of a VLM to obtain the bag-of-regions embedding, which is learned to be aligned to the corresponding features extracted by a frozen VLM.

Ranked #7 on Open Vocabulary Object Detection on MSCOCO (using extra training data)

Object object-detection +2

160

Paper
Code

PERF: Panoramic Neural Radiance Field from a Single Panorama

1 code implementation • 25 Oct 2023 • Guangcong Wang, Peng Wang, Zhaoxi Chen, Wenping Wang, Chen Change Loy, Ziwei Liu

In this paper, we present PERF, a 360-degree novel view synthesis framework that trains a panoramic neural radiance field from a single panorama.

Novel View Synthesis Text to 3D

160

Paper
Code

Contextual Object Detection with Multimodal Large Language Models

1 code implementation • 29 May 2023 • Yuhang Zang, Wei Li, Jun Han, Kaiyang Zhou, Chen Change Loy

Moreover, we present ContextDET, a unified multimodal model that is capable of end-to-end differentiable modeling of visual-language contexts, so as to locate, identify, and associate visual objects with language inputs for human-AI interaction.

Cloze Test Image Captioning +6

158

Paper
Code

Face alignment by coarse-to-fine shape searching

1 code implementation • 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) 2015 • Shizhan Zhu, Cheng Li, Chen Change Loy, Xiaoou Tang

We present a novel face alignment framework based on coarse-to-fine shape searching.

Ranked #20 on Face Alignment on AFLW-19

Face Alignment regression

156

Paper
Code

Image Aesthetic Assessment: An Experimental Survey

1 code implementation • 4 Oct 2016 • Yubin Deng, Chen Change Loy, Xiaoou Tang

This survey aims at reviewing recent computer vision techniques used in the assessment of image aesthetic quality.

Binary Classification

153

Paper
Code

Video K-Net: A Simple, Strong, and Unified Baseline for Video Segmentation

1 code implementation • CVPR 2022 • Xiangtai Li, Wenwei Zhang, Jiangmiao Pang, Kai Chen, Guangliang Cheng, Yunhai Tong, Chen Change Loy

We hope this simple, yet effective method can serve as a new, flexible baseline in unified video segmentation design.

Ranked #1 on Video Panoptic Segmentation on KITTI-STEP (using extra training data)

Image Segmentation Instance Segmentation +5

150

Paper
Code

Dense Intrinsic Appearance Flow for Human Pose Transfer

1 code implementation • CVPR 2019 • Yining Li, Chen Huang, Chen Change Loy

Unlike existing methods, we propose to estimate dense and intrinsic 3D appearance flow to better guide the transfer of pixels between poses.

Pose Transfer

147

Paper
Code

A Shading-Guided Generative Implicit Model for Shape-Accurate 3D-Aware Image Synthesis

1 code implementation • NeurIPS 2021 • Xingang Pan, Xudong Xu, Chen Change Loy, Christian Theobalt, Bo Dai

Motivated by the observation that a 3D object should look realistic from multiple viewpoints, these methods introduce a multi-view constraint as regularization to learn valid 3D radiance fields from 2D images.

3D-Aware Image Synthesis 3D Shape Reconstruction +2

146

Paper
Code

Self-Supervised Learning via Conditional Motion Propagation

1 code implementation • CVPR 2019 • Xiaohang Zhan, Xingang Pan, Ziwei Liu, Dahua Lin, Chen Change Loy

Instead of explicitly modeling the motion probabilities, we design the pretext task as a conditional motion propagation problem.

Human Parsing Instance Segmentation +2

137

Paper
Code

CLIPSelf: Vision Transformer Distills Itself for Open-Vocabulary Dense Prediction

1 code implementation • 2 Oct 2023 • Size Wu, Wenwei Zhang, Lumin Xu, Sheng Jin, Xiangtai Li, Wentao Liu, Chen Change Loy

However, when transferring the vision-language alignment of CLIP from global image representation to local region representation for the open-vocabulary dense prediction tasks, CLIP ViTs suffer from the domain shift from full images to local image regions.

Ranked #3 on Open Vocabulary Semantic Segmentation on PASCAL Context-59

Image Classification Image Segmentation +7

133

Paper
Code

Delving into High-Quality Synthetic Face Occlusion Segmentation Datasets

3 code implementations • 12 May 2022 • Kenny T. R. Voo, Liming Jiang, Chen Change Loy

This paper performs comprehensive analysis on datasets for occlusion-aware face segmentation, a task that is crucial for many downstream applications.

Segmentation Synthetic Data Generation +1

121

Paper
Code

PGDiff: Guiding Diffusion Models for Versatile Face Restoration via Partial Guidance

1 code implementation • NeurIPS 2023 • Peiqing Yang, Shangchen Zhou, Qingyi Tao, Chen Change Loy

When combined with a diffusion prior, this partial guidance can deliver appealing results across a range of restoration tasks.

118

Paper
Code

StyleLight: HDR Panorama Generation for Lighting Estimation and Editing

1 code implementation • 29 Jul 2022 • Guangcong Wang, Yinuo Yang, Chen Change Loy, Ziwei Liu

To tackle this problem, we propose a coupled dual-StyleGAN panorama synthesis network (StyleLight) that integrates LDR and HDR panorama synthesis into a unified framework.

Lighting Estimation

113

Paper
Code

Inter-Region Affinity Distillation for Road Marking Segmentation

1 code implementation • CVPR 2020 • Yuenan Hou, Zheng Ma, Chunxiao Liu, Tak-Wai Hui, Chen Change Loy

We study the problem of distilling knowledge from a large deep teacher network to a much smaller student network for the task of road marking segmentation.

Ranked #1 on Semantic Segmentation on ApolloScape

Knowledge Distillation Lane Detection +1

112

Paper
Code

Flare7K: A Phenomenological Nighttime Flare Removal Dataset

1 code implementation • 12 Oct 2022 • Yuekun Dai, Chongyi Li, Shangchen Zhou, Ruicheng Feng, Chen Change Loy

In this paper, we introduce, Flare7K, the first nighttime flare removal dataset, which is generated based on the observation and statistics of real-world nighttime lens flares.

Ranked #2 on Flare Removal on Flare7K

Flare Removal

110

Paper
Code

Flare7K++: Mixing Synthetic and Real Datasets for Nighttime Flare Removal and Beyond

1 code implementation • 7 Jun 2023 • Yuekun Dai, Chongyi Li, Shangchen Zhou, Ruicheng Feng, Yihang Luo, Chen Change Loy

To address this issue, we additionally provide the annotations of light sources in Flare7K++ and propose a new end-to-end pipeline to preserve the light source while removing lens flares.

Flare Removal

110

Paper
Code

Not All Pixels Are Equal: Difficulty-aware Semantic Segmentation via Deep Layer Cascade

1 code implementation • CVPR 2017 • Xiaoxiao Li, Ziwei Liu, Ping Luo, Chen Change Loy, Xiaoou Tang

Third, in comparison to MC, LC is an end-to-end trainable framework, allowing joint learning of all sub-models.

Ranked #22 on Semantic Segmentation on PASCAL VOC 2012 test

Semantic Segmentation

108

Paper
Code

Tube-Link: A Flexible Cross Tube Framework for Universal Video Segmentation

2 code implementations • ICCV 2023 • Xiangtai Li, Haobo Yuan, Wenwei Zhang, Guangliang Cheng, Jiangmiao Pang, Chen Change Loy

Our framework is a near-online approach that takes a short subclip as input and outputs the corresponding spatial-temporal tube masks.

Ranked #3 on Video Semantic Segmentation on VSPW

Contrastive Learning Segmentation +4

105

Paper
Code

MosaicFusion: Diffusion Models as Data Augmenters for Large Vocabulary Instance Segmentation

1 code implementation • 22 Sep 2023 • Jiahao Xie, Wei Li, Xiangtai Li, Ziwei Liu, Yew Soon Ong, Chen Change Loy

We present MosaicFusion, a simple yet effective diffusion-based data augmentation approach for large vocabulary instance segmentation.

Data Augmentation Instance Segmentation +1

105

Paper
Code

Removing Diffraction Image Artifacts in Under-Display Camera via Dynamic Skip Connection Network

1 code implementation • CVPR 2021 • Ruicheng Feng, Chongyi Li, Huaijin Chen, Shuai Li, Chen Change Loy, Jinwei Gu

Recent development of Under-Display Camera (UDC) systems provides a true bezel-less and notch-free viewing experience on smartphones (and TV, laptops, tablets), while allowing images to be captured from the selfie camera embedded underneath.

Image Restoration

Paper
Code

Everything's Talkin': Pareidolia Face Reenactment

1 code implementation • 7 Apr 2021 • Linsen Song, Wayne Wu, Chaoyou Fu, Chen Qian, Chen Change Loy, Ran He

We present a new application direction named Pareidolia Face Reenactment, which is defined as animating a static illusory face to move in tandem with a human face in the video.

Face Reenactment Texture Synthesis

Paper
Code

AnimeRun: 2D Animation Visual Correspondence from Open Source 3D Movies

1 code implementation • 10 Nov 2022 • Li SiYao, Yuhang Li, Bo Li, Chao Dong, Ziwei Liu, Chen Change Loy

Existing correspondence datasets for two-dimensional (2D) cartoon suffer from simple frame composition and monotonic movements, making them insufficient to simulate real animations.

Optical Flow Estimation

Paper
Code

When StyleGAN Meets Stable Diffusion: a $\mathscr{W}_+$ Adapter for Personalized Image Generation

1 code implementation • 29 Nov 2023 • Xiaoming Li, Xinyu Hou, Chen Change Loy

Text-to-image diffusion models have remarkably excelled in producing diverse, high-quality, and photo-realistic images.

Attribute Disentanglement +1

Paper
Code

Position-Guided Point Cloud Panoptic Segmentation Transformer

1 code implementation • 23 Mar 2023 • Zeqi Xiao, Wenwei Zhang, Tai Wang, Chen Change Loy, Dahua Lin, Jiangmiao Pang

DEtection TRansformer (DETR) started a trend that uses a group of learnable queries for unified visual perception.

Ranked #1 on Panoptic Segmentation on SemanticKITTI

Instance Segmentation Panoptic Segmentation +3

Paper
Code

Delving Deep Into Hybrid Annotations for 3D Human Recovery in the Wild

1 code implementation • ICCV 2019 • Yu Rong, Ziwei Liu, Cheng Li, Kaidi Cao, Chen Change Loy

Specifically, we focus on the challenging task of in-the-wild 3D human recovery from single images when paired 3D annotations are not fully available.

Paper
Code

NTIRE 2021 Challenge on Quality Enhancement of Compressed Video: Methods and Results

1 code implementation • 21 Apr 2021 • Ren Yang, Radu Timofte, Jing Liu, Yi Xu, Xinjian Zhang, Minyi Zhao, Shuigeng Zhou, Kelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change Loy, Xin Li, Fanglong Liu, He Zheng, Lielin Jiang, Qi Zhang, Dongliang He, Fu Li, Qingqing Dang, Yibin Huang, Matteo Maggioni, Zhongqian Fu, Shuai Xiao, Cheng Li, Thomas Tanay, Fenglong Song, Wentao Chao, Qiang Guo, Yan Liu, Jiang Li, Xiaochao Qu, Dewang Hou, Jiayu Yang, Lyn Jiang, Di You, Zhenyu Zhang, Chong Mou, Iaroslav Koshelev, Pavel Ostyakov, Andrey Somov, Jia Hao, Xueyi Zou, Shijie Zhao, Xiaopeng Sun, Yiting Liao, Yuanzhi Zhang, Qing Wang, Gen Zhan, Mengxi Guo, Junlin Li, Ming Lu, Zhan Ma, Pablo Navarrete Michelini, Hai Wang, Yiyun Chen, Jingyu Guo, Liliang Zhang, Wenming Yang, Sijung Kim, Syehoon Oh, Yucong Wang, Minjie Cai, Wei Hao, Kangdi Shi, Liangyan Li, Jun Chen, Wei Gao, Wang Liu, XiaoYu Zhang, Linjie Zhou, Sixin Lin, Ru Wang

This paper reviews the first NTIRE challenge on quality enhancement of compressed video, with a focus on the proposed methods and results.

Paper
Code

BRACE: The Breakdancing Competition Dataset for Dance Motion Synthesis

1 code implementation • 20 Jul 2022 • Davide Moltisanti, Jinyi Wu, Bo Dai, Chen Change Loy

Estimating human keypoints from these videos is difficult due to the complexity of the dance, as well as the multiple moving cameras recording setup.

Motion Synthesis Pose Estimation

Paper
Code

Nighttime Smartphone Reflective Flare Removal Using Optical Center Symmetry Prior

1 code implementation • CVPR 2023 • Yuekun Dai, Yihang Luo, Shangchen Zhou, Chongyi Li, Chen Change Loy

With the dataset, neural networks can be trained to remove the reflective flares effectively.

Flare Removal

Paper
Code

Masked Frequency Modeling for Self-Supervised Visual Pre-Training

3 code implementations • 15 Jun 2022 • Jiahao Xie, Wei Li, Xiaohang Zhan, Ziwei Liu, Yew Soon Ong, Chen Change Loy

We present Masked Frequency Modeling (MFM), a unified frequency-domain-based approach for self-supervised pre-training of visual models.

Image Classification Image Restoration +2

Paper
Code

Unsupervised Object-Level Representation Learning from Scene Images

1 code implementation • NeurIPS 2021 • Jiahao Xie, Xiaohang Zhan, Ziwei Liu, Yew Soon Ong, Chen Change Loy

Extensive experiments on COCO show that ORL significantly improves the performance of self-supervised learning on scene images, even surpassing supervised ImageNet pre-training on several downstream tasks.

Object Representation Learning +2

Paper
Code

Unified Vision and Language Prompt Learning

1 code implementation • 13 Oct 2022 • Yuhang Zang, Wei Li, Kaiyang Zhou, Chen Huang, Chen Change Loy

Prompt tuning, a parameter- and data-efficient transfer learning paradigm that tunes only a small number of parameters in a model's input space, has become a trend in the vision community since the emergence of large vision-language models like CLIP.

Domain Generalization Few-Shot Learning +2

Paper
Code

Monocular 3D Object Reconstruction with GAN Inversion

1 code implementation • 20 Jul 2022 • Junzhe Zhang, Daxuan Ren, Zhongang Cai, Chai Kiat Yeo, Bo Dai, Chen Change Loy

Reconstruction is achieved by searching for a latent space in the 3D GAN that best resembles the target mesh in accordance with the single view observation.

3D Object Reconstruction Object

Paper
Code

Explore In-Context Learning for 3D Point Cloud Understanding

2 code implementations • NeurIPS 2023 • Zhongbin Fang, Xiangtai Li, Xia Li, Joachim M. Buhmann, Chen Change Loy, Mengyuan Liu

With the rise of large-scale models trained on broad data, in-context learning has become a new learning paradigm that has demonstrated significant potential in natural language processing and computer vision tasks.

In-Context Learning

Paper
Code

Point-In-Context: Understanding Point Cloud via In-Context Learning

1 code implementation • 18 Apr 2024 • Mengyuan Liu, Zhongbin Fang, Xia Li, Joachim M. Buhmann, Xiangtai Li, Chen Change Loy

With the emergence of large-scale models trained on diverse datasets, in-context learning has emerged as a promising paradigm for multitasking, notably in natural language processing and image processing.

In-Context Learning

Paper
Code

Learning to Steer by Mimicking Features from Heterogeneous Auxiliary Networks

2 code implementations • 7 Nov 2018 • Yuenan Hou, Zheng Ma, Chunxiao Liu, Chen Change Loy

In this paper, we considerably improve the accuracy and robustness of predictions through heterogeneous auxiliary networks feature mimicking, a new and effective training method that provides us with much richer contextual signals apart from steering direction.

Ranked #1 on Steering Control on BDD100K val

Image Segmentation Multi-Task Learning +3

Paper
Code

Panoptic Video Scene Graph Generation

3 code implementations • CVPR 2023 • Jingkang Yang, Wenxuan Peng, Xiangtai Li, Zujin Guo, Liangyu Chen, Bo Li, Zheng Ma, Kaiyang Zhou, Wayne Zhang, Chen Change Loy, Ziwei Liu

PVSG relates to the existing video scene graph generation (VidSGG) problem, which focuses on temporal interactions between humans and objects grounded with bounding boxes in videos.

Graph Generation Panoptic Scene Graph Generation +5

Paper
Code

Robust and Fast Decoding of High-Capacity Color QR Codes for Mobile Applications

1 code implementation • 21 Apr 2017 • Zhibo Yang, Huanle Xu, Jianyuan Deng, Chen Change Loy, Wing Cheong Lau

Particularly, we further discover a new type of chromatic distortion in high-density color QR codes, cross-module color interference, caused by the high density which also makes the geometric distortion correction more challenging.

Paper
Code

DeformToon3D: Deformable 3D Toonification from Neural Radiance Fields

1 code implementation • 8 Sep 2023 • Junzhe Zhang, Yushi Lan, Shuai Yang, Fangzhou Hong, Quan Wang, Chai Kiat Yeo, Ziwei Liu, Chen Change Loy

In this paper, we address the challenging problem of 3D toonification, which involves transferring the style of an artistic domain onto a target 3D face with stylized geometry and texture.

Paper
Code

Aesthetic-Driven Image Enhancement by Adversarial Learning

1 code implementation • 17 Jul 2017 • Yubin Deng, Chen Change Loy, Xiaoou Tang

We introduce EnhanceGAN, an adversarial learning based model that performs automatic image enhancement.

Image Cropping Image Enhancement

Paper
Code

Compression Artifacts Reduction by a Deep Convolutional Network

4 code implementations • ICCV 2015 • Chao Dong, Yubin Deng, Chen Change Loy, Xiaoou Tang

Lossy compression introduces complex compression artifacts, particularly the blocking artifacts, ringing effects and blurring.

Ranked #4 on JPEG Artifact Correction on ICB (Quality 20 Grayscale)

Blocking Denoising +2

Paper
Code

Generating Aligned Pseudo-Supervision from Non-Aligned Data for Image Restoration in Under-Display Camera

1 code implementation • CVPR 2023 • Ruicheng Feng, Chongyi Li, Huaijin Chen, Shuai Li, Jinwei Gu, Chen Change Loy

Due to the difficulty in collecting large-scale and perfectly aligned paired training data for Under-Display Camera (UDC) image restoration, previous methods resort to monitor-based image systems or simulation-based methods, sacrificing the realness of the data and introducing domain gaps.

Image Restoration

Paper
Code

Betrayed by Captions: Joint Caption Grounding and Generation for Open Vocabulary Instance Segmentation

2 code implementations • ICCV 2023 • Jianzong Wu, Xiangtai Li, Henghui Ding, Xia Li, Guangliang Cheng, Yunhai Tong, Chen Change Loy

Experiments on the COCO dataset with two settings: Open Vocabulary Instance Segmentation (OVIS) and Open Set Panoptic Segmentation (OSPS) demonstrate the superiority of the CGG.

Caption Generation Instance Segmentation +2

Paper
Code

Path-Restore: Learning Network Path Selection for Image Restoration

1 code implementation • 23 Apr 2019 • Ke Yu, Xintao Wang, Chao Dong, Xiaoou Tang, Chen Change Loy

To leverage this, we propose Path-Restore, a multi-path CNN with a pathfinder that can dynamically select an appropriate route for each image region.

Denoising Image Restoration +1

Paper
Code

MIPI 2022 Challenge on Quad-Bayer Re-mosaic: Dataset and Report

1 code implementation • 15 Sep 2022 • Qingyu Yang, Guang Yang, Jun Jiang, Chongyi Li, Ruicheng Feng, Shangchen Zhou, Wenxiu Sun, Qingpeng Zhu, Chen Change Loy, Jinwei Gu

A detailed description of all models developed in this challenge is provided in this paper.

SSIM

Paper
Code

MIPI 2022 Challenge on RGB+ToF Depth Completion: Dataset and Report

1 code implementation • 15 Sep 2022 • Wenxiu Sun, Qingpeng Zhu, Chongyi Li, Ruicheng Feng, Shangchen Zhou, Jun Jiang, Qingyu Yang, Chen Change Loy, Jinwei Gu

A detailed description of all models developed in this challenge is provided in this paper.

Depth Completion

Paper
Code

MIPI 2022 Challenge on Under-Display Camera Image Restoration: Methods and Results

1 code implementation • 15 Sep 2022 • Ruicheng Feng, Chongyi Li, Shangchen Zhou, Wenxiu Sun, Qingpeng Zhu, Jun Jiang, Qingyu Yang, Chen Change Loy, Jinwei Gu

In this paper, we summarize and review the Under-Display Camera (UDC) Image Restoration track on MIPI 2022.

Image Restoration

Paper
Code

MIPI 2022 Challenge on RGBW Sensor Fusion: Dataset and Report

1 code implementation • 15 Sep 2022 • Qingyu Yang, Guang Yang, Jun Jiang, Chongyi Li, Ruicheng Feng, Shangchen Zhou, Wenxiu Sun, Qingpeng Zhu, Chen Change Loy, Jinwei Gu

A detailed description of all models developed in this challenge is provided in this paper.

Sensor Fusion SSIM

Paper
Code

MIPI 2022 Challenge on RGBW Sensor Re-mosaic: Dataset and Report

1 code implementation • 15 Sep 2022 • Qingyu Yang, Guang Yang, Jun Jiang, Chongyi Li, Ruicheng Feng, Shangchen Zhou, Wenxiu Sun, Qingpeng Zhu, Chen Change Loy, Jinwei Gu

A detailed description of all models developed in this challenge is provided in this paper.

SSIM

Paper
Code

Temporally Consistent Video Colorization with Deep Feature Propagation and Self-regularization Learning

1 code implementation • 9 Oct 2021 • Yihao Liu, Hengyuan Zhao, Kelvin C. K. Chan, Xintao Wang, Chen Change Loy, Yu Qiao, Chao Dong

We address this problem from a new perspective, by jointly considering colorization and temporal consistency in a unified framework.

Colorization Image Colorization

Paper
Code

Siamese DETR

1 code implementation • CVPR 2023 • Zeren Chen, Gengshi Huang, Wei Li, Jianing Teng, Kun Wang, Jing Shao, Chen Change Loy, Lu Sheng

In this work, we present Siamese DETR, a Siamese self-supervised pretraining approach for the Transformer architecture in DETR.

MULTI-VIEW LEARNING Representation Learning

Paper
Code

Quantifying Facial Age by Posterior of Age Comparisons

1 code implementation • 31 Aug 2017 • Yunxuan Zhang, Li Liu, Cheng Li, Chen Change Loy

We introduce a novel approach for annotating large quantity of in-the-wild facial images with high-quality posterior age distribution as labels.

Ranked #7 on Age Estimation on MORPH Album2 (using extra training data)

Age And Gender Classification Age Estimation

Paper
Code

FASA: Feature Augmentation and Sampling Adaptation for Long-Tailed Instance Segmentation

1 code implementation • ICCV 2021 • Yuhang Zang, Chen Huang, Chen Change Loy

We propose a simple yet effective method, Feature Augmentation and Sampling Adaptation (FASA), that addresses the data scarcity issue by augmenting the feature space especially for rare classes.

Instance Segmentation Segmentation +2

Paper
Code

Dense Siamese Network for Dense Unsupervised Learning

1 code implementation • 21 Mar 2022 • Wenwei Zhang, Jiangmiao Pang, Kai Chen, Chen Change Loy

It also extracts a batch of region embeddings that correspond to some sub-regions in the overlapped area to be contrasted for region consistency.

Ranked #2 on Unsupervised Semantic Segmentation on COCO-All (mIoU metric)

Self-Supervised Learning Unsupervised Semantic Segmentation

Paper
Code

Transformer with Implicit Edges for Particle-based Physics Simulation

1 code implementation • 22 Jul 2022 • Yidi Shao, Chen Change Loy, Bo Dai

Consequently, in this paper we propose a novel Transformer-based method, dubbed as Transformer with Implicit Edges (TIE), to capture the rich semantics of particle interactions in an edge-free manner.

Paper
Code

Mind the Gap in Distilling StyleGANs

1 code implementation • 18 Aug 2022 • Guodong Xu, Yuenan Hou, Ziwei Liu, Chen Change Loy

To further enhance the semantic consistency between the teacher and student model, we present a latent-direction-based distillation loss that preserves the semantic relations in latent space.

Knowledge Distillation

Paper
Code

Correlational Image Modeling for Self-Supervised Visual Pre-Training

1 code implementation • CVPR 2023 • Wei Li, Jiahao Xie, Chen Change Loy

We introduce Correlational Image Modeling (CIM), a novel and surprisingly effective approach to self-supervised visual pre-training.

Paper
Code

Computation-Efficient Knowledge Distillation via Uncertainty-Aware Mixup

1 code implementation • 17 Dec 2020 • Guodong Xu, Ziwei Liu, Chen Change Loy

Our goal is to achieve a performance comparable to conventional knowledge distillation with a lower computation cost during training.

Informativeness Knowledge Distillation +2

Paper
Code

ReconfigISP: Reconfigurable Camera Image Processing Pipeline

1 code implementation • ICCV 2021 • Ke Yu, Zexian Li, Yue Peng, Chen Change Loy, Jinwei Gu

Image Signal Processor (ISP) is a crucial component in digital cameras that transforms sensor signals into images for us to perceive and understand.

Image Restoration Neural Architecture Search +2

Paper
Code

StyleInV: A Temporal Style Modulated Inversion Network for Unconditional Video Generation

1 code implementation • ICCV 2023 • YuHan Wang, Liming Jiang, Chen Change Loy

In this paper, we introduce a novel motion generator design that uses a learning-based inversion network for GAN.

Style Transfer Unconditional Video Generation

Paper
Code

DST-Det: Simple Dynamic Self-Training for Open-Vocabulary Object Detection

1 code implementation • 2 Oct 2023 • Shilin Xu, Xiangtai Li, Size Wu, Wenwei Zhang, Yunhai Tong, Chen Change Loy

We refer to this approach as the self-training strategy, which enhances recall and accuracy for novel classes without requiring extra annotations, datasets, and re-training.

Novel Object Detection Object +5

Paper
Code

Merge or Not? Learning to Group Faces via Imitation Learning

1 code implementation • 13 Jul 2017 • Yue He, Kaidi Cao, Cheng Li, Chen Change Loy

Given a large number of unlabeled face images, face grouping aims at clustering the images into individual identities present in the data.

Clustering Imitation Learning

Paper
Code

CLIM: Contrastive Language-Image Mosaic for Region Representation

1 code implementation • 18 Dec 2023 • Size Wu, Wenwei Zhang, Lumin Xu, Sheng Jin, Wentao Liu, Chen Change Loy

Our experimental results demonstrate that CLIM improves different baseline open-vocabulary object detectors by a large margin on both OV-COCO and OV-LVIS benchmarks.

Ranked #6 on Open Vocabulary Object Detection on LVIS v1.0

Object object-detection +1

Paper
Code

RGB-D Salient Object Detection with Cross-Modality Modulation and Selection

1 code implementation • ECCV 2020 • Chongyi Li, Runmin Cong, Yongri Piao, Qianqian Xu, Chen Change Loy

Second, we propose an adaptive feature selection (AFS) module to select saliency-related features and suppress the inferior ones.

Ranked #8 on RGB-D Salient Object Detection on NJU2K

feature selection object-detection +3

Paper
Code

Rethinking CLIP-based Video Learners in Cross-Domain Open-Vocabulary Action Recognition

1 code implementation • 3 Mar 2024 • Kun-Yu Lin, Henghui Ding, Jiaming Zhou, Yi-Xing Peng, Zhilin Zhao, Chen Change Loy, Wei-Shi Zheng

To answer this, we establish a CROSS-domain Open-Vocabulary Action recognition benchmark named XOV-Action, and conduct a comprehensive evaluation of five state-of-the-art CLIP-based video learners under various types of domain gaps.

Open Vocabulary Action Recognition

Paper
Code

A Large-Scale Car Dataset for Fine-Grained Categorization and Verification

3 code implementations • CVPR 2015 • Linjie Yang, Ping Luo, Chen Change Loy, Xiaoou Tang

Updated on 24/09/2015: This update provides preliminary experiment results for fine-grained classification on the surveillance data of CompCars.

Ranked #5 on Fine-Grained Image Classification on CompCars

Fine-Grained Image Classification General Classification

Paper
Code

Deep Convolution Networks for Compression Artifacts Reduction

2 code implementations • 9 Aug 2016 • Ke Yu, Chao Dong, Chen Change Loy, Xiaoou Tang

Lossy compression introduces complex compression artifacts, particularly blocking artifacts, ringing effects and blurring.

Blocking Transfer Learning

Paper
Code

Disentangling Content and Style via Unsupervised Geometry Distillation

1 code implementation • ICLR Workshop DeepGenStruct 2019 • Wayne Wu, Kaidi Cao, Cheng Li, Chen Qian, Chen Change Loy

It is challenging to disentangle an object into two orthogonal spaces of content and style since each can influence the visual observation differently and unpredictably.

Disentanglement

Paper
Code

Discover and Learn New Objects from Documentaries

1 code implementation • CVPR 2017 • Kai Chen, Hang Song, Chen Change Loy, Dahua Lin

Despite the remarkable progress in recent years, detecting objects in a new context remains a challenging task.

Object Weakly-supervised Learning

Paper
Code

An Embarrassingly Simple Approach for Knowledge Distillation

1 code implementation • 5 Dec 2018 • Mengya Gao, Yujun Shen, Quanquan Li, Junjie Yan, Liang Wan, Dahua Lin, Chen Change Loy, Xiaoou Tang

Knowledge Distillation (KD) aims at improving the performance of a low-capacity student model by inheriting knowledge from a high-capacity teacher model.

Face Recognition Knowledge Distillation +3

Paper
Code

Deep Imbalanced Learning for Face Recognition and Attribute Prediction

1 code implementation • 1 Jun 2018 • Chen Huang, Yining Li, Chen Change Loy, Xiaoou Tang

Data for face analysis often exhibit highly-skewed class distribution, i. e., most data belong to a few majority classes, while the minority classes only contain a scarce amount of instances.

Attribute Face Recognition +1

Paper
Code

WIDER FACE: A Face Detection Benchmark

1 code implementation • CVPR 2016 • Shuo Yang, Ping Luo, Chen Change Loy, Xiaoou Tang

Face detection is one of the most studied topics in the computer vision community.

Ranked #34 on Face Detection on WIDER Face (Medium)

Face Detection

Paper
Code

Deep Network Interpolation for Continuous Imagery Effect Transition

2 code implementations • CVPR 2019 • Xintao Wang, Ke Yu, Chao Dong, Xiaoou Tang, Chen Change Loy

Deep convolutional neural network has demonstrated its capability of learning a deterministic mapping for the desired imagery effect.

Image Restoration Image-to-Image Translation +2

Paper
Code

Deep Fourier Up-Sampling

1 code implementation • 11 Oct 2022 • Man Zhou, Hu Yu, Jie Huang, Feng Zhao, Jinwei Gu, Chen Change Loy, Deyu Meng, Chongyi Li

Existing convolutional neural networks widely adopt spatial down-/up-sampling for multi-scale modeling.

Image Dehazing Image Segmentation +4

Paper
Code

Video Object Segmentation with Joint Re-identification and Attention-Aware Mask Propagation

no code implementations • ECCV 2018 • Xiaoxiao Li, Chen Change Loy

The problem of video object segmentation can become extremely challenging when multiple instances co-exist.

Semantic Segmentation Video Object Segmentation +1

Paper
Add Code

Mix-and-Match Tuning for Self-Supervised Semantic Segmentation

no code implementations • 2 Dec 2017 • Xiaohang Zhan, Ziwei Liu, Ping Luo, Xiaoou Tang, Chen Change Loy

The key of this new form of learning is to design a proxy task (e. g. image colorization), from which a discriminative loss can be formulated on unlabeled data.

Colorization Image Colorization +3

Paper
Add Code

From Facial Expression Recognition to Interpersonal Relation Prediction

no code implementations • 21 Sep 2016 • Zhanpeng Zhang, Ping Luo, Chen Change Loy, Xiaoou Tang

Unlike existing models that typically learn from facial expression labels alone, we devise an effective multitask network that is capable of learning from rich auxiliary attributes such as gender, age, and head pose, beyond just facial expression data.

Attribute Facial Expression Recognition +2

Paper
Add Code

Be Your Own Prada: Fashion Synthesis with Structural Coherence

no code implementations • ICCV 2017 • Shizhan Zhu, Sanja Fidler, Raquel Urtasun, Dahua Lin, Chen Change Loy

In the second stage, a generative model with a newly proposed compositional mapping layer is used to render the final image with precise regions and textures conditioned on this map.

Fashion Synthesis Semantic Segmentation +1

Paper
Add Code

Faceness-Net: Face Detection through Deep Facial Part Responses

no code implementations • 29 Jan 2017 • Shuo Yang, Ping Luo, Chen Change Loy, Xiaoou Tang

We propose a deep convolutional neural network (CNN) for face detection leveraging on facial attributes based supervision.

Face Detection

Paper
Add Code

Deep Learning Markov Random Field for Semantic Segmentation

no code implementations • 23 Jun 2016 • Ziwei Liu, Xiaoxiao Li, Ping Luo, Chen Change Loy, Xiaoou Tang

Semantic segmentation tasks can be well modeled by Markov Random Field (MRF).

Segmentation Semantic Segmentation +2

Paper
Add Code

Face Detection through Scale-Friendly Deep Convolutional Networks

no code implementations • 9 Jun 2017 • Shuo Yang, Yuanjun Xiong, Chen Change Loy, Xiaoou Tang

Specifically, our method achieves 76. 4 average precision on the challenging WIDER FACE dataset and 96% recall rate on the FDDB dataset with 7 frames per second (fps) for 900 * 1300 input image.

Face Detection

Paper
Add Code

Local Similarity-Aware Deep Feature Embedding

no code implementations • NeurIPS 2016 • Chen Huang, Chen Change Loy, Xiaoou Tang

Existing deep embedding methods in vision tasks are capable of learning a compact Euclidean space from images, where Euclidean distances correspond to a similarity metric.

Ranked #27 on Metric Learning on CUB-200-2011

Image Retrieval Retrieval +2

Paper
Add Code

Deep Cascaded Bi-Network for Face Hallucination

no code implementations • 18 Jul 2016 • Shizhan Zhu, Sifei Liu, Chen Change Loy, Xiaoou Tang

We present a novel framework for hallucinating faces of unconstrained poses and with very low resolution (face size as small as 5pxIOD).

Ranked #5 on Image Super-Resolution on VggFace2 - 8x upscaling

Face Hallucination Hallucination

Paper
Add Code

Discriminative Sparse Neighbor Approximation for Imbalanced Learning

no code implementations • 3 Feb 2016 • Chen Huang, Chen Change Loy, Xiaoou Tang

These methods further deteriorate on small, imbalanced data that has a large degree of class overlap.

Paper
Add Code

Reading Scene Text in Deep Convolutional Sequences

1 code implementation • 14 Jun 2015 • Pan He, Weilin Huang, Yu Qiao, Chen Change Loy, Xiaoou Tang

We develop a Deep-Text Recurrent Network (DTRN) that regards scene text reading as a sequence labelling problem.

Scene Text Recognition

Paper
Code

Towards Arbitrary-View Face Alignment by Recommendation Trees

no code implementations • 20 Nov 2015 • Shizhan Zhu, Cheng Li, Chen Change Loy, Xiaoou Tang

The unified framework seamlessly handles different viewpoints and landmark protocols, and it is trained by optimising directly on landmark locations, thus yielding superior results on arbitrary-view face alignment.

Face Alignment Head Pose Estimation +1

Paper
Add Code

An Empirical Study of Recent Face Alignment Methods

no code implementations • 16 Nov 2015 • Heng Yang, Xuhui Jia, Chen Change Loy, Peter Robinson

In this paper, we carry out a rigorous evaluation of these methods by making the following contributions: 1) we proposes a new evaluation metric for face alignment on a set of images, i. e., area under error distribution curve within a threshold, AUC$_\alpha$, given the fact that the traditional evaluation measure (mean error) is very sensitive to big alignment error.

Face Alignment Face Detection

Paper
Add Code

Semantic Image Segmentation via Deep Parsing Network

no code implementations • ICCV 2015 • Ziwei Liu, Xiaoxiao Li, Ping Luo, Chen Change Loy, Xiaoou Tang

This paper addresses semantic image segmentation by incorporating rich information into Markov Random Field (MRF), including high-order relations and mixture of label contexts.

Ranked #89 on Semantic Segmentation on Cityscapes test

Image Segmentation Semantic Segmentation

Paper
Add Code

From Facial Parts Responses to Face Detection: A Deep Learning Approach

1 code implementation • ICCV 2015 • Shuo Yang, Ping Luo, Chen Change Loy, Xiaoou Tang

In this paper, we propose a novel deep convolutional network (DCN) that achieves outstanding performance on FDDB, PASCAL Face, and AFW.

Face Detection

Paper
Code

Learning Social Relation Traits from Face Images

no code implementations • ICCV 2015 • Zhanpeng Zhang, Ping Luo, Chen Change Loy, Xiaoou Tang

Social relation defines the association, e. g, warm, friendliness, and dominance, between two or more people.

Attribute Relation

Paper
Add Code

Learning Deep Representation for Face Alignment with Auxiliary Attributes

no code implementations • 18 Aug 2014 • Zhanpeng Zhang, Ping Luo, Chen Change Loy, Xiaoou Tang

In this study, we show that landmark detection or face alignment task is not a single and independent problem.

Ranked #13 on Unsupervised Facial Landmark Detection on MAFL

Attribute Face Alignment

Paper
Add Code

Boosting Optical Character Recognition: A Super-Resolution Approach

no code implementations • 7 Jun 2015 • Chao Dong, Ximei Zhu, Yubin Deng, Chen Change Loy, Yu Qiao

Text image super-resolution is a challenging yet open research problem in the computer vision community.

Image Super-Resolution Optical Character Recognition +1

Paper
Add Code

Learning to Recognize Pedestrian Attribute

no code implementations • 5 Jan 2015 • Yubin Deng, Ping Luo, Chen Change Loy, Xiaoou Tang

Learning to recognize pedestrian attributes at far distance is a challenging problem in visual surveillance since face and body close-shots are hardly available; instead, only far-view image frames of pedestrian are given.

Attribute Informativeness

Paper
Add Code

Learning from Multiple Sources for Video Summarisation

no code implementations • 13 Jan 2015 • Xiatian Zhu, Chen Change Loy, Shaogang Gong

Many visual surveillance tasks, e. g. video summarisation, is conventionally accomplished through analysing imagerybased features.

Clustering Video Understanding

Paper
Add Code

Crowd Saliency Detection via Global Similarity Structure

no code implementations • 14 Oct 2014 • Mei Kuan Lim, Ven Jyn Kok, Chen Change Loy, Chee Seng Chan

This paper proposes a novel framework to identify and localize salient regions in a crowd scene, by transforming low-level features extracted from crowd motion field into a global similarity structure.

Saliency Detection

Paper
Add Code

Transferring Landmark Annotations for Cross-Dataset Face Alignment

no code implementations • 2 Sep 2014 • Shizhan Zhu, Cheng Li, Chen Change Loy, Xiaoou Tang

We show extensive results on combining various popular databases (LFW, AFLW, LFPW, HELEN) for improved cross-dataset and unseen data alignment.

Face Alignment Object Recognition

Paper
Add Code

Cannot find the paper you are looking for? You can Submit a new open access paper.