Is this you? Claim this profile to correct it, add a bio and choose the work people see first.

Claim this profile

Academic lineage

Possible advisorsa guess from early papers, not confirmed

  • Chen Change Loy

    Possible advisor · last author on 7 of their early first-author papers, 2021–2022

    Suggested from co-authorship

Is this you? Claim this profile to confirm or dismiss it.

Works41 from public data

TitleCited by
  • EDVR: Video Restoration With Enhanced Deformable Convolutional Networks

    Xintao Wang, Kelvin C. K. Chan, K. Yu, Chao Dong, Chen Change Loy

    IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) · 2019

    This work proposes a novel Video Restoration framework with Enhanced Deformable convolutions, termed EDVR, and proposes a Temporal and Spatial Attention (TSA) fusion module, in which attention is applied both temporally and spatially, so as to emphasize important features for subsequent restoration.

    1,341
  • Exploiting Diffusion Prior for Real-World Image Super-Resolution

    Jianyi Wang, Zongsheng Yue, Shang-Chen Zhou, Kelvin C. K. Chan, Chen Change Loy

    International Journal of Computer Vision · 2024

    A novel approach to leverage prior knowledge encapsulated in pre-trained text-to-image diffusion models for blind super-resolution by employing the time-aware encoder can achieve promising restoration results without altering the pre-trained synthesis model, thereby preserving the generative prior and minimizing training cost.

    715
  • Exploring CLIP for Assessing the Look and Feel of Images

    Proceedings of the AAAI Conference on Artificial Intelligence · 2023

    709
  • BasicVSR: The Search for Essential Components in Video Super-Resolution and Beyond

    Kelvin C. K. Chan, Xintao Wang, Ke Yu, Chao Dong, Chen Change Loy

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2021

    A succinct pipeline is shown that achieves appealing improvements in terms of speed and restoration quality in comparison to many state-of-the-art algorithms and can serve as strong baselines for future VSR approaches.

    698
  • BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and Alignment

    Kelvin C. K. Chan, Shang-Chen Zhou, Xiangyu Xu, Chen Change Loy

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2022

    This study redesigns BasicVsr by proposing second-order grid propagation and flow-guided deformable alignment, and shows that by empowering the re-current framework with enhanced propagation and align-ment, one can exploit spatiotemporal information across misaligned video frames more effectively.

    687
  • GLEAN: Generative Latent Bank for Large-Factor Image Super-Resolution

    Kelvin C. K. Chan, Xintao Wang, Xiangyu Xu, Jinwei Gu, Chen Change Loy

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2021

    This work shows that pre-trained Generative Adversarial Networks (GANs), e.g., StyleGAN, can be used as a latent bank to improve the restoration quality of large-factor image super-resolution (SR) and shows clear improvements in terms of fidelity and texture faithfulness in comparison to existing methods.

    305
  • ProPainter: Improving Propagation and Transformer for Video Inpainting

    Shang-Chen Zhou, Chongyi Li, Kelvin C. K. Chan, Chan-Ge Chen

    IEEE/CVF International Conference on Computer Vision (ICCV) · 2023

    This work introduces dual-domain propagation that combines the advantages of image and feature warping, exploiting global correspondences reliably, and proposes a mask-guided sparse video Transformer, which achieves high efficiency by discarding unnecessary and redundant tokens.

    267
  • Understanding Deformable Alignment in Video Super-Resolution

    Kelvin C. K. Chan, Xintao Wang, K. Yu, Chao Dong, Chen Change Loy

    Proceedings of the AAAI Conference on Artificial Intelligence · 2021

    It is shown that deformable convolution can be decomposed into a combination of spatial warping and convolution, which reveals the commonality of deformable alignment and flow-based alignment in formulation, but with a key difference in their offset diversity.

    195
  • Collaborative Diffusion for Multi-Modal Face Generation and Editing

    Ziqi Huang, Kelvin C. K. Chan, Yu-Ming Jiang, Ziwei Liu

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2023

    191
  • Investigating Tradeoffs in Real-World Video Super-Resolution

    Kelvin C. K. Chan, Shang-Chen Zhou, Xiang-Yu Xu, Chen Change Loy

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2022

    A stochastic degradation scheme that reduces up to 40% of training time without sacrificing performance is proposed and it is suggested that employing longer sequences rather than larger batches during training allows more effective uses of temporal information, leading to more stable performance during inference.

    185
  • Robust Reference-based Super-Resolution via C 2 -Matching

    Yu-Ming Jiang, Kelvin C. K. Chan, Xin-Tao Wang, Chen Change Loy, Zi-Wei Liu

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2021

    The proposed C2-Matching significantly outperforms state of the arts by over 1dB on the standard CUFED5 benchmark and shows great generalizability on WR-SR dataset as well as robustness across large scale and rotation transformations.

    149
  • Taming Encoder for Zero Fine-tuning Image Customization with Text-to-Image Diffusion Models

    Xu-Hui Jia, Yang Zhao, Kelvin C. K. Chan, Yan-Dong Li, Han-Ying Zhang, Boqing Gong, Tingbo Hou, H. Wang, +1 more

    arXiv · 2023

    134
  • ReVersion: Diffusion-Based Relation Inversion from Images

    Ziqi Huang, Tianxing Wu, Yu-Ming Jiang, Kelvin C. K. Chan, Ziwei Liu

    SIGGRAPH Asia Conference Papers · 2024

    This work proposes the Relation Inversion task, which aims to learn a specific relation (represented as “relation prompt”) from exemplar images and proposes a novel “relation-steering contrastive learning” scheme to steer the relation prompt towards relation-dense regions, and disentangle it away from object appearances.

    101
  • Instruct-Imagen: Image Generation with Multi-modal Instruction

    Hexiang Hu, Kelvin C.K. Chan, Yu-Chuan Su, Wenhu Chen, Yan-Dong Li, Kihyuk Sohn, Yang Zhao, Xue Ben, +4 more

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2024

    Human evaluation on various image generation datasets re-veals that Instruct-Imagen matches or surpasses prior task-specific models in-domain and demonstrates promising generalization to unseen and more complex tasks.

    98
  • Temporally consistent video colorization with deep feature propagation and self-regularization learning

    Yi-Hao Liu, Hengyuan Zhao, Kelvin C. K. Chan, Xin-Tao Wang, Chen Change Loy, Y. Qiao, Chao Dong

    Computational Visual Media · 2024

    A novel temporally consistent video colorization (TCVC) framework that effectively propagates frame-level deep features in a bidirectional way to enhance the temporal consistency of colorization and introduces a self-regularization learning (SRL) scheme to minimize the differences in predictions obtained using different time steps.

    67
  • NTIRE 2019 Challenge on Video Super-Resolution: Methods and Results

    S. Nah, R. Timofte, Shu-Hang Gu, Sungyong Baik, Seokil Hong, Gyeongsik Moon, S. Son, Kyoung Mu Lee, +22 more

    IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) · 2019

    45
  • 44
  • NTIRE 2021 Challenge on Quality Enhancement of Compressed Video: Methods and Results

    Ren Yang, R. Timofte, Jing Liu, Yi Xu, Xinjian Zhang, Minyi Zhao, Shui-Geng Zhou, Kelvin C. K. Chan, +22 more

    IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) · 2021

    This paper reviews the first NTIRE challenge on quality enhancement of compressed video, with a focus on the proposed methods and results, and gauge the state of the art of video quality enhancement.

    43
  • DreamInpainter: Text-Guided Subject-Driven Image Inpainting with Diffusion Models

    Shao-An Xie, Yang Zhao, Zhisheng Xiao, Kelvin C.K. Chan, Yan-Dong Li, Yanwu Xu, Kun Zhang, Tingbo Hou

    arXiv · 2023

    This paper proposes a two-step approach DreamInpainter, which employs a discriminative token selection module to eliminate redundant subject details, preserving the subject's identity while allowing changes according to other conditions such as mask shape and text prompts.

    41
  • NTIRE 2019 Challenge on Video Deblurring: Methods and Results

    S. Nah, R. Timofte, Sungyong Baik, Seokil Hong, Gyeongsik Moon, S. Son, Kyoung Mu Lee, Xintao Wang, +22 more

    IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) · 2019

    41
  • Dual Associated Encoder for Face Restoration

    Yu-Ju Tsai, Yu-Lun Liu, Lu Qi, Kelvin C. K. Chan, Ming Yang

    arXiv · 2023

    This work proposes a novel dual-branch framework named DAEFR, which introduces an auxiliary LQ branch that extracts crucial information from the LQ inputs and incorporates association training to promote effective synergy between the two branches, enhancing code prediction and output quality.

    36
  • GLEAN: Generative Latent Bank for Image Super-Resolution and Beyond

    Kelvin C. K. Chan, Xiangyu Xu, Xintao Wang, Jinwei Gu, Chen Change Loy

    IEEE Transactions on Pattern Analysis and Machine Intelligence · 2022

    The method, Generative LatEnt bANk (GLEAN), goes beyond existing practices by directly leveraging rich and diverse priors encapsulated in a pre-trained GAN, and extends to different tasks including image colorization and blind image restoration.

    36
  • NTIRE 2021 Challenge on Video Super-Resolution

    S. Son, Suyoung Lee, S. Nah, R. Timofte, K. Lee, Kelvin C. K. Chan, Shang-Chen Zhou, Xiangyu Xu, +22 more

    IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) · 2021

    This paper presents evaluation results from two competition tracks as well as the proposed solutions to the NTIRE 2021 Challenge on Video Super-Resolution, and develops conventional video SR methods focusing on the restoration quality.

    36
  • On the Generalization of BasicVSR++ to Video Deblurring and Denoising

    Kelvin C. K. Chan, Shang-Chen Zhou, Xiangyu Xu, Chen Change Loy

    arXiv · 2022

    The proposed framework achieves compelling performance with great efficiency in various video restoration tasks including video deblurring and denoising and achieves comparable performance to Transformer-based approaches with up to 79% of parameter reduction and 44x speedup.

    33
  • Reference-based Image and Video Super-Resolution via $C^{2}$-Matching

    Yu-Ming Jiang, Kelvin C. K. Chan, Xintao Wang, Chen Change Loy, Zi-Wei Liu

    IEEE Transactions on Pattern Analysis and Machine Intelligence · 2022

    A contrastive correspondence network, which learns transformation-robust correspondences using augmented views of the input image, and a dynamic aggregation module to address the potential misalignment issue between input images and reference images are designed.

    25
  • Multi-task Image Restoration Guided By Robust DINO Features

    Xin Lin, Jingtong Yue, Kelvin C. K. Chan, Lu Qi, Chao Ren, Jinshan Pan, Ming-Hsuan Yang

    arXiv · 2023

    It is observed that the features of DINOv2 can effectively model semantic information and are independent of degradation factors, and a multi-task image restoration approach leveraging robust features extracted from DINOv2 to solve multi-task image restoration simultaneously is proposed.

    24
  • A Simple Approach to Unifying Diffusion-based Conditional Generation

    Xirui Li, Charles Herrmann, Kelvin C. K. Chan, Yin-Xiao Li, Deqing Sun, Chao Ma, Ming-Hsuan Yang

    arXiv · 2024

    This work introduces a simple, unified framework to handle diverse conditional generation tasks involving a specific image-condition correlation, and demonstrates that multiple models can be effectively combined for multi-signal conditional generation.

    19
  • Re-Boosting Self-Collaboration Parallel Prompt GAN for Unsupervised Image Restoration

    Xin Lin, Yuyan Zhou, Jingtong Yue, Chao Ren, Kelvin C. K. Chan, Lu Qi, Ming-Hsuan Yang

    IEEE Transactions on Pattern Analysis and Machine Intelligence · 2025

    A self-collaboration (SC) strategy for existing restoration models is proposed, achieving significant performance improvement without increasing the framework’s inference complexity.

    16
  • Improving Subject-Driven Image Synthesis with Subject-Agnostic Guidance

    Kelvin C.K. Chan, Yang Zhao, Xu-Hui Jia, Ming-Hsuan Yang, Huisheng Wang

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2024

    This work shows that through constructing a subject-agnostic condition and applying their proposed dual classifier-free guidance, one could obtain outputs consistent with both the given subject and input text prompts, and demonstrates its applicability in second-order customization methods, where an encoder-based model is fine-tuned with DreamBooth.

    10
  • Identity Encoder for Personalized Diffusion

    Yu-Chuan Su, Kelvin C. K. Chan, Yan-Dong Li, Yang Zhao, Han-Ying Zhang, Boqing Gong, H. Wang, Xu-Hui Jia

    arXiv · 2023

    This work learns an identity encoder which can extract an identity representation from a set of reference images of a subject, together with a diffusion generator that can generate new images of the subject conditioned on the identity representation.

    10
  • HoliSDiP: Image Super-Resolution via Holistic Semantics and Diffusion Prior

    Li-Yuan Tsao, Hao-Wei Chen, Hao-Wei Chung, De-Qing Sun, Chun-Yi Lee, Kelvin C. K. Chan, Ming-Hsuan Yang

    arXiv · 2024

    HoliSDiP is presented, a framework that leverages semantic segmentation to provide both precise textual and spatial guidance for diffusion-based Real-ISR through reduced prompt noise and enhanced spatial control.

    9
  • KITTEN: A Knowledge-Intensive Evaluation of Image Generation on Visual Entities

    Hsin-Ping Huang, Xinyi Wang, Yonatan Bitton, Hagai Taitelbaum, Gaurav Singh Tomar, Ming-Wei Chang, Xu-Hui Jia, Kelvin C. K. Chan, +3 more

    arXiv · 2024

    This work proposes KITTEN, a benchmark for Knowledge-Inensive image generaTion on real-world ENtities, and conducts a systematic study of the latest text-to-image models and retrieval-augmented models, focusing on their ability to generate real-world visual entities, such as landmarks and animals.

    7
  • A Convex Model for Edge-Histogram Specification with Applications to Edge-Preserving Smoothing

    Kelvin C. K. Chan, R. Chan, M. Nikolova

    Axioms · 2018

    This paper directly considers the image gradients and proposes a convex model based on them that allows us to compute the output image efficiently using either Alternating Direction Method of Multipliers or Fast Iterative Shrinkage-Thresholding Algorithm.

    4
  • From Prompt to Progression: Taming Video Diffusion Models for Seamless Attribute Transition

    Ling Lo, Kelvin C. K. Chan, Wen-Huang Cheng, Ming-Hsuan Yang

    IEEE/CVF International Conference on Computer Vision (ICCV) · 2025

    This work proposes a simple yet effective method to extend existing models for smooth and consistent attribute transitions, through introducing frame-wise guidance during the denoising process, and presents the Controlled-Attribute-Transition Benchmark (CAT-Bench), which integrates both attribute and motion dynamics.

    3
  • AdaIR: Exploiting Underlying Similarities of Image Restoration Tasks with Adapters

    Hao-Wei Chen, Yu-Syuan Xu, Kelvin C. K. Chan, Hsien-Kai Kuo, Chun-Yi Lee, Ming-Hsuan Yang

    arXiv · 2024

    AdaIR is proposed, a novel framework that enables low storage cost and efficient training without sacrificing performance and achieves outstanding results on multi-task restoration while utilizing significantly fewer parameters and less training time.

    3
  • CoCoIns: Consistent Subject Generation via Contrastive Instantiated Concepts

    Hsin-Ying Lee, Kelvin C. K. Chan, Ming-Hsuan Yang

    arXiv · 2025

    This work introduces Contrastive Concept Instantiation (CoCoIns), a framework that effectively synthesizes consistent subjects across multiple independent generations and proposes a contrastive learning approach that trains the network to distinguish between different combinations of prompts and latent codes.

    1
  • Effective Adapter for Face Recognition in the Wild

    Yunhao Liu, Lu Qi, Yu-Ju Tsai, Xiang-Tai Li, Kelvin C. K. Chan, Ming-Hsuan Yang

    arXiv · 2023

    This paper proposes an effective adapter for augmenting existing face recognition models trained on high-quality facial datasets using two similar structures, one fixed and the other trainable, to process both the unrefined and enhanced images using two similar structures.

    1
  • 1
  • Towards Robust Video Frame Interpolation with Long-Term Propagation

    Zi-Qi Huang, Kelvin C. K. Chan, Bihan Wen, Zi-Wei Liu

    Lecture notes in computer science · 2025

    –
  • Multimodal Face Generation and Manipulation with Collaborative Diffusion Models

    Advances in computer vision and pattern recognition · 2025

    –
  • RealSR-Nikon

    TIB Data Manager · 2024

    –

Publication data from OpenAlex, with missing venues and authors filled in from Crossref; citation counts are the higher of OpenAlex and Semantic Scholar, last synced 2026-10-10. One-sentence summaries under some papers are written by Semantic Scholar’s model. Citation counts may be lower than on Google Scholar, which indexes more sources.

Report an error

Wrong papers, two people merged into one, or a profile that should not be here? Tell us and we will fix or hide it. You will be asked to sign in.