Is this you? Claim this profile to correct it, add a bio and choose the work people see first.

Claim this profile

Academic lineage

Possible advisorsa guess from early papers, not confirmed

  • Ping Luo

    Possible advisor · last author on 3 of their early first-author papers, 2022

    Suggested from co-authorship
  • Liang Lin

    Possible advisor · last author on 4 of their early first-author papers, 2017–2018

    Suggested from co-authorship

Is this you? Claim this profile to confirm or dismiss it.

Works22 from public data

TitleCited by
  • MotionCtrl: A Unified and Flexible Motion Controller for Video Generation

    Zhouxia Wang, Ziyang Yuan, Xintao Wang, Yaowei Li, Tianshui Chen, Menghan Xia, Ping Luo, Ying Shan

    ACM SIGGRAPH Conference Papers (SIGGRAPH) · 2024

    MotionCtrl is presented, a unified and flexible motion controller for video generation designed to effectively and independently control camera and object motion and is a relatively generalizable model that can adapt to a wide array of camera poses and trajectories once trained.

    673
  • Multi-label Image Recognition by Recurrently Discovering Attentional Regions

    Zhouxia Wang, Tianshui Chen, Guanbin Li, Ruijia Xu, Liang Lin

    IEEE International Conference on Computer Vision (ICCV) · 2017

    This work achieves the interpretable and contextualized multi-label image classification by developing a recurrent memorized-attention module that demonstrates superior performances over other existing state-of-the-arts in both accuracy and efficiency.

    338
  • Deep Reasoning with Knowledge Graph for Social Relationship Understanding

    Zhouxia Wang, Tianshui Chen, Jimmy Ren, Weihao Yu, Hui Cheng, Liang Lin

    International Joint Conference on Artificial Intelligence · 2018

    This work has found that the interplay between these two factors can be effectively modeled by a novel structured knowledge graph with proper message propagation and attention and can be efficiently integrated into the deep neural network architecture to promote social relationship understanding.

    198
  • RestoreFormer: High-Quality Blind Face Restoration from Undegraded Key-Value Pairs

    Zhouxia Wang, Jiawei Zhang, Runjian Chen, Wenping Wang, Ping Luo

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2022

    This work proposes a method, RestoreFormer, which explores fully-spatial attentions to model contextual information and surpasses existing works that use local operators and outperforms advanced state-of-the-art methods on one synthetic dataset and three real-world datasets.

    177
  • Recurrent Attentional Reinforcement Learning for Multi-Label Image Recognition

    Tianshui Chen, Zhouxia Wang, Guanbin Li, Liang Lin

    Proceedings of the AAAI Conference on Artificial Intelligence · 2018

    A recurrent attention reinforcement learning framework to iteratively discover a sequence of attentional and informative regions that are related to different semantic objects and further predict label scores conditioned on these regions to facilitate multi-label recognition.

    176
  • LSTM Pose Machines

    Yue Gang Luo, Jimmy Ren, Zhouxia Wang, Wenxiu Sun, Jinshan Pan, Jianbo Liu, Jiahao Pang, Liang Lin

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2018

    It is shown that if the weight sharing scheme is imposed to the multi-stage CNN, it could be re-written as a Recurrent Neural Network (RNN), which decouples the relationship among multiple network stages and results in significantly faster speed in invoking the network for videos.

    148
  • FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models

    Haonan Qiu, Zhaoxi Chen, Zhouxia Wang, Yingqing He, Menghan Xia, Ziwei Liu

    arXiv · 2024

    A tuning-free framework to achieve trajectory-controllable video generation, by imposing guidance on both noise construction and attention computation, and proposes FreeTraj, a tuning-free approach that enables trajectory control by modifying noise sampling and attention mechanisms.

    61
  • RestoreFormer++: Towards Real-World Blind Face Restoration From Undegraded Key-Value Pairs

    Zhouxia Wang, Jiawei Zhang, Tianshui Chen, Wenping Wang, Ping Luo

    IEEE Transactions on Pattern Analysis and Machine Intelligence · 2023

    This work proposes RestoreFormer++, which on the one hand introduces fully-spatial attention mechanisms to model the contextual information and the interplay with the priors, and on the other hand explores an extending degrading model to help generate more realistic degraded face images to alleviate the synthetic-to-real-world gap.

    56
  • StyleAdapter: A Unified Stylized Image Generation Model

    Zhouxia Wang, Xintao Wang, Liangbin Xie, Zhongang Qi, Ying Shan, Wenping Wang, Ping Luo

    International Journal of Computer Vision · 2024

    The StyleAdapter is proposed, a unified stylized image generation model capable of producing a variety of stylized images that match both the content of a given prompt and the style of reference images, without the need for per-style fine-tuning.

    52
  • Diffusion-based Blind Text Image Super-Resolution

    Yuzhe Zhang, Jiawei Zhang, Hao Li, Zhouxia Wang, Luwei Hou, Dongqing Zou, Liheng Bian

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2024

    Extensive experiments on synthetic and real-world datasets demonstrate that the Diffusion-based Blind Text Image Super-Resolution (DiffTSR) can restore text images with more accurate text structures as well as more realistic appearances simultaneously.

    43
  • Learning a Reinforced Agent for Flexible Exposure Bracketing Selection

    Zhouxia Wang, Jiawei Zhang, Mude Lin, Jiong Wang, Ping Luo, Jimmy Ren

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2020

    A novel deep neural network to automatically select exposure bracketing, named EBSNet, which is sufficiently flexible without having the above restrictions and can be jointly trained to produce favorable results against recent state-of-the-art approaches.

    24
  • RoVer: Robot Reward Model as Test-Time Verifier for Vision-Language-Action Model

    Mingtong Dai, Lingbo Liu, Bai, Yongjie, Yang Liu, Zhouxia Wang, Rui Jing Su, Chunjie Chen, Lin, Liang, +1 more

    arXiv · 2025

    The approach effectively transforms available computing resources into better action decision-making, realizing the benefits of test-time scaling without extra training overhead.

    21
  • Multi-label image recognition with attentive transformer-localizer module

    Lin Nie, Tianshui Chen, Zhouxia Wang, Wenxiong Kang, Liang Lin

    Multimedia Tools and Applications · 2022

    A novel attentive transformer-localizer (ATL) module is applied to progressively localize the attentive regions from the convolutional feature maps in a proposal-free manner, and the LSTM network sequentially predicts label scores for the localized regions and updates the parameters of the ATL module while capturing the global dependencies among these regions.

    18
  • ObjCtrl-2.5D: Training-free Object Control with Camera Poses

    Zhouxia Wang, Yushi Lan, Shangchen Zhou, Chen Change Loy

    International Journal of Computer Vision · 2026

    17
  • Denoising as Adaptation: Noise-Space Domain Adaptation for Image Restoration

    Kang Liao, Zongsheng Yue, Zhouxia Wang, Chen Change Loy

    arXiv · 2024

    This paper shows that it is possible to perform domain adaptation via the noise space using diffusion models and derives a meaningful diffusion loss that guides the restoration model in progressively aligning both restored synthetic and real-world outputs with a target clean distribution.

    17
  • Novel Antimicrobial Nano Bacteriocin: Lactic Acid Bacteria‐Derived, Self‐Assembled, and Enhanced for Superior Antimicrobial Activity

    Lanhua Yi, Shengyang Li, Miaomiao Xie, Bozhou Chen, K J Chen, Zhouxia Wang, Min Zeng, Qian Zhao, +9 more

    Advanced Materials · 2026

    The carrier‐free self‐assembly approach overcomes AMP stability and solubility limitations and paves the way for next‐generation antimicrobial therapies.

    14
  • Image Conductor: Precision Control for Interactive Video Synthesis

    Yaowei Li, Xintao Wang, Zhaoyang Zhang, Zhouxia Wang, Ziyang Yuan, Liangbin Xie, Ying Shan, Yuexian Zou

    Proceedings of the AAAI Conference on Artificial Intelligence · 2025

    Image Conductor, a method for precise control of camera transitions and object movements to generate video assets from a single image, is proposed, and a trajectory-oriented video motion data curation pipeline for training is developed.

    11
  • Recovering Extremely Degraded Faces by Joint Super-Resolution and Facial Composite

    Xiu Li, Guichun Duan, Zhouxia Wang, Jimmy Ren, Yongbing Zhang, Jiawei Zhang, Kaixiang Song

    Proceedings - International Conference on Tools with Artificial Intelligence, TAI · 2019

    8
  • Analysis and Benchmarking of Extending Blind Face Image Restoration to Videos

    Zhouxia Wang, Jiawei Zhang, Xintao Wang, Tianshui Chen, Ying Shan, Wenping Wang, Ping Luo

    IEEE Transactions on Image Processing · 2024

    A Temporal Consistency Network (TCN) cooperated with alignment smoothing to reduce jitters and flickers in restored videos is proposed, a flexible component that can be seamlessly plugged into the most advanced face image restoration algorithms, ensuring the quality of image-based restoration is maintained as closely as possible.

    7
  • Image Deblurring Aided by Low-Resolution Events

    Zhouxia Wang, Jimmy Ren, Jiawei Zhang, Ping Luo

    Electronics · 2022

    An alternately performed model is proposed in this paper to deblur high-resolution images with the help of low-resolution events and enhances the quality of events with EventSRNet by extracting the structure information in the generated sharp image.

    6
  • Precise Object and Effect Removal with Adaptive Target-Aware Attention

    Jixin Zhao, Zhouxia Wang, Peiqing Yang, Shangchen Zhou

    arXiv · 2025

    1
  • Bridge-WA: Learning Action-Relevant World Dynamics for Robotic Manipulation

    Yongjie Bai, Mingtong Dai, Zhouxia Wang, Hanting Wang, Qijun Zhong, Feng Yan, Yang Liu, Liang Lin

    arXiv · 2026

    –

Publication data from OpenAlex; citation counts are the higher of OpenAlex and Semantic Scholar, last synced 2026-10-11. One-sentence summaries under some papers are written by Semantic Scholar’s model. Citation counts may be lower than on Google Scholar, which indexes more sources.

Report an error

Wrong papers, two people merged into one, or a profile that should not be here? Tell us and we will fix or hide it. You will be asked to sign in.