Publications

Publications in reverse chronological order.

2026

  1. arXiv
    Tri-Prompting: Video Diffusion with Unified Control over Scene, Subject, and Motion
    Tri-Prompting: Video Diffusion with Unified Control over Scene, Subject, and Motion
    Z. Zhou, X. Zhan, Z. Chen, S. Y. Kim, N. Zhao, H. Zheng, Q. Liu, H. Zhang, Z. Lin, Y. Zhou, and J. Luo
    arXiv preprint arXiv:2603.15614, 2026
  2. CVPR
    HBridge: H-Shape Bridging of Heterogeneous Experts for Unified Multimodal Understanding and Generation
    HBridge: H-Shape Bridging of Heterogeneous Experts for Unified Multimodal Understanding and Generation
    X. Wang, Z. Zhang, H. Zhang, Z. Lin, Y. Zhou, Q. Liu, S. Zhang, Y. Li, S. Liu, H. Zheng, J. Kuen, Y. Wang, C. Gao, and N. Sang
    In IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2026

2025

  1. AAAI
    RealUHR: Harnessing Patch-Cascade Flows for Photorealistic Ultra-High-Resolution Synthesis
    RealUHR: Harnessing Patch-Cascade Flows for Photorealistic Ultra-High-Resolution Synthesis
    Y. Yu, H. Zheng, Z. Lin, C. Barnes, Y. Zhou, and J. Luo
    In Proceedings of the AAAI Conference on Artificial Intelligence (AAAI), 2025
  2. NeurIPS
    PixPerfect: Seamless Latent Diffusion Local Editing with Discriminative Pixel-Space Refinement
    PixPerfect: Seamless Latent Diffusion Local Editing with Discriminative Pixel-Space Refinement
    H. Zheng, Y. Yao, Y. Yu, Y. Zhou, Z. Lin, and J. Luo
    In Advances in Neural Information Processing Systems (NeurIPS), 2025
  3. ICCV
    omnipaint.gif
    OmniPaint: Mastering Object-Oriented Editing via Disentangled Insertion-Removal Inpainting
    Y. Yu, Z. Zeng, H. Zheng, and J. Luo
    In IEEE/CVF International Conference on Computer Vision (ICCV), 2025
  4. ICCV
    dollar.gif
    DOLLAR: Few-Step Video Generation via Distillation and Latent Reward Optimization
    Z. Ding, C. Jin, D. Liu, H. Zheng, K. Singh, Q. Zhang, Y. Kang, Z. Lin, and Y. Liu
    In IEEE/CVF International Conference on Computer Vision (ICCV), 2025
  5. CVPR
    TurboFill: Adapting Few-step Text-to-image Model for Fast Image Inpainting
    TurboFill: Adapting Few-step Text-to-image Model for Fast Image Inpainting
    L. Xie, D. Pakhomov, Z. Wang, Z. Wu, Z. Chen, Y. Zhou, H. Zheng, Z. Zhang, Z. Lin, J. Zhou, and C. Dong
    In IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2025
  6. CVPR
    MetaShadow: Object-Centered Shadow Detection, Removal, and Synthesis
    MetaShadow: Object-Centered Shadow Detection, Removal, and Synthesis
    T. Wang, J. Zhang, H. Zheng, Z. Ding, S. Cohen, Z. Lin, W. Xiong, C.-W. Fu, L. Figueroa, and S. Y. Kim
    In IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2025
  7. arXiv
    ZipIR: Latent Pyramid Diffusion Transformer for High-Resolution Image Restoration
    ZipIR: Latent Pyramid Diffusion Transformer for High-Resolution Image Restoration
    Y. Yu, H. Zheng, Z. Zhang, J. Zhang, Y. Zhou, C. Barnes, Y. Liu, W. Xiong, Z. Lin, and J. Luo
    arXiv preprint arXiv:2504.08591, 2025

2024

  1. SIGGRAPH
    deocclusion.jpg
    Object-level Scene Deocclusion
    Z. Liu, Q. Liu, C. Chang, J. Zhang, D. Pakhomov, H. Zheng, Z. Lin, D. Cohen-Or, and C. Fu
    In ACM SIGGRAPH 2024 Conference Papers, 2024
  2. TPAMI
    Structure-Guided Image Inpainting with Image-level and Object-level Semantic Discriminators
    Structure-Guided Image Inpainting with Image-level and Object-level Semantic Discriminators
    H. Zheng, Z. Lin, J. Lu, S. Cohen, E. Shechtman, J. Zhang, N. Xu, S. Amirghodsi, and J. Luo
    IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI), 2024

2023

  1. ICIP
    Point Cloud Denoising via Momentum Ascent in Gradient Fields
    Point Cloud Denoising via Momentum Ascent in Gradient Fields
    Y. Zhao, H. Zheng, Z. Wang, J. Luo, and E. Y. Lam
    In IEEE International Conference on Image Processing (ICIP), 2023
  2. ICIP
    Improving Video Colorization by Test-Time Tuning
    Improving Video Colorization by Test-Time Tuning
    Y. Zhao, H. Zheng, J. Luo, and E. Y. Lam
    In IEEE International Conference on Image Processing (ICIP), 2023
  3. ACM MM
    Jurassic World Remake: Bringing Ancient Fossils Back to Life via Zero-Shot Long Image-to-Image Translation
    Jurassic World Remake: Bringing Ancient Fossils Back to Life via Zero-Shot Long Image-to-Image Translation
    A. Martin, H. Zheng, J. An, and J. Luo
    In ACM International Conference on Multimedia (ACM MM), 2023

2022

  1. ECCV
    Image Inpainting with Cascaded Modulation GAN and Object-Aware Training
    Image Inpainting with Cascaded Modulation GAN and Object-Aware Training
    H. Zheng, Z. Lin, J. Lu, S. Cohen, E. Shechtman, C. Barnes, J. Zhang, N. Xu, S. Amirghodsi, and J. Luo
    In European Conference on Computer Vision (ECCV), 2022
  2. TPAMI
    Semantic Layout Manipulation with High-Resolution Sparse Attention
    Semantic Layout Manipulation with High-Resolution Sparse Attention
    H. Zheng, Z. Lin, J. Lu, S. Cohen, J. Zhang, N. Xu, and J. Luo
    IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI), 2022
  3. CVPR
    spaceedit.jpg
    SpaceEdit: Learning a Unified Editing Space for Open-Domain Image Color Editing
    J. Shi, N. Xu, H. Zheng, A. Smith, J. Luo, and C. Xu
    In IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2022
  4. ICIP
    manet.jpg
    MANet: Improving Video Denoising with a Multi-Alignment Network
    Y. Zhao, H. Zheng, Z. Wang, J. Luo, and E. Y. Lam
    In IEEE International Conference on Image Processing (ICIP), 2022
  5. ICPR
    Learning to Aggregate and Refine Noisy Labels for Visual Sentiment Analysis
    Learning to Aggregate and Refine Noisy Labels for Visual Sentiment Analysis
    W. Zhu, Z. Zheng, H. Zheng, H. Lyu, and J. Luo
    In International Conference on Pattern Recognition (ICPR), 2022
  6. CICAI
    Cross-Camera Deep Colorization
    Cross-Camera Deep Colorization
    Y. Zhao, H. Zheng, M. Ji, and R. Huang
    In CAAI International Conference on Artificial Intelligence (CICAI), 2022

2021

  1. Big Data
    fashion.jpg
    Personalized Fashion Recommendation from Personal Social Media Data: An Item-to-Set Metric Learning Approach
    H. Zheng, K. Wu, J. Park, W. Zhu, and J. Luo
    In IEEE International Conference on Big Data (Big Data), 2021
  2. ICCV
    biasinvariant.jpg
    Learning Bias-Invariant Representation by Cross-Sample Mutual Information Minimization
    W. Zhu, H. Zheng, H. Liao, W. Li, and J. Luo
    In IEEE/CVF International Conference on Computer Vision (ICCV), 2021

2020

  1. ECCV
    Example-Guided Image Synthesis using Masked Spatial-Channel Attention and Self-Supervision
    Example-Guided Image Synthesis using Masked Spatial-Channel Attention and Self-Supervision
    H. Zheng, H. Liao, L. Chen, W. Xiong, and J. Luo
    In European Conference on Computer Vision (ECCV), 2020
  2. TIP
    poseflow.gif
    Pose Flow Learning from Person Images for Pose Guided Synthesis
    H. Zheng, L. Chen, C. Xu, and J. Luo
    IEEE Transactions on Image Processing (TIP), 2020
  3. TPAMI
    CrossNet++: Cross-scale Large-parallax Warping for Reference-based Super-resolution
    CrossNet++: Cross-scale Large-parallax Warping for Reference-based Super-resolution
    Y. Tan, H. Zheng, Y. Zhu, and L. Fang
    IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI), 2020
  4. ICPR
    Cost-Effective Adversarial Attacks against Scene Text Recognition
    M. Yang, H. Zheng, X. Bai, and J. Luo
    In International Conference on Pattern Recognition (ICPR), 2020
  5. ACM MM
    sentiment.jpg
    Image Sentiment Transfer
    T. Chen, W. Xiong, H. Zheng, and J. Luo
    In ACM International Conference on Multimedia (ACM MM), 2020

2018

  1. ECCV
    CrossNet: An End-to-end Reference-based Super Resolution Network using Cross-scale Warping
    CrossNet: An End-to-end Reference-based Super Resolution Network using Cross-scale Warping
    H. Zheng, M. Ji, H. Wang, Y. Liu, and L. Fang
    In European Conference on Computer Vision (ECCV), 2018

2017

  1. BMVC
    Learning Cross-scale Correspondence and Patch-based Synthesis for Reference-based Super-Resolution
    Learning Cross-scale Correspondence and Patch-based Synthesis for Reference-based Super-Resolution
    H. Zheng, M. Ji, Z. Xu, H. Wang, Y. Liu, and L. Fang
    In British Machine Vision Conference (BMVC), 2017
  2. ICCV
    surfacenet.jpg
    SurfaceNet: an End-to-end 3D Neural Network for Multiview Stereopsis
    M. Ji, G. Juergen, H. Zheng, Y. Liu, and L. Fang
    In IEEE International Conference on Computer Vision (ICCV), 2017
  3. ICCVW
    Combining Exemplar-based Approach and Learning-based Approach for Light Field Super-resolution Using a Hybrid Imaging System
    Combining Exemplar-based Approach and Learning-based Approach for Light Field Super-resolution Using a Hybrid Imaging System
    H. Zheng, M. Guo, Y. Liu, and L. Fang
    In IEEE International Conference on Computer Vision Workshops (ICCVW), 2017
  4. GlobalSIP
    Utilizing High-level Visual Feature for Indoor Shopping Mall Navigation
    Utilizing High-level Visual Feature for Indoor Shopping Mall Navigation
    Z. Xu, H. Zheng, M. Pang, Y. Zhu, X. Su, G. Zhou, and L. Fang
    In IEEE Global Conference on Signal and Information Processing (GlobalSIP), 2017

2016

  1. TMM
    Deep Learning for Surface Material Classification Using Haptic and Visual Information
    Deep Learning for Surface Material Classification Using Haptic and Visual Information
    H. Zheng, L. Fang, M. Ji, M. Strese, Y. Oezer, and E. Steinbach
    IEEE Transactions on Multimedia (TMM), 2016
  2. TCSVT
    Computation and Memory Efficient Image Segmentation
    Computation and Memory Efficient Image Segmentation
    Y. Zhou, T. T. Do, H. Zheng, L. Fang, and N. M. Cheung
    IEEE Transactions on Circuits and Systems for Video Technology (TCSVT), 2016

2015

  1. APSIPA
    Analysis of Sports Statistics via Graph-Signal Smoothness Prior
    Analysis of Sports Statistics via Graph-Signal Smoothness Prior
    H. Zheng, G. Cheung, and L. Fang
    In Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC), 2015
  2. MLSP
    Preprocessing-free Surface Material Classification Using Convolutional Neural Networks Pretrained by Sparse Autoencoder
    M. Ji, L. Fang, H. Zheng, M. Strese, and E. Steinbach
    In IEEE International Workshop on Machine Learning for Signal Processing (MLSP), 2015

2014

  1. EMBC
    Early Melanoma Diagnosis with Mobile Imaging
    T. T. Do, Y. Zhou, H. Zheng, N. M. Cheung, and D. Koh
    In International Conference of the IEEE Engineering in Medicine and Biology Society (EMBC), 2014