Is this you? Claim this profile to correct it, add a bio and choose the work people see first.
Claim this profileAcademic lineage
Possible advisorsa guess from early papers, not confirmed
- Haifeng HuSuggested from co-authorship
Is this you? Claim this profile to confirm or dismiss it.
Works36 from public data
- 94
- 37
- 34
- 34
- 34
- 31
- 29
- 27
- InteractDiffusion: Interaction Control in Text-to-Image Diffusion Models26
A pluggable interaction control model is proposed that extends existing pre-trained T2I diffusion models to enable them being better conditioned on interactions, and outperforms existing baselines by a large margin in HOI detection score, as well as fidelity in FID and KID.
- 26
- Discriminant Deep Feature Learning based on joint supervision Loss and Multi-layer Feature Fusion for heterogeneous face recognition25
A new loss function called Scatter Loss (SL) is proposed, which embeds both inter- and intra-class information for effectively training the deep model, to enhance the discriminative power of the deeply learned features.
- 20
- 17
- Heterogeneous Face Recognition Based on Multiple Deep Networks With Scatter Loss and Diversity Combination14
This paper designs a multiple deep networks (MDN) structure for feature extraction and proposes a joint decision strategy called diversity combination (DC) to adaptively adjust the weights of each deep network and make a joint classification decision.
- 13
- Syncretic Space Learning Network for NIR-VIS Face Recognition10
This work simultaneously synthesizes NIR and VIS images into modality-independent syncretic images and proposes a novelsyncretic space learning (SSL) model to eliminate the modal gap, and develops the Syncretic Distribution Consistency (SDC), which can enhance the intra-class compactness and learn discriminative representations.
- 9
- Fine Tuning Dual Streams Deep Network with Multi-scale Pyramid Decision for Heterogeneous Face Recognition9
A novel method called fine tuning dual streams deep network (FTDSDN) with multi-scale pyramid decision (MsPD) with powerful joint decision strategy called MsPD to adaptively adjust the weight of sub structure and obtain more robust classification performance.
- 8
- 7
- 7
- E3RG: Building Explicit Emotion-driven Empathetic Response Generation System with Multimodal Large Language Model7
E3RG, an Explicit Emotion-driven Empathetic Response Generation System based on multimodal LLMs which decomposes MERG task into three parts: multimodal empathy understanding, empathy memory retrieval, and multimodal response generation is proposed.
- Cascaded Dynamic Memory Refinement and Semantic Alignment for Exo-to-Ego Cross-View Video Generation6
- Boundary Voting Network for Ambiguity-Aware Timestamp-Supervised Action Segmentation5
The boundary voting network is introduced that mitigates feature ambiguity by hierarchically propagating video-level global prior knowledge into local action-transiting regions by generating key action representations as votes throughout the video and targeting action-transiting regions.
- Heterogeneous face recognition based on modality‐independent Kernel Fisher discriminant analysis joint sparse auto‐encoder5
A novel method called modality-independent Kernel discriminant analysis joint sparse auto-encoder, for solving heterogeneous face recognition problem is proposed, which does not require the data correspondences when collecting external cross-modal data and is practical for real-world cross- modal classification problem.
Publication data from OpenAlex; citation counts are the higher of OpenAlex and Semantic Scholar, last synced 2026-10-11. One-sentence summaries under some papers are written by Semantic Scholar’s model. Citation counts may be lower than on Google Scholar, which indexes more sources.
Report an error
Wrong papers, two people merged into one, or a profile that should not be here? Tell us and we will fix or hide it. You will be asked to sign in.