Is this you? Claim this profile to correct it, add a bio and choose the work people see first.
Claim this profileAcademic lineage
Possible advisorsa guess from early papers, not confirmed
- Xinchao WangSuggested from co-authorship
Is this you? Claim this profile to confirm or dismiss it.
Works12 from public data
- OminiControl: Minimal and Universal Control for Diffusion Transformer386
OminiControl, a novel approach that rethinks how image conditions are integrated into Diffusion Transformer (DiT) architectures, demonstrates that effective image control can be achieved without architectural complexity, opening new possibilities for efficient and versatile image generation systems.
- MindBridge: A Cross-Subject Brain Decoding Framework109
This paper presents a novel approach, MindBridge, that achieves cross-subject brain decoding by employing only one model, and establishes a generic paradigm capable of addressing challenges of brain decoding, by introducing biological-inspired aggregation function and novel cyclic fMRI reconstruction mechanism for subject-invariant representation learning.
- MEMO: Memory-Guided Diffusion for Expressive Talking Video Generation44
Memory-guided EMOtion-aware diffusion (MEMO), an end-to-end audio-driven portrait animation approach to generate identity-consistent and expressive talking videos that outperforms state-of-the-art methods in overall quality, audio-lip synchronization, identity consistency, and expression-emotion alignment.
- 7
- ViFeEdit: A Video-Free Tuner of Your Video Diffusion Transformer2
A video-free tuning framework termed ViFeEdit for video diffusion transformers that enables visually faithful editing while maintaining temporal consistency with only minimal additional parameters, and delivers promising results of controllable video generation and editing with only minimal training on 2D image data.
- Gated Condition Injection without Multimodal Attention: Towards Controllable Linear-Attention Transformers2
This paper proposes a novel controllable diffusion framework tailored for linear attention backbones like SANA, which achieves state-of-the-art controllable generation performance based on linear-attention models, surpassing existing methods in terms of fidelity and controllability.
- 2
- 1
- 1
- –
- –
- –
Publication data from OpenAlex; citation counts are the higher of OpenAlex and Semantic Scholar, last synced 2026-10-11. One-sentence summaries under some papers are written by Semantic Scholar’s model. Citation counts may be lower than on Google Scholar, which indexes more sources.
Report an error
Wrong papers, two people merged into one, or a profile that should not be here? Tell us and we will fix or hide it. You will be asked to sign in.