Is this you? Claim this profile to correct it, add a bio and choose the work people see first.
Claim this profileWorks8 from public data
- Biased-Predicate Annotation Identification via Unbiased Visual Predicate Representation33
This work proposes a novel method that utilizes unbiased visual predicate representations for Biased-Annotation Identification (BAI) as a fundamental step for PSG/SGG tasks and shows great generalization and effectiveness on multiple datasets.
- Mrtnet: Multi-Resolution Temporal Network for Video Sentence Grounding25
This work introduces MRTNet, a multi-resolution grounding network with four key components: a feature encoder, a Multi-Resolution Temporal (MRT) module, a Query-aware Attention (QAM) module, and a predictor.
- Exploration Entropy for Reinforcement Learning21
The Exploration Entropy is a new criterion to analyse and manage the training process of RL and the theoretical analysis and experiment results illustrate that the curve of exploration Entropy contains more information than the existing analytical methods.
- 9
- Grounding is All You Need? Dual Temporal Grounding for Video Dialog3
The Dual Temporal Grounding-enhanced Video Dialog model (DTGVD) is introduced, designed to bridge the gap between these two approaches to video dialog response generation and enables a more nuanced understanding of conversational dynamics.
- SRDiff: A Cross-Modal Diffusion Model for Satellite-to-Radar Translation in Precipitation Nowcasting3
- 2
- –
Publication data from OpenAlex; citation counts are the higher of OpenAlex and Semantic Scholar, last synced 2026-10-11. One-sentence summaries under some papers are written by Semantic Scholar’s model. Citation counts may be lower than on Google Scholar, which indexes more sources.
Report an error
Wrong papers, two people merged into one, or a profile that should not be here? Tell us and we will fix or hide it. You will be asked to sign in.