Is this you? Claim this profile to correct it, add a bio and choose the work people see first.
Claim this profileAcademic lineage
Possible advisorsa guess from early papers, not confirmed
- Wenxiu SunSuggested from co-authorship
Is this you? Claim this profile to confirm or dismiss it.
Works36 from public data
- Review of Drug Repositioning Approaches and Resources621
Computational approaches are reviewed and highlighted their characteristics to provide references for researchers to develop more powerful approaches and to summarized 76 important resources about drug repositioning.
- GRNet: Gridding Residual Network for Dense Point Cloud Completion468
This work devise two novel differentiable layers, named Gridding and Gridding Reverse, to convert between point clouds and 3D grids without losing structural information, and presents the differentiable Cubic Feature Sampling layer to extract features of neighboring points, which preserves context information.
- Pix2Vox: Context-Aware 3D Reconstruction From Single and Multi-View Images392
Experimental results on the ShapeNet and Pix3D benchmarks indicate that the proposed Pix2Vox outperforms state-of-the-arts by a large margin, and the proposed method is 24 times faster than 3D-R2N2 in terms of backward inference time.
- Spatio-Temporal Filter Adaptive Network for Video Deblurring237
The proposed Spatio-Temporal Filter Adaptive Network (STFAN) takes both blurry and restored images of the previous frame as well as blurry image of the current frame as input, and dynamically generates the spatially adaptive filters for the alignment and deblurring.
- Efficient Regional Memory Network for Video Object Segmentation183
The proposed RM-Net effectively alleviates the ambiguity of similar objects in both memory and query frames, which allows the information to be passed from the regional memory to the query region efficiently and effectively.
- CityDreamer: Compositional Generative Model of Unbounded 3D Cities116
The proposed CityDreamer achieves state-of-the-art performance not only in generating realistic 3D cities but also in local-ized editing within the generated cities.
- DAVANet: Stereo Deblurring With View Aggregation108
This work proposes a novel stereo image deblurring network with Depth Awareness and View Aggregation, named DAVANet, which outperforms state-of-the-art methods in terms of accuracy, speed, and model size.
- Comparison among dimensionality reduction techniques based on Random Projection for cancer classification72
This work attempts to improve classification accuracy of RP through combining other reduction dimension methods such as Principle Component Analysis (PCA), Linear Discriminant Analysis (LDA), and Feature Selection (FS).
- Generative Gaussian Splatting for Unbounded 3D City Generation59
GaussianCity, a generative Gaussian splatting framework dedicated to efficiently synthesizing unbounded 3D cities with a single feed-forward pass, is proposed and introduces BEV-Point as a highly compact intermediate representation, ensuring that the growth in VRAM usage for unbounded scenes remains constant, thus enabling unbounded city generation.
- 49
- 48
- DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation45
The Dynamic Object Manipulation (DOM) benchmark is introduced, built with an automated collection pipeline that gathers 200K synthetic episodes across 2.8K scenes and 206 objects, and enables fast collection of 2K real-world episodes without teleoperation.
- 3D Scene Generation: A Survey44
A systematic overview of state-of-the-art 3D scene generation approaches, organizing them into four paradigms: procedural generation, neural 3D-based generation, image-based generation, and video-based generation, and highlights promising directions at the intersection of generative AI, 3D vision, and embodied intelligence.
- 43
- 42
- Long-Range Feature Propagating for Natural Image Matting41
This work proposes Long-Range Feature Propagating Network (LFPNet), which learns the long-range context features outside the reception fields for alpha matte estimation and presents Center-Surround Pyramid Pooling (CSPP) that explicitly propagates the context features from the surrounding context image patch to the inner center image patch.
- Learning Geometric Transformation for Point Cloud Completion40
A simple yet effective geometric transformation network (GTNet) that exploits the repetitive geometric structures in common 3D objects to recover the complete shapes, which contains three sub-networks: geometric patch network, structure transformation network, and detail refinement network.
- 39
- SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture28
This work introduces SenseNova-U1, a native unified multimodal paradigm built upon NEO-unify, in which understanding and generation evolve as synergistic views of a single underlying process that points toward a broader roadmap where models do not translate between modalities, but think and act across them in a native manner.
- Toward 3D object reconstruction from stereo images24
A new deep learning framework for reconstructing the 3D shape of an object from a pair of stereo images, which reasons about the3D structure of the object by taking bidirectional disparities and feature correspondences between the two views into account is proposed.
- Compositional Generative Model of Unbounded 4D Cities23
The proposed CityDreamer4D is a compositional generative model specifically tailored for generating unbounded 4D cities that supports a range of downstream applications, such as instance editing, city stylization, and urban simulation, while delivering state-of-the-art performance in generating realistic 4D cities.
- 2D Semantic-Guided Semantic Scene Completion22
A Semantic-guided Semantic Scene Completion framework, dubbed SG-SSC, which involves Semantic-guided Fusion (SGF) and Volume-guided Semantic Predictor (VGSP), which outperforms existing state-of-the-art methods on the NYU, NYUCAD, and SemanticKITTI datasets.
- 14
- Weighted voxel12
A new voxel representation, named Weighted Voxel, is defined, which provides more abundant information, facilitating the subsequent learning and generalization steps, and makes full use of the structure information of voxels.
- DynamicCity: Large-Scale 4D Occupancy Generation from Dynamic Scenes10
The Masked Rollout Operation reorganizes HexPlane features for DiT-based diffusion, enabling versatile 4D scene generation and demonstrates that DynamicCity surpasses state-of-the-art methods in both reconstruction and generation.
Publication data from OpenAlex; citation counts are the higher of OpenAlex and Semantic Scholar, last synced 2026-10-11. One-sentence summaries under some papers are written by Semantic Scholar’s model. Citation counts may be lower than on Google Scholar, which indexes more sources.
Report an error
Wrong papers, two people merged into one, or a profile that should not be here? Tell us and we will fix or hide it. You will be asked to sign in.