Is this you? Claim this profile to correct it, add a bio and choose the work people see first.

Claim this profile

Works27 from public data

TitleCited by
  • OpenESS: Event-Based Semantic Scene Understanding with Open Vocabularies

    Lingdong Kong, Youquan Liu, Lai Xing Ng, Benoit R. Cottereau, Wei Tsang Ooi

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2024

    This work synergize information from image, text, and event-data domains and introduces OpenESS to enable scalable ESS in an open-world, annotation-efficient manner and proposes a frame-to-event contrastive distillation and a text-to-event semantic consistency regularization.

    50
  • WorldLens: Full-Spectrum Evaluations of Driving World Models in Real World

    Liang, Ao, Kong, Lingdong, Tianyi Yan, Liu, Hongsi, Yang, Wesley, Huang, Ziqi, Wei Yin, Zuo, Jialong, +14 more

    arXiv · 2025

    This work introduces WorldLens, a full-spectrum benchmark evaluating how well a model builds, understands, and behaves within its generated world, and develops WorldLens-Agent, an evaluation model distilled from these annotations to enable scalable, explainable scoring.

    29
  • The RoboDrive Challenge: Drive Anytime Anywhere in Any Condition

    Kong, Lingdong, Shaoyuan Xie, Hanjiang Hu, Yaru Niu, Wei Tsang Ooi, Benoit R. Cottereau, Lai Xing Ng, Yuexin Ma, +22 more

    arXiv · 2024

    This challenge has set a new benchmark in the field, providing a rich repository of techniques expected to guide future research in this field, and introduced a range of innovative approaches including advanced data augmentation, multi-sensor fusion, self-supervised learning for error correction, and new algorithmic strategies to enhance sensor robustness.

    28
  • Adversarial Robustness in Two-Stage Learning-to-Defer: Algorithms and Guarantees

    Yannis Montreuil, Axel Carlier, Lai Xing Ng, Wei Tsang Ooi

    arXiv · 2025

    This work proposes SARD, a convex learning algorithm built on a family of surrogate losses that are provably Bayes-consistent and $(\mathcal{R}, \mathcal{G})$-consistent, and demonstrates that these guarantees hold across classification, regression, and multi-task settings.

    23
  • The RoboDepth Challenge: Methods and Advancements Towards Robust Depth Estimation

    Kong, Lingdong, Yaru Niu, Shaoyuan Xie, Hanjiang Hu, Lai Xing Ng, Benoit R. Cottereau, Liangjun Zhang, Hesheng Wang, +22 more

    arXiv · 2023

    Out of more than two hundred participants, nine unique and top-performing solutions have appeared, with novel designs ranging from the following aspects: spatial- and frequency-domain augmentations, masked image modeling, image restoration and super-resolution, adversarial training, diffusion-based noise suppression, vision-language pre-training, learned model ensembling, and hierarchical feature enhancement.

    19
  • EventFly: Event Camera Perception from Ground to the Sky

    Lingdong Kong, Dongyue Lu, Xiang Xu, Lai Xing Ng, Wei Tsang Ooi, Benoit R. Cottereau

    IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings - IEEE Computer Society Conference on Computer Vision and Pattern Recognition/Proceedings · 2025

    This work introduces EventFly, a framework for robust cross-platform adaptation in event camera perception, and introduces EXPo, a large-scale benchmark with diverse samples across vehicle, drone, and quadruped platforms to holistically assess cross-platform adaptation abilities.

    16
  • RoboDepth: Robust Out-of-Distribution Depth Estimation under Corruptions

    Lingdong Kong, Shaoyuan Xie, Hanjiang Hu, Lai Xing Ng, Benoit R. Cottereau, Wei Tsang Ooi

    neural information processing systems · 2023

    12
  • AI for Auto-Research: Roadmap & User Guide

    Lingdong Kong, Xian Sun, Wei Chow, Linfeng Li, Kevin Qinghong Lin, Xuan Billy Zhang, Song Wang, Li, Rong, +12 more

    arXiv · 2026

    It is shown that greater automation can obscure rather than eliminate failure modes, making human-governed collaboration the most credible deployment paradigm, and a structured taxonomy, benchmark suite, and tool inventory, cross-stage design principles, and a practitioner-oriented playbook are provided.

    11
  • Beyond Augmented-Action Surrogates for Multi-Expert Learning-to-Defer

    Yannis Montreuil, Axel Carlier, Lai Xing Ng, Wei Tsang Ooi

    arXiv · 2026

    This work proposes a decoupled surrogate: a softmax classifier head and an independent sigmoid head per expert, mirroring the two natural objects of the problem, and proves an excess-risk bound with calibration constant, the first multi-expert L2D guarantee whose constant does not grow with the expert pool when the per-expert weight is held fixed.

    9
  • Online Learning-to-Defer with Varying Experts

    Dang Hoang Duy, Yannis Montreuil, Maxime Meyer, Axel Carlier, Lai Xing Ng, Wei Tsang Ooi

    arXiv · 2026

    An online multiclass L2D algorithm that combines queried-action bandit feedback with a dynamically varying pool of experts is introduced that achieves expected true-deferral regret under a concentrated-score condition.

    7
  • Learning-to-Defer in Non-Stationary Time Series via Switching State-Space Models

    Yannis Montreuil, Letian Yu, Axel Carlier, Lai Xing Ng, Wei Tsang Ooi

    arXiv · 2026

    L2D-SLDS proves sublinear regret against a changing conditional-risk oracle without exploration, when the candidate models are accurate and either the archive separates them or live feedback reveals cost differences.

    7
  • Using Visual Intelligence to Automate Maintenance Task Guidance and Monitoring on a Head-mounted Display

    Lai Xing Ng, Jamie Ng, Keith T. W. Tang, Liyuan Li, Mark Rice, Marcus Wan

    Proceedings of the 2019 5th International Conference on Robotics and Artificial Intelligence · 2019

    An Augmented Reality Visual Intelligence (ARVI) framework, which combines visual perception with cognitive task reasoning to monitor user performance and provide contextualised guidance for maintenance tasks, is presented.

    7
  • Consistent Learning-to-Defer with Expert-Conditional Advice

    Yannis Montreuil, Leina Montreuil, Axel Carlier, Lai Xing Ng, Wei Tsang Ooi

    arXiv · 2026

    5
  • Is Your Driving World Model an All-Around Player?

    Lingdong Kong, Ao Liang, Tianyi Yan, Hongsi Liu, Wesley Yang, Ziqi Huang, Xian Sun, Wei Yin, +15 more

    arXiv · 2026

    WorldLens is introduced, a unified benchmark that measures world-model fidelity across the full spectrum, from pixel quality and 4D geometry to closed-loop driving and human perceptual alignment, through five complementary aspects and 24 standardized dimensions and forms a unified ecosystem for assessing generated worlds not merely by visual appeal, but by physical and behavioral fidelity.

    5
  • EventDrive: Event Cameras for Vision-Language Driving Intelligence

    D H Lu, R K Li, Ao Liang, Lingdong Kong, Wei Yin, Lai Xing Ng, Benoit R. Cottereau, Camille Simon-Chane, +1 more

    arXiv · 2026

    Comprehensive evaluation across diverse tasks shows that event streams provide substantial gains in temporal precision, motion awareness, and robustness, bringing event sensing into the center of driving intelligence.

    2
  • Learning to Remove Lens Flare in Event Camera

    Han, Haiqian, Lingdong Kong, Junmin Li, Ao Liang, Chengtao Zhu, Lyu, Jiacheng, Lai Xing Ng, Xiangyang Ji, +2 more

    arXiv · 2025

    E-Deflare is presented, the first systematic framework for removing lens flare from event camera data, and is designed to design E-DeflareNet, which achieves state-of-the-art restoration performance.

    2
  • Automated Multi-Camera Inspection System for Aircraft

    Mark Rice, Gu Ying, Kelvin Lim, Qing Yu. Hoo, Liyuan Li, Lee Jue Ying, Jacky Jie Wei Tan, Lai Xing Ng, +1 more

    Proceedings of the AAAI Conference on Artificial Intelligence · 2026

    An automated visual inspection system designed to detect defects on the upper surface of an aircraft airframe using a multi-camera PTZ set-up to capture and process images from designated regions is presented.

    –
  • SpikeCLR: Contrastive Self-Supervised Learning for Few-Shot Event-Based Vision using Spiking Neural Networks

    Maxime Vaillant, Axel Carlier, Lai Xing Ng, Christophe Hurter, Benoit R. Cottereau

    arXiv · 2026

    SpikeCLR is introduced, a contrastive self-supervised learning framework that enables SNNs to learn robust visual representations from unlabeled event data and shows that learned representations transfer across datasets, contributing to efforts for powerful event-based models in label-scarce settings.

    –
  • The RoboSense Challenge: Sense Anything, Navigate Anywhere, Adapt Across Platforms

    Lingdong Kong, Shaoyuan Xie, Zeying Gong, Ye Li, Meng Chu, Liang, Ao, Yuhao Dong, Tianshuai Hu, +22 more

    arXiv · 2026

    –
  • MemoVision: A Digital Catalog for Everyday Interactions

    Lai Xing Ng, Keith T. W. Tang, Jacky Jie Wei Tan

    Proceedings of the AAAI Conference on Artificial Intelligence · 2026

    MemoVision is presented, a digital catalog system that captures semantic, spatial, temporal and interaction information as users move around physical environments using client devices such as smart glasses, enabling more contextualized responses compared to current multimodal large language models.

    –
  • Toward On-Chip Training of Spiking Neural Networks for Dense Event-Based Vision

    Maxime Vaillant, Axel Carlier, Lai Xing Ng, Christophe Hurter, Benoit R. Cottereau

    arXiv · 2026

    –
  • A Query Is Not a Commitment: Learning to Correct Expert Answers in Online Deferral

    Yannis Montreuil, Axel Carlier, Lai Xing Ng, Wei Tsang Ooi

    arXiv · 2026

    –
  • Interactive Content Retrieval in Egocentric Videos Based on Vague Semantic Queries

    Linda Ablaoui, Wilson E. Marcílio-Jr, Lai Xing Ng, Christophe Jouffrais, Christophe Hurter

    Multimodal Technologies and Interaction · 2025

    This work investigates the requirements for an egocentric video content retrieval framework that helps users handle vague queries and proposes a zero-shot, user-centered video content retrieval framework that leverages a VLM to provide video data and query representations that users can incrementally combine to refine queries.

    –
  • DVSim : un simulateur de vision événementielle pour l'apprentissage de réseaux de neurones à impulsions

    Maxime Vaillant, Axel Carlier, Benoit R. Cottereau, Lai Xing Ng, Christophe Hurter

    HAL (Le Centre pour la Communication Scientifique Directe) · 2025

    –
  • One-Stage Top-$k$ Learning-to-Defer: Score-Based Surrogates with Theoretical Guarantees

    Yannis Montreuil, Axel Carlier, Lai Xing Ng, Wei Tsang Ooi

    arXiv · 2025

    –
  • Talk2Event: Grounded Understanding of Dynamic Scenes from Event Cameras

    Lingdong Kong, Dongyue Lu, Ao Liang, Rong Li, Yuhao Dong, Tianshuai Hu, Lai Xing Ng, Wei Tsang Ooi, +1 more

    neural information processing systems · 2025

    –
  • A Two-Stage Learning-to-Defer Approach for Multi-Task Learning

    Yannis Montreuil, Yeo, Shu Heng, Axel Carlier, Lai Xing Ng, Wei Tsang Ooi

    arXiv · 2024

    –

Publication data from OpenAlex, with missing venues and authors filled in from Crossref; citation counts are the higher of OpenAlex and Semantic Scholar, last synced 2026-10-11. One-sentence summaries under some papers are written by Semantic Scholar’s model. Citation counts may be lower than on Google Scholar, which indexes more sources.

Report an error

Wrong papers, two people merged into one, or a profile that should not be here? Tell us and we will fix or hide it.

Sign in to report