Senior Staff Engineer at the Huawei Singapore Norbert Wiener Research Centre, working at the intersection of Embodied AI, reinforcement learning, and robotics. My research focuses on multi-task learning and world models for long-horizon robotic tasks, with the goal of developing intelligent agents that can reason, adapt, and act reliably in complex environments.
Senior Staff Engineer at Huawei Singapore Norbert Wiener Research Centre. Building learning systems for long-horizon robotic tasks, with a focus on multi-task learning, world models, and scalable embodied intelligence.
Advancing reasoning, generalization, and long-horizon control for embodied agents. I am interested in how learned world models and reusable skills can help robots plan, coordinate, and recover across diverse tasks.
My research connects learning, reasoning, and robotic autonomy so embodied agents can plan and recover across complex real-world tasks.
A framework that aligns remote sensing imagery with ground-level visual priors to improve robotic search efficiency using test-time adaptation.
Research on generating 3D maps for complex indoor environments that do not follow the Manhattan world assumption and using a sparse LiDAR sensor.
This project develops a high-fidelity framework for embodied object navigation by leveraging incremental 3D Scene Graphs and foundational Vision-Language Models (VLMs). By moving beyond flat occupancy maps, the system builds a hierarchical semantic representation of the world that captures objects, rooms, and their relationships. The system utilizes Knowledge-Augmented Generation (KAG) to predict structural and semantic properties of unobserved regions, enabling robots to perform complex, cross-modal search missions based on categories, natural language descriptions, or visual exemplars in completely unknown topographies.
This project explores the integration of foundational Vision-Language Models (VLMs) and Large Language Models (LLMs) to redefine the cognitive architecture of intelligent navigation. We investigated a framework that leverages the zero-shot reasoning capabilities of internet-scale foundational models alongside a structured spatial-semantic memory. This enables embodied agents to perform complex, language-driven semantic search tasks—such as finding specific objects described in everyday natural language—by reasoning over environmental uncertainty and past observations in completely novel environments.
This project introduces MARVEL (Multi-Agent Reinforcement Learning for Constrained Field-of-View Multi-Robot Exploration), a framework for high-performance, decentralized coordination in large-scale environments. By leveraging Graph Attention mechanisms, MARVEL enables robot teams to reason about teammate intent and spatial dependencies under restricted sensing constraints. Our approach focuses on information-theoretic action pruning to optimize coverage and mission efficiency, facilitating complex collaborative maneuvers in completely unknown topographies without a central controller.
This project develops a decentralized, learning-based framework for visibility-based pursuit-evasion in challenging outdoor environments. We focus on enabling teams of mobile agents to systematically clear contaminated spaces and capture adversarial evaders within high-density urban terrains. By integrating multi-agent reinforcement learning with advanced spatial reasoning, the system addresses the critical challenges of building-induced occlusions and limited sensor ranges, allowing for real-time coordinated maneuvers without the need for a central controller.
This project introduces STAR (Swarm Technology for Aerial Robotics), a modular, open-source infrastructure designed to bridge the gap between simulation and high-fidelity physical deployments. STAR integrates decentralized task allocation with robust vision-based landmark localization to manage fleets of nano-quadrotors (e.g., Crazyflies) in cluttered environments. The framework provides a high-throughput ROS 2-based communication layer and a hardware-in-the-loop (HIL) sim-to-real pipeline, enabling researchers to validate complex multi-agent algorithms, reactive obstacle avoidance, and swarm behaviors on physical robotic collectives.
Don't have time to read a 20-page paper right now? Upload any research PDF to instantly generate a beautifully structured, comprehensive summary powered by Google Gemini.
Try the Tool