Learning from demonstration videos
Ongoing research into extracting objects, actions and reusable skills from unstructured human demonstrations, connecting with multimodal grounding and skill learning.
RESEARCH · PROJECTS · EXPLORATIONS
Research, engineering projects and earlier experiments, from 2017 to today.
Research
Grounding “this” and “here” in speech and gestures, so robots can identify what to move and where it belongs.
CIS-RAM 2026 →Industry collaboration
Robotic cable insertion through visual registration and force feedback, in collaboration with Desay SV Singapore.
Research & engineering
A hardware-agnostic vision-servoed system for automated, precise injection of implantable microdevices.
ICRA 2026 →Engineering
Connecting visual-language perception and arm control to find, grasp and organize objects on a cluttered desk.
Publication
Expert Systems with Applications · 294, 128736
Publication
Journal of Manufacturing Systems · 82, 1213–1226
Publication
IEEE ICHMS · 223–228
Engineering
Integrating a robot arm, mobile base and vision sensors into a platform for natural-language-guided manipulation.
Publication
IEEE Transactions on Industrial Informatics · 20(9), 11372–11383
Research
Turning instructions and interaction history into reusable skills for longer, compositional robot tasks.
CIS-RAM 2024 · 14–19 · Best Paper →Publication
IEEE Robotics and Automation Letters · 7(4), 12451–12458
Research & engineering
Proximity-based slowing and stopping, FCL self-collision checks, and an admittance controller that turns obstacle distance into avoidance motion.
Research & engineering
Skeleton-based collision avoidance with Nuitrack and MoveIt, GGCNN grasping and handover, and position-based visual servoing on a UR3e.
Robotics
Ground-to-air robot coordination based on visual navigation.
Publication
IEEE ITAIC · Chongqing, China
Engineering
A Raspberry Pi and STM32 hexapod with 18 servos, ROS interfaces and multimodal sensing for walking, obstacle avoidance and uneven terrain.
Machine learning
Deep reinforcement learning for video games; related work published at IEEE ITAIC 2019.
Human–machine interaction
EEG-based teleoperation with visual signal stimulation.
Ongoing research into extracting objects, actions and reusable skills from unstructured human demonstrations, connecting with multimodal grounding and skill learning.
Engineering work on unloading order, trajectories and suction-based handling, with factory-site integration and demonstration experience in 2023.
A wearable navigation prototype integrating Raspberry Pi, ToF sensing, voice interaction, stereo cameras and ORB-SLAM3.