Responsible for the core architecture design and world model design of this open-source Vision-Language-Action platform.
About Me
I am a Ph.D. candidate in Computer Science at The Hong Kong University of Science and Technology (Guangzhou), advised by Prof. Hui Xiong. I am currently working on Safe Embodied AI Pre-training within Ant Group's Security Department. Previously, I was a Research Intern and Co-Lead of core VLA research at AI² Robotics.
My research focuses on Embodied AI, particularly Vision-Language-Action models, spatial perception, understanding, and modeling. I aim to build generalizable robot intelligence that can remember, reason, and act reliably in real-world environments.
Education


Industry



Honors & Awards
- Outstanding Graduate, Shenzhen University2024
- National Scholarship, Ministry of Education2023
- Reviewer for CV, robotics, and ML venuesOngoing
Research Interests
- Vision-Language-Action Models
- Embodied AI & Robot Learning
- Spatial Memory & World Models
- 3D and Event-based Perception
News
SOMA is accepted to ICML 2026.
AlphaBrain reaches 180+ GitHub stars.
See&Trek is accepted to NeurIPS 2025.
Joined AI² Robotics as a Research Intern.
Implicit Grasp Diffusion is accepted to CoRL 2024.
Robot Trajectron is accepted to ICRA 2024.
Selected Publications (view all )
[TMM 2026] DeblurSplat: SfM-Free 3D Gaussian Splatting with Event Camera for Robust Deblurring
Pengteng Li, Y. Lu, Pinhao Song, W. Guo, Huizai Yao, F. Richard Yu, Hui Xiong
[IJCAI 2024] OTOcc: Optimal Transport for Occupancy Prediction
Pengteng Li, Ying He, F. Richard Yu, Pinhao Song, Xingchen Zhou, Guang Zhou
[ACM MM 2023] IGG: Improved Graph Generation for Domain Adaptive Object Detection
Pengteng Li, Ying He, F. Richard Yu, Pinhao Song, Dongfu Yin, Guang Zhou
[ICASSP 2023] Bagging R-CNN: Ensemble for Object Detection in Complex Traffic Scenes
Pengteng Li, Ying He, Dongfu Yin, F. Richard Yu, Pinhao Song



