I'm now a Research Scientist at Tencent Singapore, via the Talent Project UP (also known as "青云计划"), working on Agentic AI, in the Frontier Lab of Hunyuan. Before that, I was a Machine Learning Engineer at TikTok Singapore.
I received my Ph.D. degree from National University of Singapore (NUS), advised by Professor Leong Tze Yun. I received my Bachelor's degree from Xi'an Jiaotong University, Qian Xuesen Honors College (钱学森实验班), supervised by Professor Liu Jun.
My research interests include Reinforcement Learning, Reward Shaping, Agentic AI, Large Language Models, and Robotics.
Tech. Report
UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations
Tencent Hy Frontier Team
UI-Mate Technical Report.
NeurIPS 2025
Centralized Reward Agent for Knowledge Sharing and Transfer in Multi-Task Reinforcement Learning
Haozhe Ma, Zhengding Luo, Thanh Vinh Vo, Kuankuan Sima, Tze-Yun Leong.
39th Annual Conference on Neural Information Processing Systems (NeurIPS), 2025
ICML 2025
Catching Two Birds with One Stone: Reward Shaping with Dual Random Networks for Balancing Exploration and Exploitation
Haozhe Ma, Fangling Li, Jing Yu Lim, Zhengding Luo, Thanh Vinh Vo, Tze-Yun Leong.
42nd International Conference on Machine Learning (ICML), 2025
ICLR 2025
Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning
Haozhe Ma, Zhengding Luo, Thanh Vinh Vo, Kuankuan Sima, Tze-Yun Leong.
13th International Conference on Learning Representations (ICLR), 2025
ICML 2024
Reward Shaping for Reinforcement Learning with An Assistant Reward Agent.
Haozhe Ma, Kuankuan Sima, Thanh Vinh Vo, Di Fu, Tze-Yun Leong.
41st International Conference on Machine Learning (ICML), 2024
Preprint
Exploration by Random Reward Perturbation
Haozhe Ma, Guoji Fu, Zhengding Luo, Jiele Wu, Tze-Yun Leong.
AAMAS 2024
Oral
Mixed-Initiative Bayesian Sub-Goal Optimization in Hierarchical Reinforcement Learning.
Haozhe Ma, Thanh Vinh Vo, Tze-Yun Leong.
23rd International Conference on Autonomous Agents and Multiagent Systems (AAMAS), 2024
NN 2024
GFANC-RL: Reinforcement Learning-based Generative Fixed-filter Active Noise Control.
Zhengding Luo*, Haozhe Ma*, Dongyuan Shi, Woon-Seng Gan.
*Equal contribution.
Neural Networks Journal, 2024. (Impact Factor: 6.0)
AAMAS 2023
Hierarchical Reinforcement Learning with Human-AI Collaborative Sub-Goals Optimization.
Haozhe Ma, Thanh Vinh Vo, Tze-Yun Leong.
22rd International Conference on Autonomous Agents and Multiagent Systems (AAMAS), 2023
AAAI Sym. 2023
Human-AI Collaborative Sub-Goal Optimization in Hierarchical Reinforcement Learning.
Haozhe Ma, Thanh Vinh Vo, Tze-Yun Leong.
Inaugural Summer Symposium Series 2023 of AAAI, 2023
SP 2026
Reinforcement learning-based selective fixed-filter active noise control (RL-SFANC): From theory to real-time headphone implementation.
Zhengding Luo, Haozhe Ma , Boxiang Wang, Dongyuan Shi, Woon-Seng Gan.
Signal Processing
IEEE RA-L
MacroNav: Multi-Task Context Representation Learning Enables Efficient Navigation in Unknown Environments.
Kuankuan Sima, Longbin Tang, Zhenyu Yang, Haozhe Ma, Lin Zhao.
IEEE Robotics and Automation Letters
ACL Fdgs 2025
Robustness via referencing: Defending against prompt injection attacks by referencing the executed instruction.
Yulin Chen, Haoran Li, Yuan Sui, Yue Liu, Yufei He, Xiaoling Bai, Chi Fei, Yabo Li, Haozhe Ma, Yangqiu Song, Bryan Hooi.
The 64th Annual Meeting of the Association for Computational Linguistics, 2025
Ph.D. Degree
School of Computing
Computer Science, School of Computing
Master Degree
Computer Science, School of Computing
School of Computing
Bachelor Degree
Computer Science and Engineering
Computer Science, Qian Xuesen Honors College (钱学森实验班)
Chinese Government Award for Outstanding Students Abroad. (中国教育部国家优秀留学生奖学金)
Ministry of Education of China's Chunhui Innovation Outstanding Achievement Award. (中国教育部留学中心春晖创新创业优秀成果奖)
Research Incentive Award of School of Computing of National University of Singapore (AY2022-2023)
Research Scholarship from the Ministry of Education in Singapore (2022-2026).
Scholarship of Annual Conference on Neural Information Processing Systems (NeurIPS 2025).
Scholarship of International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2024).
Outstanding Student Scholarship of Xi'an Jiaotong University.
SiYuan Student Scholarship of Xi'an Jiaotong University.
Student Area Search Committee for faculty recruitment for School of Computing, National University of Singapore.
Reviewers of conferences and journals: ICLR, ICML, NeurIPS, TPAMI, TNNLS, TSP, NN, etc.
Teaching Assistant of the undergraduate course Foundations of Artificial Intelligence.
Teaching Assistant of the graduate course AI Planning and Decision Making.