About Me

Conceptual human-robot interaction research illustration

My Work

I am an Early Research Career Fellow at Technology, CSIRO, Australia, working in the Human-Robot Interaction team. My research focuses on deep reinforcement learning, reward shaping, and robotic manipulation, with an emphasis on helping learning agents acquire robust behavior from limited interaction.

During my Ph.D., I worked on adaptive potential functions for reinforcement learning, exploring how task guidance and value estimation can help agents learn more efficiently. My current work studies how robotic imitation learning can generalize better under distribution shifts. In particular, I investigate how policies can reduce causal confusion in observations by focusing on task-relevant information and avoiding spurious correlations in demonstrations. This aims to make learned robot policies more reliable when the environment changes, such as under different backgrounds, distractors, object appearances, or task conditions.

My Experiences

1/2024 - present

Early Research Career Fellow

CSIRO Technology, Australia.

  • Human-Robot Interaction team.
  • Research topic: robotic manipulation, imitation learning, causal confusion, and reinforcement learning.
10/2018 - 4/2025

Ph.D Degree in Deep Reinforcement Learning

University of Groningen, The Netherlands.

  • Thesis: Beyond Value Estimation: Adaptive Potential Functions For Reinforcement Learning.
  • Research topic: deep reinforcement learning and reward shaping.
9/2015 - 6/2018

Master of Engineering Degree in Visible Optical Communication

Zhejiang University, China.

  • Research topic: Underwater Wireless Optical Communication.
  • Publish one journal paper in Optics Express (2019 IF: 3.669) as the 1st author.
9/2011 - 6/2015

Bachelor Degree in Electrical Information Engineering

Nanjing University of Posts and Telecommunications, China.

  • Research topic: Speed control of DC motors based on PID algorithm.

My Publications

* corresponding author.
Boosting Reinforcement Learning Algorithms in Continuous Robotic Reaching Tasks Using Adaptive Potential Functions Yifei Chen*, Lambert Schomaker, Francisco Cruz AJCAI 2024
Improving Proximal Policy Optimization Algorithm in Interactive Multi-Agent Systems Yi Shang, Yifei Chen, Francisco Cruz ICDL 2024, pp. 1-6
Reinforcement Learning with Potential Functions Trained to Discriminate Good and Bad States Yifei Chen*, Hamidreza Kasaei, Lambert Schomaker, Marco Wiering IJCNN 2021
An Investigation into the Effect of the Learning Rate on Overestimation Bias of Connectionist Q-learning Yifei Chen*, Lambert Schomaker, Marco Wiering ICAART 2021
Improving Generalization Ability of Robotic Imitation Learning by Resolving Causal Confusion in Observations Yifei Chen, Yuzhe Zhang, Giovanni Durso, Nicholas Lawrance, Brendan Tidd In Progress
26 m/5.5 Gbps Air-Water Optical Wireless Communication Based on an OFDM-Modulated 520-nm Laser Diode Yifei Chen, M. W. Kong, T. Ali, J. L. Wang, R. Sarwar, J. Han, C. Y. Guo, B. Sun, N. Deng, J. Xu Optics Express, 2017

Contact Me

If you have any questions or comments, please don't hesitate to contact me.

Mon - Fri 09:00 – 18:00