首页 /研究 /Enhancing Reinforcement Learning via Causally Correct Input Identification and Targeted Intervention
LEARNING

Enhancing Reinforcement Learning via Causally Correct Input Identification and Targeted Intervention

Jiwei Shen, Lu Hu, Hao Zhang, Shujing Lyu, Yue Lu

发表年份
2024
引用次数
2

摘要

Causal confusion, characterized by the learning of spurious correlations, detrimentally affects the generalization and effectiveness of reinforcement learning (RL) algorithms, especially in environments without latent confounders often encountered in robot autonomous navigation tasks. This study addresses this gap by developing a causal structure within a Partially Observable Markov Decision Process (POMDP). Subsequently, we introduce a targeted intervention that mitigates the influence of spurious correlations by isolating causally significant state variables and discarding irrelevant inputs. Testing in three real-world scenarios confirms the approach’s feasibility and superiority in enhancing the RL algorithms’ performance and generalization ability, signifying a promising step towards more robust online RL frameworks.

关键词

Reinforcement learningIdentification (biology)Computer scienceIntervention (counseling)ReinforcementArtificial intelligenceMachine learningPsychologySocial psychology

相关论文

查看 LEARNING 分类全部论文