首页 /研究 /Too Much of a Good Thing: When sim2real Efforts Impede Policy Learning (And What to Do About It)

LEARNING

Too Much of a Good Thing: When sim2real Efforts Impede Policy Learning (And What to Do About It)

Kyle Morgenstein, Bharath Masetty, Stephen Welch, Luis Sentis

发表年份: 2026
访问权限: 开放获取

摘要

While sim2real efforts are necessary for effective policy transfer to hardware, there is such a thing as too much of a good thing. We argue that sim2real efforts have led to misaligned incentives with policy learning, resulting in simulator lock in and poor policy exploration due to the unreasonable constraints imposed by the real world. We offer a diagnosis and explanation of the current status of the problem, and propose a potential solution via a sim2sim2real paradigm that leverages the robot's kinematics as the sole design constraint.

关键词

cs.ROcs.AI

Too Much of a Good Thing: When sim2real Efforts Impede Policy Learning (And What to Do About It)

摘要

关键词

相关论文

The Organization of Behavior

Fractional Brownian Motions, Fractional Noises and Applications

Review of deep learning: concepts, CNN architectures, challenges, applications, future directions

A guide to deep learning in healthcare