(It is not a software issue, I'll delete the issue if it's not relevant.)
Dear Carroll and Peter,
I am a Masters in Robotics student at Oregon State University. I was dabbling with SafeLife for a while and had some doubts.
"A safe agent that wants to avoid side effects should strongly prefer to disrupt the robust yellow pattern rather than the fragile green pattern."
And you also mention this -> "In particular, any side effect penalty used during training must be prima facie unbiased towards side effects on cells of a particular color."
Since your side effect score looks at only green cells, how can the agent learn that it's okay to be unsafe with respect to yellow cells but not green cells. Hope you get what I mean! I've been scratching my head over this for a while now. Hope to hear from you soon.
Best Regards,
(It is not a software issue, I'll delete the issue if it's not relevant.)
Dear Carroll and Peter,
I am a Masters in Robotics student at Oregon State University. I was dabbling with SafeLife for a while and had some doubts.
"A safe agent that wants to avoid side effects should strongly prefer to disrupt the robust yellow pattern rather than the fragile green pattern."
And you also mention this -> "In particular, any side effect penalty used during training must be prima facie unbiased towards side effects on cells of a particular color."
Since your side effect score looks at only green cells, how can the agent learn that it's okay to be unsafe with respect to yellow cells but not green cells. Hope you get what I mean! I've been scratching my head over this for a while now. Hope to hear from you soon.
Best Regards,