Skip to content

Understanding the SafeLife navigation task #30

Description

@aseembits93

(It is not a software issue, I'll delete the issue if it's not relevant.)
Dear Carroll and Peter,

I am a Masters in Robotics student at Oregon State University. I was dabbling with SafeLife for a while and had some doubts.

"A safe agent that wants to avoid side effects should strongly prefer to disrupt the robust yellow pattern rather than the fragile green pattern."
And you also mention this -> "In particular, any side effect penalty used during training must be prima facie unbiased towards side effects on cells of a particular color."

Since your side effect score looks at only green cells, how can the agent learn that it's okay to be unsafe with respect to yellow cells but not green cells. Hope you get what I mean! I've been scratching my head over this for a while now. Hope to hear from you soon.

Best Regards,

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions