Skip to content

Question about the relation between standing mask and zero velocity commands #5

Description

@love1994love

I am studying how you can achieve walking and standing in one policy. When I study your code, I find that in some cases the velocity might be zero due to "set small commands to zero” but the standing mask might be false. However, the standing mask does not belong to the observations. This means that when the robot receives zero velocity command, its reward will be very different between true standing mask and false standing mask even uses the same actions. Is this a problem when we train the robot? Thanks.

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions