Skip to content

Experiments on cheetah-dir and ant-dir #2

Description

@Lagrant

Hi,

I've tried normalizing environments, revising reward functions, upgrading/downgrading MuJoCo versions, but still not able to reproduce the performance declared in your paper on ant-dir. The average return just fluctuates at a very low level around 10. Besides, experiments on cheetah-dir get an average training return of 1300 but an average testing return of -1400 which never happens on other enviroments. There seems to be something wrong with the environment. Could you also check it out?

My experiment logs, models and configurations are uploaded to google file for your reference.

Any reply would be much appreciated!

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions