facebookresearch / facebookresearch/ScaDiver

Questions about training.

Open
#8 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
165
Forks
19
PR merge metrics
No merged PRs in 30d

Description

Hi,
I trained on the single test motion clip `data/motion/test.bvh` with default configuration in `test_env_humanoid_imitation.yaml`. The total reward does not increase after 12 hours training. However, the character still fall. I wonder if there are some parameters in configuration need to be tuned.

commond:
`python rllib_driver.py --mode train --spec data/spec/test_env_humanoid_imitation.yaml --project_dir ./`

reward:
![Screenshot 2021-05-27 202821](https://user-images.githubusercontent.com/32760532/119825954-1cf0b200-bf2a-11eb-933f-13bd51519f6e.png)

test results:
![2](https://user-images.githubusercontent.com/32760532/119826714-e5363a00-bf2a-11eb-9386-bc8339fc1ee6.gif)

Thanks !

Contributor guide

Open the contributing guide

Research direction

Reproduce the report with data/motion/test.bvh, data/spec/test_env_humanoid_imitation.yaml, and the command using rllib_driver.py. Inspect the training output and reward plot alongside the falling-character result; the issue provides no parameter target or defined success condition beyond improved reward and a character that does not fall.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
game-dev, machine-learning
Issue type
Bug
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.