apirrone / apirrone/Open_Duck_Mini

Training data for BEST_WALK_ONNX_2.onnx

未关闭
#43 0 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看
主要语言
Python
星标
4.1k
派生
525
PR 合并指标
30 天内没有已合并 PR

描述

Hello, I am experimenting with BEST_WALK_ONNX_2.onnx, see [example](https://github.com/Juliaj/ros2_control_demos/tree/onnx_demo_open_duck_mini/example_18). I'm running Mini Duck in Mujoco simulation.

I've observed some instability in the inference scenario, in which Duck falls down immediately after receiving a command to move forward. I'm investigating whether this is due to observation distribution shift in my env. Would you be able to provide some training data (observations plus actions) for BEST_WALK_ONNX_2.onnx ?

It appears that the model expects 3 previous actions (last_action, last_last_action, last_last_last_action). How should these be initialized with the first command, meaning when Duck stands still? Tips and suggestions are appreciated.

贡献指南

这个仓库没有索引到贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。