apirrone / apirrone/Open_Duck_Mini

Training data for BEST_WALK_ONNX_2.onnx

未關閉
#43 0 則留言 0 個 reaction 已指派 0 人 在 GitHub 檢視
主要語言
Python
星號
4.1k
分支
526
PR 合併指標
30 天內沒有已合併 PR

描述

Hello, I am experimenting with BEST_WALK_ONNX_2.onnx, see [example](https://github.com/Juliaj/ros2_control_demos/tree/onnx_demo_open_duck_mini/example_18). I'm running Mini Duck in Mujoco simulation.

I've observed some instability in the inference scenario, in which Duck falls down immediately after receiving a command to move forward. I'm investigating whether this is due to observation distribution shift in my env. Would you be able to provide some training data (observations plus actions) for BEST_WALK_ONNX_2.onnx ?

It appears that the model expects 3 previous actions (last_action, last_last_action, last_last_last_action). How should these be initialized with the first command, meaning when Duck stands still? Tips and suggestions are appreciated.

貢獻指南

這個儲存庫沒有索引到貢獻指南

評估

這個 Issue 還沒有評估資料。

把新 issue 寄到你的電子郵件信箱

精選適合新手參與的 GitHub issue 摘要。