apirrone / apirrone/Open_Duck_Mini

Training data for BEST_WALK_ONNX_2.onnx

オープン
#43 コメント 0 件 リアクション 0 件 担当者 0 名 GitHub で見る
主要言語
Python
スター
4.1k
フォーク
525
PR マージ指標
30日以内にマージされた PR はありません

説明

Hello, I am experimenting with BEST_WALK_ONNX_2.onnx, see [example](https://github.com/Juliaj/ros2_control_demos/tree/onnx_demo_open_duck_mini/example_18). I'm running Mini Duck in Mujoco simulation.

I've observed some instability in the inference scenario, in which Duck falls down immediately after receiving a command to move forward. I'm investigating whether this is due to observation distribution shift in my env. Would you be able to provide some training data (observations plus actions) for BEST_WALK_ONNX_2.onnx ?

It appears that the model expects 3 previous actions (last_action, last_last_action, last_last_last_action). How should these be initialized with the first command, meaning when Duck stands still? Tips and suggestions are appreciated.

コントリビューションガイド

このリポジトリのコントリビューションガイドは索引されていません

評価

この issue はまだ評価されていません。

新しい issue をメールで受け取る

初心者向けの GitHub issue を短くまとめたダイジェスト。