AI4Finance-Foundation / AI4Finance-Foundation/FinRL
Unpredictable Rewards in Stable Baseline 3 DDPG and TD3 Models - Seeking Clarification
未關閉
discussion
- 主要語言
- Jupyter Notebook
- 星號
- 16.3k
- 分支
- 3.5k
- PR 合併指標
- 30 天內沒有已合併 PR
描述
Hello,
Thank you for creating the library, and I appreciate your excellent work.
I've been experimenting with the Stable Baseline 3 DDPG and TD3 Models. When I run the training script, I'm experiencing unpredictable rewards, sometimes they calculate correctly, and other times they stay at 0. If I stop and rerun the script, the rewards may still be 0. Could you clarify if this is an inherent issue with the models or if I should review the code?
I'm using one week of 1-minute data for a single stock in my training.
Best regards
貢獻指南
這個儲存庫沒有索引到貢獻指南
評估
這個 Issue 還沒有評估資料。