AI4Finance-Foundation / AI4Finance-Foundation/FinRL
Unpredictable Rewards in Stable Baseline 3 DDPG and TD3 Models - Seeking Clarification
- Lingua principale
- Jupyter Notebook
- Stelle
- 16.3k
- Fork
- 3.5k
- Metriche di merge delle PR
- Nessuna PR unita negli ultimi 30g
Descrizione
Hello,
Thank you for creating the library, and I appreciate your excellent work.
I've been experimenting with the Stable Baseline 3 DDPG and TD3 Models. When I run the training script, I'm experiencing unpredictable rewards, sometimes they calculate correctly, and other times they stay at 0. If I stop and rerun the script, the rewards may still be 0. Could you clarify if this is an inherent issue with the models or if I should review the code?
I'm using one week of 1-minute data for a single stock in my training.
Best regards
Guida per i contributori
Nessuna guida per i contributori indicizzata per questo repository
Valutazione
Questa issue non è ancora stata valutata.