AI4Finance-Foundation / AI4Finance-Foundation/FinRL

Unpredictable Rewards in Stable Baseline 3 DDPG and TD3 Models - Seeking Clarification

Aperta
#1,138 2 commenti 0 reazioni 1 assegnatario Rivendicata da @zhumingpassional Vedi su GitHub
discussion
Lingua principale
Jupyter Notebook
Stelle
16.3k
Fork
3.5k
Metriche di merge delle PR
Nessuna PR unita negli ultimi 30g

Descrizione

Hello,

Thank you for creating the library, and I appreciate your excellent work.

I've been experimenting with the Stable Baseline 3 DDPG and TD3 Models. When I run the training script, I'm experiencing unpredictable rewards, sometimes they calculate correctly, and other times they stay at 0. If I stop and rerun the script, the rewards may still be 0. Could you clarify if this is an inherent issue with the models or if I should review the code?

I'm using one week of 1-minute data for a single stock in my training.

Best regards

Guida per i contributori

Nessuna guida per i contributori indicizzata per questo repository

Valutazione

Questa issue non è ancora stata valutata.

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.