AI4Finance-Foundation / AI4Finance-Foundation/FinRL

Unpredictable Rewards in Stable Baseline 3 DDPG and TD3 Models - Seeking Clarification

Aberta
#1,138 2 comentários 0 reações 1 responsável Reivindicada por @zhumingpassional Ver no GitHub
discussion
Linguagem predominante
Jupyter Notebook
Estrelas
16.3k
Forks
3.5k
Métricas de merge de PRs
Nenhum PR com merge em 30d

Descrição

Hello,

Thank you for creating the library, and I appreciate your excellent work.

I've been experimenting with the Stable Baseline 3 DDPG and TD3 Models. When I run the training script, I'm experiencing unpredictable rewards, sometimes they calculate correctly, and other times they stay at 0. If I stop and rerun the script, the rewards may still be 0. Could you clarify if this is an inherent issue with the models or if I should review the code?

I'm using one week of 1-minute data for a single stock in my training.

Best regards

Guia de contribuição

Nenhum guia de contribuição indexado para este repositório

Avaliação

Esta issue ainda não foi avaliada.

Receba novas issues na sua caixa de entrada

Um resumo curto de issues do GitHub para quem está começando.