JdeRobot / JdeRobot/BehaviorMetrics

Introduce Continuous Control F1RL Environment

Open
#140 3 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
64
Forks
52
PR merge metrics
No merged PRs in 30d

Description

Current Setup:
DQN and Q-learning are already done with Discrete Control Actions

To be introduced:
Policy Gradient Algorithm with Continuous Control Actions

Major Changes:
1. Add a continuous control environment derived from DQN environment by Change the action space
2. Add a brain to the f1rl module to execute the actions

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.