Eclectic-Sheep / Eclectic-Sheep/sheeprl

enabling self play

Open
#241 4 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
439
Forks
66
PR merge metrics
No merged PRs in 30d

Description

hi,
i tried out this project and it is one of the few that actually works off the shelf, thank you for your work.
Is there a way to enable self play when training an agent? My usecase is to use DreamerV3 as a alternative to algorithms such as muzero to train agents for boardgames.
I have looked around the repo but this feature does not seem trivially available out of the box.

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by locating the DreamerV3 training entry point and the environment setup used during training, then determine how self-play would fit the existing agent workflow. Done would require a clearly defined self-play behavior and validation that it supports the requested board-game training use case.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, pytorch
Domain
machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.