EleutherAI / EleutherAI/minetest-baselines
Finetune VPT model for Minetest
- Dominant language
- Python
- Stars
- 9
- Forks
- 2
- PR merge metrics
- No merged PRs in 30d
Description
Finetuning would consist in a combination of
- imitation learning
- reinforcement learning
VPT is a recurrent model, which requires special treatment of the hidden state
Since VPT was developed for Minecraft it's action space has to be suitable mapped to that of Minetest.
Contributor guide
No contributing guide indexed for this repository
Research direction
The issue names no files, tests, or entry points. Start by locating the VPT recurrent model and the Minetest action-space mapping, then determine how imitation learning, reinforcement learning, and hidden-state handling fit together. The issue does not define a concrete completion criterion, so the scope and expected finetuning result need clarification.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- game-dev, machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100