EleutherAI / EleutherAI/minetest-baselines

Finetune VPT model for Minetest

Open
#1 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
9
Forks
2
PR merge metrics
No merged PRs in 30d

Description

Finetuning would consist in a combination of
- imitation learning
- reinforcement learning

VPT is a recurrent model, which requires special treatment of the hidden state

Since VPT was developed for Minecraft it's action space has to be suitable mapped to that of Minetest.

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue names no files, tests, or entry points. Start by locating the VPT recurrent model and the Minetest action-space mapping, then determine how imitation learning, reinforcement learning, and hidden-state handling fit together. The issue does not define a concrete completion criterion, so the scope and expected finetuning result need clarification.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
game-dev, machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.