ikostrikov / ikostrikov/implicit_q_learning

Code for Behavior cloning policy

Open
#11 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
337
Forks
49
PR merge metrics
No merged PRs in 30d

Description

Could you please provide the code implementation related to BC in Table 1 of the paper? It looks like it gets great performance in walker2d-medium-expert-v2 dataset and is better than BC in other papers, e.g., Transformer Decision. Some description of the implementation would also be very useful for me. Thank you very much for your outstanding work.

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by comparing the BC result in Table 1 of the paper with the Python code in this repository. Determine whether the requested implementation is absent and what implementation details can be documented. Done means providing the BC code and an accompanying explanation for the reported walker2d-medium-expert-v2 performance.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.