JuliaPOMDP / JuliaPOMDP/BasicPOMCP.jl

Average reward after learning a strategy

Open
#15 4 comments 0 reactions 0 assignees View on GitHub
Dominant language
Julia
Stars
39
Forks
18
PR merge metrics
No merged PRs in 30d

Description

Hello I used BasicPOMCP to find optimal strategy in quite large game. I used example to calculate 10000 tree queries, but even tho i see the tree, I am mostly interested in average reward. I know there is function simulate, however i feel like results from this method vary more than i expect (but maybe taking n simulations and then do some kind of average is a good solution).
Simply put is it possible to get average reward immediately after solver solves a game?

Thank you for response

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.