google / google/paxml

How to continue training from a checkpoint?

Open
#37 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
560
Forks
72
PR merge metrics
No merged PRs in 30d

Description

I am trying to continue training my model from a checkpoint, using paxml.

Does paxml not support `restore_checkpoint_dir` or `restore_checkpoint_step` for train mode?

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.