facebookresearch / facebookresearch/SpinQuant

CUDA out of memory during Training

Open
#43 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
428
Forks
93
PR merge metrics
No merged PRs in 30d

Description

I've changed your training code to get optimal rotation matrix only by setting groupsize for weight and activations to 32. However, I consistently encounter CUDA out of memory error during forward pass. I personally think that this is due to large activations memory required during forward pass. I 've also included your FSDP config file. Could you give me an explanation for this?

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.