Comfy-Org / Comfy-Org/ComfyUI

Very High VRAM usage when using lora with flux

Open
#4,681 14 comments 1 reaction 0 assignees View on GitHub
Potential Bug
Dominant language
Python
Stars
133k
Forks
15.7k
Avg merge
1d 7h
Merged PRs (30d)
158

Description

### Expected Behavior

Not 10Gb Vram eaten using the lora.

### Actual Behavior

I have flux fp8 schnell on a 3090, I run two loras rank 64 onto the model, but it uses all VRAM until it starts offloading and generations of course slow down.

### Steps to Reproduce

Just add two 64 rank lora onto flux schnell fp8

### Debug Logs

```powershell
None
```

### Other

_No response_

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.