Is there an existing issue for this problem?
Install method
Invoke's Launcher
Operating system
Linux
GPU vendor
Nvidia (CUDA)
GPU model
RTX 3050 Laptop
GPU VRAM
4
Version number
6.3.0
Browser
No response
System Information
No response
What happened
I have enabled the recommended "enable_partial_loading: true", tried enabling "pytorch_cuda_alloc_conf: "backend:cudaMallocAsync"" and setting device_working_mem_gb: 3", nothing helps, I keep getting OOM errors when using more than 3 reference images, sometimes (randomly) when only 2 of them.
What you expected to happen
I was hoping the enable_partial_loading option would help with offloading VRAM, as it had previously worked perfectly, so i'll be able to use more than 2 reference images for the Flux Kontext.
How to reproduce the problem
No response
Additional context
No response
Discord username
No response
Is there an existing issue for this problem?
Install method
Invoke's Launcher
Operating system
Linux
GPU vendor
Nvidia (CUDA)
GPU model
RTX 3050 Laptop
GPU VRAM
4
Version number
6.3.0
Browser
No response
System Information
No response
What happened
I have enabled the recommended "enable_partial_loading: true", tried enabling "pytorch_cuda_alloc_conf: "backend:cudaMallocAsync"" and setting device_working_mem_gb: 3", nothing helps, I keep getting OOM errors when using more than 3 reference images, sometimes (randomly) when only 2 of them.
What you expected to happen
I was hoping the enable_partial_loading option would help with offloading VRAM, as it had previously worked perfectly, so i'll be able to use more than 2 reference images for the Flux Kontext.
How to reproduce the problem
No response
Additional context
No response
Discord username
No response