Hello,
I know it is very little memory, but it is what I have by now.
By default, the demo code won't inference because of cuda out of memory. I tried to reduce the batch size of the inference to just 1, but is not enough.
Do you know a way to reduce the memory consumption running the inference?
I know that the best solution is to upgrade the GPU to a RTX 3090/4090/A6000, but before that I would like to try another way if possible.
Thank you!
David Martin Rius
Hello,
I know it is very little memory, but it is what I have by now.
By default, the demo code won't inference because of cuda out of memory. I tried to reduce the batch size of the inference to just 1, but is not enough.
Do you know a way to reduce the memory consumption running the inference?
I know that the best solution is to upgrade the GPU to a RTX 3090/4090/A6000, but before that I would like to try another way if possible.
Thank you!
David Martin Rius