Running on <4gb Vram
#3
by Jit2024 - opened
Hi, I found out a way of running this model on <4gb Vram. And I have not used any quantization.
Repo: https://github.com/Jit-Roy/WeeLLM
thanks for sharing; diffusers still a good choice
Hi, I found out a way of running this model on <4gb Vram. And I have not used any quantization.
Repo: https://github.com/Jit-Roy/WeeLLM
thanks for sharing; diffusers still a good choice