From the source
Running the Vicuna 13B open-source chatbot model on a single AMD GPU using ROCm, including quantization with GPTQ to reduce memory footprint.
From the source
From the source

Running the Vicuna 13B open-source chatbot model on a single AMD GPU using ROCm, including quantization with GPTQ to reduce memory footprint.
From the source
In this blog, we will delve into the world of Vicuna, and explain how to run the Vicuna 13B model on a single AMD GPU with ROCm.
huggingface.co