Can I deploy it on the 4*RTX A6000 GPU?

#3
by Yuejuan - opened

I plan to use the VLLM framework to run this model. My GPU consists of 4 RTX A6000 cards, each with 48 GB of GDDR6 memory (totaling 192 GB), and the NVIDIA Ampere architecture (SM 8.6).Is it appropriate?

Sign up or log in to comment