How to use from
Docker Model Runner
docker model run hf.co/TransformerTales/llama-2-7b-8bit-nested
Quick Links

I used Google Colab to quantize/nest the Llama 2 7B model. Should help out those who wish to use Llama 2 7B on a low-end computer. GPU is still recomended...

Downloads last month
6
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support