154 MB
jasperan's picture
Add 4-bit-CPU-path LoRA GGUF (language_model tensors, f16) for llama.cpp inference
78371dd verified