Commit History

Add Q8_0 and Q4_K_M GGUF quantizations with custom llama.cpp runtime
e57b37d
verified

HCHs commited on