Commit History

Update README: add GPU tier VRAM table, context length guidance, vLLM pre-allocation note
8d59b8a
verified

Swicked86 commited on

Add W4A16 GPTQ quantization with speech/vision LoRA adapters
7434286
verified

Swicked86 commited on

initial commit
11f130b
verified

Swicked86 commited on