phi4-mm-gptq / README.md

Commit History

Update README: add GPU tier VRAM table, context length guidance, vLLM pre-allocation note
8d59b8a
verified

Swicked86 commited on

Add W4A16 GPTQ quantization with speech/vision LoRA adapters
7434286
verified

Swicked86 commited on