view article Article Accelerating vision-language models with LFM2.5-VL-DSpark LiquidAI • 6 days ago • 47
Pym Collection Aggressive mixed-precision GGUF quants — shrink models to a fraction of their size, keep their edge and speed. Runs on llama.cpp (AMD/Apple too). • 2 items • Updated Aug 15
Pym Collection Aggressive mixed-precision GGUF quants — shrink models to a fraction of their size, keep their edge and speed. Runs on llama.cpp (AMD/Apple too). • 2 items • Updated Aug 15