'Make knowledge free for everyone'

Experimental

Use this llama.cpp branch: https://github.com/csabakecskemeti/llama.cpp/tree/instella-moe

Quantized version of: amd/Instella-MoE-16B-A3B-Think

Buy Me a Coffee at ko-fi.com

Downloads last month
1,302
GGUF
Model size
16B params
Architecture
instella-moe
Hardware compatibility
Log In to add your hardware

2-bit

3-bit

4-bit

5-bit

6-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for DevQuasar/amd.Instella-MoE-16B-A3B-Think-GGUF

Quantized
(2)
this model