GGUF
conversational

This is GGUF quantization of BLR2/Qwen3.5-9B-Eagle3-ShareGPT.

You can use it as Eagle3 speculative decoding drafter for any Qwen3.5-9B quantization.

Now that llama.cpp has merged support for Eagle3 in Qwen, you can use any build past b9723.

Downloads last month
115
GGUF
Model size
0.4B params
Architecture
eagle3
Hardware compatibility
Log In to add your hardware

4-bit

5-bit

6-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for dobriak/Eagle3-Qwen3.5-9B

Quantized
(1)
this model