Fabian Heller
Ununnilium
AI & ML interests
None yet
Organizations
None yet
TP5 RTX 6000?
8
#1 opened 2 months ago
by
jpsequeira
is possible to create Qwen3.6-27B-IQ4_XS-pure-GGUF with MTP support?
3
#1 opened 3 months ago
by
vinimuchulski
Running models with vLLM on the RTX Pro 6000 - SM120
ππ 2
12
#28 opened 3 months ago
by
liku2001
Works good with vLLM, just no tool calling
1
#1 opened 12 months ago
by
Ununnilium
How many GPU Memory AWQ need?
5
#1 opened about 1 year ago
by
hermitg
Does it possible to create a version without MTP layer to save some VRAM
π 1
1
#3 opened about 1 year ago
by
adonishong