pathosethoslogos
pathosethoslogos
·
AI & ML interests
None yet
Recent Activity
new activity about 2 hours ago
QUASAR-QAT/Qwen3.8-27B-QUASAR-NVFP4:This model keeps quitting mid-way and/or doesn't answer new activity about 12 hours ago
QUASAR-QAT/Qwen3.8-27B-QUASAR-NVFP4:Quant request?Organizations
None yet
This model keeps quitting mid-way and/or doesn't answer
#6 opened about 2 hours ago
by
pathosethoslogos
Quant request?
#5 opened about 12 hours ago
by
pathosethoslogos
One-command deployment on a single DGX Spark (34-42 tok/s, prefix caching, 262K)
👍🔥 2
5
#7 opened 5 days ago
by
hasanbasbunar
Any plans for merging into mainline vLLM?
➕ 1
2
#16 opened 6 days ago
by
pathosethoslogos
No explanation on what 'bnb' is in the model card
👍 1
1
#2 opened 4 days ago
by
pathosethoslogos
When will this DFlash model be compatible or come to mainline vLLM Git?
#12 opened 7 days ago
by
pathosethoslogos
When will it come to mainline vLLM Git?
#7 opened 8 days ago
by
pathosethoslogos
run on DGX SPARK
4
#23 opened 8 days ago
by
Muyanghao
Is this really a 1B model?
4
#2 opened 16 days ago
by
mindplay
model-00016-of-00016.safetensors was just updated... What?
3
#25 opened 10 days ago
by
pathosethoslogos
</think> every response
5
#36 opened 30 days ago
by
pathosethoslogos
Inferact/Qwen3.8-27B-NVFP4 is on vLLM's official documentation
10
#8 opened 20 days ago
by
pathosethoslogos
The model goes crazy and goes into loops
1
#1 opened 14 days ago
by
pathosethoslogos
Qwen3.8-27B Serving Configs: DGX Spark vLLM NVFP4
🤗🚀 10
7
#7 opened 20 days ago
by
erdal
Does not work with vllm 0.27.1 (latest)
7
#11 opened 19 days ago
by
flaviusburca
Bench maxed -- don't be fooled by the table presented in the model card
👀 1
4
#84 opened 19 days ago
by
pathosethoslogos
unsloth/Qwen3.8-27B-NVFP4 vs. Inferact/Qwen3.8-27B-NVFP4?
🔥➕ 3
10
#1 opened 20 days ago
by
pathosethoslogos
Custom vLLM merge request to main vLLM?
👍 1
7
#6 opened 20 days ago
by
pathosethoslogos
Please submit a PR to vLLM for upstream model support?
👍 1
2
#2 opened 20 days ago
by
GadflyII