Henry
hdnh2006
AI & ML interests
Math, data & AI.
Recent Activity
liked a model 1 day ago
pottokao/Qwen-Image-2.1-DiT-GGUF new activity 4 days ago
pottokao/Qwen-Image-2.1-PE-T2I-Heretic-NVFP4:Can you provide I2I as well? new activity 4 days ago
ModelsLab/Qwen-Image-2.1-W4A4-int4:Any recipe for vLLM?Organizations
Can you provide I2I as well?
1
#1 opened 4 days ago
by
hdnh2006
Any recipe for vLLM?
1
#1 opened 4 days ago
by
hdnh2006
Any chance to release 27B version?
#4 opened 8 days ago
by
hdnh2006
VRAM required for serving with vLLM
#1 opened 9 days ago
by
hdnh2006
GGUF weights and workflow.json missed
#1 opened 13 days ago
by
hdnh2006
RuntimeError: Worker failed with error 'concat_and_cache_mla
3
#3 opened 20 days ago
by
hdnh2006
GGUF will be out ?
5
#4 opened about 1 month ago
by
marikbrest
Fantastic model! We would like to serve this model for free
🔥 1
4
#3 opened about 1 month ago
by
hdnh2006
Could you please provide the recipe used for quantize this model?
1
#1 opened about 2 months ago
by
hdnh2006
MTP is slower than the "normal" serve
4
#10 opened 4 months ago
by
hdnh2006
NVFP4 on RTX 5090: 120k Context & 8-bit KV Cache Feasibility
1
#1 opened 4 months ago
by
nsfilho
processor_config.json file upload for smooth integration with vLLM
#5 opened 4 months ago
by
hdnh2006
Recommend to upload process_config.json from original model
1
#7 opened 4 months ago
by
hdnh2006
This PR uploads the process_config.json file allowing to serve the model using vLLM without issues
#8 opened 4 months ago
by
hdnh2006
llama.cpp version?
1
#5 opened 4 months ago
by
hdnh2006
New activity in DavidAU/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking 4 months ago
vLLM version? I am getting error with some layers
2
#6 opened 4 months ago
by
hdnh2006
After applying patch, the model is unable to serve
#3 opened 4 months ago
by
hdnh2006
This model is not uncensored at all
1
#2 opened 4 months ago
by
hdnh2006
Why is it too big?
➕ 3
9
#1 opened 5 months ago
by
alexcardo
MTP?
➕ 4
3
#3 opened 4 months ago
by
tyapo