Mayank Maheshwari
mayankiit04
AI & ML interests
None yet
Recent Activity
new activity 1 day ago
unsloth/Qwen3.8-Flash-Next-GGUF:can we have a gguf varity where ngram layer 2 is at q8? new activity 1 day ago
unsloth/Qwen3.8-Flash-Next-GGUF:How can I improve the prefilling speed for this model? new activity 3 days ago
tencent/Hy4-preview:With Qwen and GLM now in sub 300B local machine level, this is a big model.Organizations
None yet
can we have a gguf varity where ngram layer 2 is at q8?
#53 opened 1 day ago
by
mayankiit04
How can I improve the prefilling speed for this model?
14
#49 opened 2 days ago
by
BipedalBit
With Qwen and GLM now in sub 300B local machine level, this is a big model.
#10 opened 3 days ago
by
mayankiit04
Is the n-gram chunk embedded in the gguf(s), and is it ~51GB independent of quantization?
14
#35 opened 5 days ago
by
dagb
How to keep n-gram table on fast nvme ssd
18
#23 opened 6 days ago
by
mayankiit04
here qwen3.8-flash-next is 125B but unsloth has 180B ... how come?
2
#17 opened 6 days ago
by
mayankiit04
Are you serious, Chatgpt puts this model at estimate 59 score AA above opus 4.8!!
#3 opened 11 days ago
by
mayankiit04
ArtificialAnalysis score of 52 outscore GLM 5.2 , Opus 4.6 and touches Opus 4.7!!!!
🔥 1
10
#143 opened 13 days ago
by
mayankiit04
Qwen 3.8-27b context usage is about 10x of 3.6-27b
👍 4
8
#45 opened 18 days ago
by
WhiteDan64
Stable MTP first release!
❤️ 15
11
#6 opened 4 months ago
by
danielhanchen
Is it possible to only download the mtp gguf (<1GB one) to use with existing ggufs?
4
#3 opened 4 months ago
by
CHNtentes
Does this perform in comparision to base b16 quantized model ?
1
#1 opened 4 months ago
by
mayankiit04
Will there be a smaller model like Qwen3.5 122 or Nemotron 3 super
➕ 7
7
#9 opened 4 months ago
by
mayankiit04
Why this release?
👍🤯 9
4
#3 opened 4 months ago
by
neoOpus
Qwen3.6 27B vs Qwen 3.6 35B A3B? Which one to go with
🔥 1
2
#12 opened 4 months ago
by
mayankiit04
This is the most underreported model when it comes to agentic coding and intelligence
➕🚀 4
3
#35 opened 5 months ago
by
mayankiit04
Not getting proper ui due to formatting on kilo code where model is run on llama.cpp
#1 opened 6 months ago
by
mayankiit04