708-145 PRO
TobDeBer
AI & ML interests
Diffusion, Causality, LLM, LMM (Large Music Model), Quantization, AI Context Databases
Recent Activity
updated a collection 3 days ago
Video liked a Space 3 days ago
Viggle/Qwen-Image-2.1-viggle-turbo updated a Space 8 days ago
TobDeBer/LagunaZeroOrganizations
None yet
Great Project!
1
#1 opened 23 days ago
by
ai-god-sai
Could you make a 122B A10B version
16
#115 opened about 1 month ago
by
mbirrell
Is it possible to offload n-gram to an NVMe SSD?
👀🚀 31
15
#11 opened about 1 month ago
by
lingyezhixing
file size (quantization)
2
#16 opened about 1 month ago
by
jacek2024
how IQ1_S is 72Gb :\
12
#5 opened about 1 month ago
by
Darkknight535
Reasoning does not work
10
#6 opened 2 months ago
by
Nerdsking
reap map
2
#12 opened 3 months ago
by
elismasilva
Is it possible to make less than 1 bit quantization?
👍🤗 2
13
#1 opened 3 months ago
by
RealBar
Is using MTP actually slower?
👍 8
5
#4 opened 4 months ago
by
jian2023
UD-TQ1_0 variants of 30B+ llms not seen lately .
1
#28 opened 4 months ago
by
TnK93
minimax 3 什么时候开源?
25
#33 opened 4 months ago
by
wpfnnnns
Gemma 4 MTP assistant/drafter models in GGUF
2
#41 opened 4 months ago
by
redLiw
Model does not support audio
👍 3
5
#1 opened 5 months ago
by
alphamerian
Update app.py
#1 opened 8 months ago
by
TobDeBer
Should UD-Q6_K_XL identical to Q6_K.gguf?
5
#1 opened 9 months ago
by
BVEsun
BF16 or Q8_K_XL - which would give more accurate coding results?
5
#6 opened 10 months ago
by
TimothyRoo
Jan 12 2026: Qwen3-Next updated with iMatrix + Improved performance!
👍 3
26
#3 opened 10 months ago
by
danielhanchen
Benchmark suggestion
2
#2 opened 10 months ago
by
FlareRebellion