0xSero
0xSero
AI & ML interests
Quantizing, benchmarking, training, and building.
Recent Activity
published an article 6 days ago
The DGX Spark Handbook liked a model 10 days ago
AikidoSec/altar-1 updated a model 14 days ago
0xSero/GLM-5.3-Flash-EXL3-SparkOrganizations
Running on 4x RTX PRO 6000 with NVMe offload for ngram
🚀 7
8
#28 opened 22 days ago
by
0xSero
Independent KLD measurement of this Q4 on a sealed 25-window panel: 0.027263 nats (5 bitwise-identical runs)
2
#1 opened about 1 month ago
by
malaiwah
Amazing KLD how?
❤️ 1
1
#2 opened about 2 months ago
by
0xSero
Terminal-Bench-2.1 & Throughput on 4x RTX PRO 6000 (gen 4 system)
❤️ 1
1
#1 opened 2 months ago
by
0xSero
Donation cost to run Heretic or otherwise Abliterate this?
3
#3 opened 3 months ago
by
dissociativity
Axis 4 reasoning/termination: 3,057 -- Can you share this calibration data?
👍 1
3
#2 opened 3 months ago
by
0xSero
Korean Multilingual is broken.
4
#6 opened 4 months ago
by
DFveloper
Starting GPTQ model on H200 fails
3
#2 opened 4 months ago
by
hjjg85
error with vLLM
6
#3 opened 5 months ago
by
bnjmnmarie
Running models with vLLM on the RTX Pro 6000 - SM120
👀👍 2
12
#28 opened 5 months ago
by
liku2001
Thank you 🙏
4
#1 opened 6 months ago
by
BlueNipples
Special Token Disaster: Your Tech Lead Has Zero Design Taste
👀🔥 3
4
#6 opened 5 months ago
by
ytgui
How are you running this?
5
#1 opened 6 months ago
by
richardhundt
use chat template here or the new one from google?
1
#5 opened 6 months ago
by
Grandys
Repetition loops with llama.cpp defaults
1
#2 opened 6 months ago
by
todaymare
"vLLM Launch Failed - Kernel Incompatibility with RTX 3090"
5
#3 opened 7 months ago
by
steppi
Add recovery report for VLM fix
#1 opened 7 months ago
by
0xSero