Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
aidendle94
/
GLM-5.2-MXFP4-Experts-GPTQ
like
6
Safetensors
vllm
glm_moe_dsa
glm
Mixture of Experts
mxfp4
fp8
gptq
mixed-precision
quantization
dgx-spark
gb10
hybrid_mxfp4_ct
License:
mit
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
GLM-5.2-MXFP4-Experts-GPTQ
/
mtp-draft
6.01 GB
Ctrl+K
Ctrl+K
1 contributor
History:
1 commit
aidendle94
Full standalone model: FP8 attention/shared + NVFP4 dense + GPTQ-MXFP4 experts + MTP draft + stitched index
5285593
verified
7 days ago
config.json
10 kB
Full standalone model: FP8 attention/shared + NVFP4 dense + GPTQ-MXFP4 experts + MTP draft + stitched index
7 days ago
model-mtp-inputscales.safetensors
86.2 kB
xet
Full standalone model: FP8 attention/shared + NVFP4 dense + GPTQ-MXFP4 experts + MTP draft + stitched index
7 days ago
model-mtp.safetensors
6.01 GB
xet
Full standalone model: FP8 attention/shared + NVFP4 dense + GPTQ-MXFP4 experts + MTP draft + stitched index
7 days ago
model.safetensors.index.json
Safe
259 kB
Full standalone model: FP8 attention/shared + NVFP4 dense + GPTQ-MXFP4 experts + MTP draft + stitched index
7 days ago