0xSero commited on
Commit
304cbb9
·
verified ·
1 Parent(s): 20084ba

Standardize model card (template rollout)

Browse files
Files changed (1) hide show
  1. README.md +44 -18
README.md CHANGED
@@ -1,27 +1,53 @@
1
  ---
2
- base_model: zai-org/GLM-4.7-Flash
 
 
 
 
 
 
3
  ---
 
4
  > [!TIP]
5
- > Support this work: **[donate.sybilsolutions.ai](https://donate.sybilsolutions.ai)**
6
- >
7
- > REAP surfaces: [GLM](https://huggingface.co/spaces/0xSero/reap-glm-family) | [MiniMax](https://huggingface.co/spaces/0xSero/reap-minimax-family) | [Qwen](https://huggingface.co/spaces/0xSero/reap-qwen-family) | [Gemma](https://huggingface.co/spaces/0xSero/reap-gemma-family) | [Paper](https://arxiv.org/abs/2510.13999) | [Code](https://github.com/CerebrasResearch/reap) | [PR17](https://github.com/CerebrasResearch/reap/pull/17) | [Cerebras Collection](https://huggingface.co/collections/cerebras/cerebras-reap)
8
 
9
- # glm-4.7-flash-sero
10
 
 
11
 
12
- <!-- SERO_MANAGED_TOP_LINKS_START -->
13
- ## Support and links
14
- - Donate: https://donate.sybilsolutions.ai
15
- - X: https://x.com/0xsero
16
- - GitHub: https://github.com/0xsero
17
- <!-- SERO_MANAGED_TOP_LINKS_END -->
18
 
19
- ## Sponsors
 
 
 
 
 
 
 
 
 
 
 
 
20
 
21
- Thank you for the kind sponsors, wouldn't be possible without them:
 
 
 
 
 
22
 
23
- - Nvidia
24
- - TNG Technology
25
- - Lambda
26
- - Prime Intellect
27
- - HotAisle
 
 
 
 
 
 
 
 
 
1
  ---
2
+ base_model:
3
+ - zai-org/GLM-4.7-Flash
4
+ license: mit
5
+ pipeline_tag: text-generation
6
+ library_name: transformers
7
+ tags:
8
+ - glm
9
  ---
10
+
11
  > [!TIP]
12
+ > **[Support this work ](https://donate.sybilsolutions.ai)** · [X](https://x.com/0xsero) · [GitHub](https://github.com/0xsero) · [REAP paper](https://arxiv.org/abs/2510.13999) · [Cerebras REAP](https://huggingface.co/collections/cerebras/cerebras-reap)
 
 
13
 
14
+ # GLM-4.7-Flash
15
 
16
+ REAP-pruned [zai-org/GLM-4.7-Flash](https://huggingface.co/zai-org/GLM-4.7-Flash).
17
 
18
+ ## At a glance
 
 
 
 
 
19
 
20
+ | | |
21
+ |---|---|
22
+ | Base model | [zai-org/GLM-4.7-Flash](https://huggingface.co/zai-org/GLM-4.7-Flash) |
23
+ | Format | BF16 |
24
+ | Total params | **30B** |
25
+ | Active / token | — |
26
+ | Experts / layer | 64 |
27
+ | Layers | 47 |
28
+ | Hidden size | 2048 |
29
+ | Context | 202,752 |
30
+ | On-disk size | 120 GB |
31
+
32
+ ## Which variant should I pick?
33
 
34
+ | Variant | Format | Link |
35
+ |---|---|---|
36
+ | `GLM-4.7-Flash` **(this)** | BF16 | [link](https://huggingface.co/0xSero/GLM-4.7-Flash) |
37
+ | `GLM-4.7-Flash-DPO` | DPO | [link](https://huggingface.co/0xSero/GLM-4.7-Flash-DPO) |
38
+ | `GLM-4.7-Flash-SFT` | SFT | [link](https://huggingface.co/0xSero/GLM-4.7-Flash-SFT) |
39
+ | `GLM-4.7-Flash-Tools` | Tools | [link](https://huggingface.co/0xSero/GLM-4.7-Flash-Tools) |
40
 
41
+ ## License & citation
42
+ License inherited from the base model.
43
+
44
+ ```bibtex
45
+ @misc{lasby2025reap,
46
+ title = {REAP the Experts: Why Pruning Prevails for One-Shot MoE Compression},
47
+ author = {Mike Lasby and Ivan Lazarevich and Nish Sinnadurai and Sean Lie and Yani Ioannou and Vithursan Thangarasa},
48
+ year = {2025}, eprint = {2510.13999}, archivePrefix = {arXiv}
49
+ }
50
+ ```
51
+
52
+ ## Sponsors
53
+ Made possible by **NVIDIA · TNG Technology · Lambda · Prime Intellect · Hot Aisle**.