ProCreations's picture
Add WebGPU pack for Maple 361db5da
93c57e5 verified
|
Raw
History Blame Contribute Delete
1 kB
---
license: mit
library_name: webgpu
base_model: deepgrove/maple-preview
tags:
- webgpu
- browser
- text-generation
- quantized
---
# Maple-Preview WebGPU Pack
Lossless browser-streaming repack of
[`deepgrove/maple-preview-2bit-mlx`](https://huggingface.co/deepgrove/maple-preview-2bit-mlx)
for [`ProCreations/maple-webgpu`](https://huggingface.co/spaces/ProCreations/maple-webgpu).
No model values are changed and no additional quantization is applied. The
official 2-bit/4-bit MLX tensors are regrouped into one aligned binary file per
transformer layer, reducing browser loading from hundreds of range requests to
a small set of sequential CDN downloads.
- Source revision: `361db5da5e74ff6fcdd852d478e1f266ce11013a`
- Packed tensor bytes: `5.31 GB`
- Container format: `maple-webgpu-pack-v1`
- Original model and weights: MIT, © DeepGrove AI
This is a runtime artifact, not a separate model release. See `manifest.json`
for exact byte offsets, shapes, dtypes, and source tensor names.