| license: mit | |
| library_name: webgpu | |
| base_model: deepgrove/maple-preview | |
| tags: | |
| - webgpu | |
| - browser | |
| - text-generation | |
| - quantized | |
| # Maple-Preview WebGPU Pack | |
| Lossless browser-streaming repack of | |
| [`deepgrove/maple-preview-2bit-mlx`](https://huggingface.co/deepgrove/maple-preview-2bit-mlx) | |
| for [`ProCreations/maple-webgpu`](https://huggingface.co/spaces/ProCreations/maple-webgpu). | |
| No model values are changed and no additional quantization is applied. The | |
| official 2-bit/4-bit MLX tensors are regrouped into one aligned binary file per | |
| transformer layer, reducing browser loading from hundreds of range requests to | |
| a small set of sequential CDN downloads. | |
| - Source revision: `361db5da5e74ff6fcdd852d478e1f266ce11013a` | |
| - Packed tensor bytes: `5.31 GB` | |
| - Container format: `maple-webgpu-pack-v1` | |
| - Original model and weights: MIT, © DeepGrove AI | |
| This is a runtime artifact, not a separate model release. See `manifest.json` | |
| for exact byte offsets, shapes, dtypes, and source tensor names. | |