--- license: mit library_name: webgpu base_model: deepgrove/maple-preview tags: - webgpu - browser - text-generation - quantized --- # Maple-Preview WebGPU Pack Lossless browser-streaming repack of [`deepgrove/maple-preview-2bit-mlx`](https://huggingface.co/deepgrove/maple-preview-2bit-mlx) for [`ProCreations/maple-webgpu`](https://huggingface.co/spaces/ProCreations/maple-webgpu). No model values are changed and no additional quantization is applied. The official 2-bit/4-bit MLX tensors are regrouped into one aligned binary file per transformer layer, reducing browser loading from hundreds of range requests to a small set of sequential CDN downloads. - Source revision: `361db5da5e74ff6fcdd852d478e1f266ce11013a` - Packed tensor bytes: `5.31 GB` - Container format: `maple-webgpu-pack-v1` - Original model and weights: MIT, © DeepGrove AI This is a runtime artifact, not a separate model release. See `manifest.json` for exact byte offsets, shapes, dtypes, and source tensor names.