How to use from the
Use from the
MLX library
# Download the model from the Hub
pip install huggingface_hub[hf_xet]

huggingface-cli download --local-dir MOSS-Music-8B-Instruct-4bit rtikw/MOSS-Music-8B-Instruct-4bit

MOSS-Music-8B-Instruct โ€” MLX 4-bit

An MLX conversion of OpenMOSS-Team/MOSS-Music-8B-Instruct (Apache-2.0), quantized to 4-bit (group size 64) with the moss_music_mlx converter from the community MLX port of the MOSS-Music inference stack โ€” the same converter and layout as the mlx-community/MOSS-Music-8B-Thinking-* quants. The language model, adapters, lm_head and token embeddings are quantized; the audio encoder is kept in bf16.

Weights are otherwise unmodified. All credit to the OpenMOSS team; see the base model card for capabilities, benchmarks and citation.

Converted for use in a local macOS music app; uploaded so its model catalog has a stable download source.

Downloads last month
-
Safetensors
Model size
2B params
Tensor type
BF16
ยท
U32
ยท
MLX
Hardware compatibility
Log In to add your hardware

Quantized

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for rtikw/MOSS-Music-8B-Instruct-4bit

Finetuned
(1)
this model