Unsloth dynamic quants for MLX: MXFP4+MXFP8 on macOS; NVFP4+E4M3 weights with A16 on DGX.
🤝 Open to Collab
Yinan Long
Brooooooklyn
AI & ML interests
None yet
Recent Activity
updated a model about 20 hours ago
Brooooooklyn/Qwen-AgentWorld-35B-A3B-mxfp4-mlx updated a model about 22 hours ago
Brooooooklyn/Qwen-AgentWorld-35B-A3B-nvfp4-mlx updated a model about 22 hours ago
Brooooooklyn/Ornith-1.0-35B-mxfp4-mlxOrganizations
None yet
Qwen-AgentWorld-35B-A3B-unsloth-mlx
Qwen-AgentWorld-35B-A3B (Qwen3.5-VL-MoE world model) for Apple Silicon MLX: legacy imatrix-AWQ Q3–Q8 + MXFP8, quality-verified vs BF16.
-
Brooooooklyn/Qwen-AgentWorld-35B-A3B-UD-Q3_K_XL-mlx
Image-Text-to-Text • 35B • Updated • 60 • 1 -
Brooooooklyn/Qwen-AgentWorld-35B-A3B-UD-Q4_K_XL-mlx
Image-Text-to-Text • 35B • Updated • 46 -
Brooooooklyn/Qwen-AgentWorld-35B-A3B-UD-Q5_K_XL-mlx
Image-Text-to-Text • 35B • Updated • 50 -
Brooooooklyn/Qwen-AgentWorld-35B-A3B-UD-Q6_K_XL-mlx
Image-Text-to-Text • 35B • Updated • 18
Gemma-4-31B-IT-unsloth-mlx
Gemma-4-31B-IT (dense vision-language) quantized for Apple Silicon (MLX) — Unsloth Dynamic 2.0 with AWQ imatrix pre-scaling.
-
Brooooooklyn/Gemma-4-31B-IT-UD-Q2_K_XL-mlx
Text Generation • 31B • Updated • 20 -
Brooooooklyn/Gemma-4-31B-IT-UD-Q3_K_XL-mlx
Text Generation • 31B • Updated • 32 -
Brooooooklyn/Gemma-4-31B-IT-UD-Q4_K_XL-mlx
Text Generation • 31B • Updated • 50 -
Brooooooklyn/Gemma-4-31B-IT-UD-MXFP4_K_XL-mlx
Text Generation • 31B • Updated • 14 • 1
Qwen-3.6-unsloth-mlx
AWQ-style pre-scaling using Unsloth's imatrix calibration data, then 3-6-bit affine quantization with the Unsloth mixed-precision recipe via MLX
-
Brooooooklyn/Qwen3.6-35B-A3B-UD-Q2_K_XL-mlx
Text Generation • 35B • Updated • 339 • 7 -
Brooooooklyn/Qwen3.6-35B-A3B-UD-Q3_K_XL-mlx
Text Generation • 35B • Updated • 415 • 4 -
Brooooooklyn/Qwen3.6-35B-A3B-UD-Q4_K_XL-mlx
Text Generation • 35B • Updated • 857 • 9 -
Brooooooklyn/Qwen3.6-35B-A3B-UD-Q5_K_XL-mlx
Text Generation • 35B • Updated • 158 • 2
NVIDIA NVFP4 → MLX
NVIDIA post-trained NVFP4 checkpoints (ModelOpt PTQ) transcoded to MLX for Apple Silicon / mlx-node, keeping NVIDIA's per-layer bit allocation.
Ornith-1.0-35B-unsloth-mlx
Ornith-1.0-35B for Apple Silicon MLX: legacy Unsloth Dynamic Q3–Q8 + MXFP8. Output quality verified vs BF16.
-
Brooooooklyn/Ornith-1.0-35B-UD-Q3_K_XL-mlx
Text Generation • 35B • Updated • 65 -
Brooooooklyn/Ornith-1.0-35B-UD-Q4_K_XL-mlx
Text Generation • 35B • Updated • 69 -
Brooooooklyn/Ornith-1.0-35B-UD-Q5_K_XL-mlx
Text Generation • 35B • Updated • 125 • 1 -
Brooooooklyn/Ornith-1.0-35B-UD-Q6_K_XL-mlx
Text Generation • 35B • Updated • 45 • 1
Gemma-4-26B-IT-unsloth-mlx
Gemma-4-26B-A4B-IT MoE quantized for Apple Silicon (MLX) — Unsloth Dynamic 2.0 with AWQ imatrix pre-scaling.
-
Brooooooklyn/Gemma-4-26B-A4B-IT-UD-Q3_K_XL-mlx
Text Generation • 26B • Updated • 64 -
Brooooooklyn/Gemma-4-26B-A4B-IT-UD-MXFP4_K_XL-mlx
Text Generation • 26B • Updated • 35 • 1 -
Brooooooklyn/Gemma-4-26B-A4B-IT-UD-Q4_K_XL-mlx
Text Generation • 26B • Updated • 123 -
Brooooooklyn/Gemma-4-26B-A4B-IT-UD-NVFP4_K_XL-mlx
Text Generation • 26B • Updated • 24
Qwen-3.5-unsloth-mlx
AWQ-style pre-scaling using Unsloth's imatrix calibration data, then 3-6-bit affine quantization with the Unsloth mixed-precision recipe via MLX
-
Brooooooklyn/Qwen3.5-9B-unsloth-mlx
Text Generation • 9B • Updated • 477 • 13 -
Brooooooklyn/Qwen3.5-27B-unsloth-mlx
Text Generation • 27B • Updated • 558 • 18 -
Brooooooklyn/Qwen3.5-35B-A3B-UD-Q2_K_XL-mlx
Text Generation • 35B • Updated • 33 -
Brooooooklyn/Qwen3.5-35B-A3B-UD-Q3_K_XL-mlx
Text Generation • 35B • Updated • 65 • 2
Unsloth NVFP4 Tensor-Class Recipe for MLX — macOS + DGX
Unsloth dynamic quants for MLX: MXFP4+MXFP8 on macOS; NVFP4+E4M3 weights with A16 on DGX.
NVIDIA NVFP4 → MLX
NVIDIA post-trained NVFP4 checkpoints (ModelOpt PTQ) transcoded to MLX for Apple Silicon / mlx-node, keeping NVIDIA's per-layer bit allocation.
Qwen-AgentWorld-35B-A3B-unsloth-mlx
Qwen-AgentWorld-35B-A3B (Qwen3.5-VL-MoE world model) for Apple Silicon MLX: legacy imatrix-AWQ Q3–Q8 + MXFP8, quality-verified vs BF16.
-
Brooooooklyn/Qwen-AgentWorld-35B-A3B-UD-Q3_K_XL-mlx
Image-Text-to-Text • 35B • Updated • 60 • 1 -
Brooooooklyn/Qwen-AgentWorld-35B-A3B-UD-Q4_K_XL-mlx
Image-Text-to-Text • 35B • Updated • 46 -
Brooooooklyn/Qwen-AgentWorld-35B-A3B-UD-Q5_K_XL-mlx
Image-Text-to-Text • 35B • Updated • 50 -
Brooooooklyn/Qwen-AgentWorld-35B-A3B-UD-Q6_K_XL-mlx
Image-Text-to-Text • 35B • Updated • 18
Ornith-1.0-35B-unsloth-mlx
Ornith-1.0-35B for Apple Silicon MLX: legacy Unsloth Dynamic Q3–Q8 + MXFP8. Output quality verified vs BF16.
-
Brooooooklyn/Ornith-1.0-35B-UD-Q3_K_XL-mlx
Text Generation • 35B • Updated • 65 -
Brooooooklyn/Ornith-1.0-35B-UD-Q4_K_XL-mlx
Text Generation • 35B • Updated • 69 -
Brooooooklyn/Ornith-1.0-35B-UD-Q5_K_XL-mlx
Text Generation • 35B • Updated • 125 • 1 -
Brooooooklyn/Ornith-1.0-35B-UD-Q6_K_XL-mlx
Text Generation • 35B • Updated • 45 • 1
Gemma-4-31B-IT-unsloth-mlx
Gemma-4-31B-IT (dense vision-language) quantized for Apple Silicon (MLX) — Unsloth Dynamic 2.0 with AWQ imatrix pre-scaling.
-
Brooooooklyn/Gemma-4-31B-IT-UD-Q2_K_XL-mlx
Text Generation • 31B • Updated • 20 -
Brooooooklyn/Gemma-4-31B-IT-UD-Q3_K_XL-mlx
Text Generation • 31B • Updated • 32 -
Brooooooklyn/Gemma-4-31B-IT-UD-Q4_K_XL-mlx
Text Generation • 31B • Updated • 50 -
Brooooooklyn/Gemma-4-31B-IT-UD-MXFP4_K_XL-mlx
Text Generation • 31B • Updated • 14 • 1
Gemma-4-26B-IT-unsloth-mlx
Gemma-4-26B-A4B-IT MoE quantized for Apple Silicon (MLX) — Unsloth Dynamic 2.0 with AWQ imatrix pre-scaling.
-
Brooooooklyn/Gemma-4-26B-A4B-IT-UD-Q3_K_XL-mlx
Text Generation • 26B • Updated • 64 -
Brooooooklyn/Gemma-4-26B-A4B-IT-UD-MXFP4_K_XL-mlx
Text Generation • 26B • Updated • 35 • 1 -
Brooooooklyn/Gemma-4-26B-A4B-IT-UD-Q4_K_XL-mlx
Text Generation • 26B • Updated • 123 -
Brooooooklyn/Gemma-4-26B-A4B-IT-UD-NVFP4_K_XL-mlx
Text Generation • 26B • Updated • 24
Qwen-3.6-unsloth-mlx
AWQ-style pre-scaling using Unsloth's imatrix calibration data, then 3-6-bit affine quantization with the Unsloth mixed-precision recipe via MLX
-
Brooooooklyn/Qwen3.6-35B-A3B-UD-Q2_K_XL-mlx
Text Generation • 35B • Updated • 339 • 7 -
Brooooooklyn/Qwen3.6-35B-A3B-UD-Q3_K_XL-mlx
Text Generation • 35B • Updated • 415 • 4 -
Brooooooklyn/Qwen3.6-35B-A3B-UD-Q4_K_XL-mlx
Text Generation • 35B • Updated • 857 • 9 -
Brooooooklyn/Qwen3.6-35B-A3B-UD-Q5_K_XL-mlx
Text Generation • 35B • Updated • 158 • 2
Qwen-3.5-unsloth-mlx
AWQ-style pre-scaling using Unsloth's imatrix calibration data, then 3-6-bit affine quantization with the Unsloth mixed-precision recipe via MLX
-
Brooooooklyn/Qwen3.5-9B-unsloth-mlx
Text Generation • 9B • Updated • 477 • 13 -
Brooooooklyn/Qwen3.5-27B-unsloth-mlx
Text Generation • 27B • Updated • 558 • 18 -
Brooooooklyn/Qwen3.5-35B-A3B-UD-Q2_K_XL-mlx
Text Generation • 35B • Updated • 33 -
Brooooooklyn/Qwen3.5-35B-A3B-UD-Q3_K_XL-mlx
Text Generation • 35B • Updated • 65 • 2