AX-Ornith-1.0-35B-MLX-AXQ-6bit

Development AXQuant (AXQ) 6bit MLX pack of deepreinforce-ai/Ornith-1.0-35B @ 5df2ed3f675c7beaa490328cc70bb573b65fb660.

Claims (read carefully)

Claim Status
AXQuant architecture-prior / development quant Yes
Qwen 3.6 Tier 1 / Tier 2 certification No
Measured coding-bench scores Not claimed
MTP acceleration Not claimed (no -MTP product label; source has no MTP weights)
Vision / VLM quality Not claimed — vision tensors preserved at BF16 only

Source

  • Model: deepreinforce-ai/Ornith-1.0-35B (MIT; twin of ornith-ai/Ornith-1.0-35B)
  • Revision: 5df2ed3f675c7beaa490328cc70bb573b65fb660
  • Architecture: qwen3_5_moe 35B-A3B-class MoE
  • AXQuant adapter: qwen35-moe-v1 (convertible, development)
  • Prep: per-expert HF weights restacked to MLX packed experts.gate_up_proj / experts.down_proj before convert

Quantization

  • Toolkit: AXQuant (MLX-native PTQ)
  • Product class: 6bit
  • Planning: architecture-prior ladder (not measured sensitivity)
  • Optimization scope: text path only
  • Vision tower: BF16-preserved (vision.safetensors)

Intended use

Local Apple Silicon inference via MLX-LM for development evaluation. Do not treat this pack as a certified AutomatosX flagship release.

Attribution

Base weights © Ornith / DeepReinforce contributors under MIT. Quantization artifacts produced with AXQuant for research and development use.

Links

Downloads last month
-
Safetensors
Model size
7B params
Tensor type
BF16
·
U32
·
MLX
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for AutomatosX/AX-Ornith-1.0-35B-MLX-AXQ-6bit

Quantized
(182)
this model

Collections including AutomatosX/AX-Ornith-1.0-35B-MLX-AXQ-6bit