One model for both halves of RAG retrieval; a strong default per size. Contact baa.ai for the optimal pick for your corpus.
AI & ML interests
Model Quantization
Recent Activity
View all activity
Organization Card
Smaller. Smarter. Sovereign.
Making frontier models run anywhere
We publish high-quality quantized models for Apple Silicon and GGUF. Our models use a proprietary optimisation method that delivers superior quality at your target memory budget.
Browse our models, or connect with us below.
models 81
baa-ai/Qwen3.8-27B-RAM-13GB-GGUF
Text Generation • 27B • Updated • 259 • 2
baa-ai/Qwen3.8-27B-RAM-31GB-GGUF
Text Generation • 27B • Updated • 72
baa-ai/Qwen3.8-27B-RAM-24GB-MLX
Image-Text-to-Text • 8B • Updated • 324 • 1
baa-ai/paddock-reader-35b-gguf
35B • Updated • 21
baa-ai/paddock-reader-9b-gguf
9B • Updated • 13
baa-ai/GLM-5.2-RAM-333GB-MLX
96B • Updated • 436
baa-ai/Merino-XL-v2
Sentence Similarity • Updated • 5
baa-ai/Merino-XL
Sentence Similarity • Updated • 4
baa-ai/Merino-Pro-4bit
Sentence Similarity • Updated • 5
baa-ai/Merino-Pro
Sentence Similarity • Updated • 17
datasets 0
None public yet