UEmbed multimodal embedding models running dense text embeddings on Apple Silicon via MLX.
Volodymyr IERMOLAIEV PRO
majentik
AI & ML interests
Bleeding-edge AI models and most prominent ML open-source projects.
Recent Activity
updated a dataset 5 days ago
majentik/SASRBench-v1 published a dataset 5 days ago
majentik/SASRBench-v1 updated a dataset 5 days ago
majentik/WildASROrganizations
Embeddings (ONNX)
Qwen3-Embedding model packs in ONNX, GGUF, and MLX formats for retrieval and embedding workloads.
-
majentik/Qwen3-Embedding-0.6B-GGUF-IQ4_XS
Feature Extraction • 0.6B • Updated • 25 • 1 -
majentik/Qwen3-Embedding-0.6B-GGUF-Q4_K_M
Feature Extraction • 0.6B • Updated • 123 -
majentik/Qwen3-Embedding-0.6B-GGUF-Q5_K_M
Feature Extraction • 0.6B • Updated • 39 -
majentik/Qwen3-Embedding-0.6B-GGUF-Q8_0
Feature Extraction • 0.6B • Updated • 56
Verified MLX releases (smoke-gated)
Session-verified MLX packs that passed a smoke gate against a reference implementation before publication.
-
majentik/MERaLiON-3-10B-MLX-8bit
Automatic Speech Recognition • 3B • Updated • 18 -
majentik/MERaLiON-3-10B-MLX-4bit
Automatic Speech Recognition • 2B • Updated • 17 -
majentik/MERaLiON-3-3B-ASR-MLX-8bit
Automatic Speech Recognition • 1B • Updated • 33 -
majentik/MOSS-Transcribe-Diarize-MLX-8bit
Automatic Speech Recognition • 0.5B • Updated • 131 • 1
OCR & Document AI (MLX)
OCR and document-understanding MLX packs, including Unlimited-OCR and the GELab-Zero preview family.
-
majentik/GELab-Zero-4B-preview-Sico-Evolution-MLX-4bit
Image-Text-to-Text • 1B • Updated • 10 -
majentik/GELab-Zero-4B-preview-Sico-Evolution-MLX-6bit
Image-Text-to-Text • 1B • Updated • 7 -
majentik/GELab-Zero-4B-preview-Sico-Evolution-MLX-8bit
Image-Text-to-Text • 2B • Updated • 3 -
majentik/Unlimited-OCR-MLX-6bit
Image-Text-to-Text • 1B • Updated • 28
MERaLiON — Singapore speech models on MLX
MERaLiON speech and ASR model packs (MLX, GGUF, quantized variants) for Apple Silicon.
-
majentik/MERaLiON-3-10B-MLX
Automatic Speech Recognition • Updated • 101 -
majentik/MERaLiON-3-10B-MLX-4bit
Automatic Speech Recognition • 2B • Updated • 17 -
majentik/MERaLiON-3-10B-MLX-8bit
Automatic Speech Recognition • 3B • Updated • 18 -
majentik/MERaLiON-3-10B-RotorQuant-MLX-2bit
Automatic Speech Recognition • Updated • 24
Qwen 3.6 — quantized
Quantized GGUF, MLX, and FP8/FP4 packs of the Qwen 3.6 model family.
Meta Muse Glimmer — MLX
Agentic 30B vision-language model from Meta, quantized for Apple Silicon with MLX.
Nemotron — quantized
Quantized GGUF and MLX packs of the Nemotron 3 model family (Nano, Nano-Omni, Super).
-
majentik/Nemotron-3-Nano-30B-A3B-RotorQuant-GGUF-IQ4_XS
Text Generation • 32B • Updated • 22 -
majentik/Nemotron-3-Nano-30B-A3B-RotorQuant-GGUF-Q2_K
Text Generation • 32B • Updated • 28 -
majentik/Nemotron-3-Nano-30B-A3B-RotorQuant-GGUF-Q3_K_M
Text Generation • 32B • Updated • 22 -
majentik/Nemotron-3-Nano-30B-A3B-RotorQuant-GGUF-Q5_K_M
Text Generation • 32B • Updated • 20
gpt-oss — rebuilt GGUFs
The verified, rebuilt gpt-oss-20b RotorQuant GGUF quant family (the retired 120b tombstones are excluded).
-
majentik/gpt-oss-20b-RotorQuant-GGUF-IQ4_XS
Text Generation • 21B • Updated • 146 -
majentik/gpt-oss-20b-RotorQuant-GGUF-Q2_K
Text Generation • 21B • Updated • 141 -
majentik/gpt-oss-20b-RotorQuant-GGUF-Q3_K_M
Text Generation • 21B • Updated • 84 -
majentik/gpt-oss-20b-RotorQuant-GGUF-Q4_K_M
Text Generation • 21B • Updated • 114
ASR on Apple Silicon (MLX)
Automatic speech recognition and transcription/diarization MLX packs that run natively on Apple Silicon.
-
majentik/cohere-transcribe-arabic-07-2026-MLX-4bit
Automatic Speech Recognition • 0.5B • Updated • 42 -
majentik/cohere-transcribe-arabic-07-2026-MLX-6bit
Automatic Speech Recognition • 0.6B • Updated • 40 -
majentik/cohere-transcribe-arabic-07-2026-MLX-8bit
Automatic Speech Recognition • 0.8B • Updated • 115 -
majentik/MERaLiON-3-3B-ASR-MLX-8bit
Automatic Speech Recognition • 1B • Updated • 33
Qwen 3.5 — quantized
Quantized GGUF and MLX packs of the Qwen 3.5 model family, including the 397B-A17B MoE.
-
majentik/Qwen3.5-27B-RotorQuant-GGUF-IQ4_XS
Text Generation • 27B • Updated • 101 -
majentik/Qwen3.5-27B-RotorQuant-GGUF-Q2_K
Text Generation • 27B • Updated • 133 • 2 -
majentik/Qwen3.5-27B-RotorQuant-GGUF-Q3_K_M
Text Generation • 27B • Updated • 120 -
majentik/Qwen3.5-27B-RotorQuant-GGUF-Q4_K_M
Text Generation • 27B • Updated • 102
Gemma 4 — quantized (GGUF + MLX)
Quantized GGUF and MLX packs of the Gemma 4 family (dense and MoE variants) for local inference.
-
majentik/gemma-4-26B-A4B-it-RotorQuant-GGUF-IQ4_XS
Image-Text-to-Text • 25B • Updated • 157 -
majentik/gemma-4-26B-A4B-it-RotorQuant-GGUF-Q2_K
Image-Text-to-Text • 25B • Updated • 394 -
majentik/gemma-4-26B-A4B-it-RotorQuant-GGUF-Q3_K_M
Image-Text-to-Text • 25B • Updated • 126 -
majentik/gemma-4-26B-A4B-it-RotorQuant-GGUF-Q4_K_M
Image-Text-to-Text • 25B • Updated • 390 • 1
UEmbed — multimodal embeddings (MLX, dense text)
UEmbed multimodal embedding models running dense text embeddings on Apple Silicon via MLX.
Meta Muse Glimmer — MLX
Agentic 30B vision-language model from Meta, quantized for Apple Silicon with MLX.
Embeddings (ONNX)
Qwen3-Embedding model packs in ONNX, GGUF, and MLX formats for retrieval and embedding workloads.
-
majentik/Qwen3-Embedding-0.6B-GGUF-IQ4_XS
Feature Extraction • 0.6B • Updated • 25 • 1 -
majentik/Qwen3-Embedding-0.6B-GGUF-Q4_K_M
Feature Extraction • 0.6B • Updated • 123 -
majentik/Qwen3-Embedding-0.6B-GGUF-Q5_K_M
Feature Extraction • 0.6B • Updated • 39 -
majentik/Qwen3-Embedding-0.6B-GGUF-Q8_0
Feature Extraction • 0.6B • Updated • 56
Nemotron — quantized
Quantized GGUF and MLX packs of the Nemotron 3 model family (Nano, Nano-Omni, Super).
-
majentik/Nemotron-3-Nano-30B-A3B-RotorQuant-GGUF-IQ4_XS
Text Generation • 32B • Updated • 22 -
majentik/Nemotron-3-Nano-30B-A3B-RotorQuant-GGUF-Q2_K
Text Generation • 32B • Updated • 28 -
majentik/Nemotron-3-Nano-30B-A3B-RotorQuant-GGUF-Q3_K_M
Text Generation • 32B • Updated • 22 -
majentik/Nemotron-3-Nano-30B-A3B-RotorQuant-GGUF-Q5_K_M
Text Generation • 32B • Updated • 20
Verified MLX releases (smoke-gated)
Session-verified MLX packs that passed a smoke gate against a reference implementation before publication.
-
majentik/MERaLiON-3-10B-MLX-8bit
Automatic Speech Recognition • 3B • Updated • 18 -
majentik/MERaLiON-3-10B-MLX-4bit
Automatic Speech Recognition • 2B • Updated • 17 -
majentik/MERaLiON-3-3B-ASR-MLX-8bit
Automatic Speech Recognition • 1B • Updated • 33 -
majentik/MOSS-Transcribe-Diarize-MLX-8bit
Automatic Speech Recognition • 0.5B • Updated • 131 • 1
gpt-oss — rebuilt GGUFs
The verified, rebuilt gpt-oss-20b RotorQuant GGUF quant family (the retired 120b tombstones are excluded).
-
majentik/gpt-oss-20b-RotorQuant-GGUF-IQ4_XS
Text Generation • 21B • Updated • 146 -
majentik/gpt-oss-20b-RotorQuant-GGUF-Q2_K
Text Generation • 21B • Updated • 141 -
majentik/gpt-oss-20b-RotorQuant-GGUF-Q3_K_M
Text Generation • 21B • Updated • 84 -
majentik/gpt-oss-20b-RotorQuant-GGUF-Q4_K_M
Text Generation • 21B • Updated • 114
OCR & Document AI (MLX)
OCR and document-understanding MLX packs, including Unlimited-OCR and the GELab-Zero preview family.
-
majentik/GELab-Zero-4B-preview-Sico-Evolution-MLX-4bit
Image-Text-to-Text • 1B • Updated • 10 -
majentik/GELab-Zero-4B-preview-Sico-Evolution-MLX-6bit
Image-Text-to-Text • 1B • Updated • 7 -
majentik/GELab-Zero-4B-preview-Sico-Evolution-MLX-8bit
Image-Text-to-Text • 2B • Updated • 3 -
majentik/Unlimited-OCR-MLX-6bit
Image-Text-to-Text • 1B • Updated • 28
ASR on Apple Silicon (MLX)
Automatic speech recognition and transcription/diarization MLX packs that run natively on Apple Silicon.
-
majentik/cohere-transcribe-arabic-07-2026-MLX-4bit
Automatic Speech Recognition • 0.5B • Updated • 42 -
majentik/cohere-transcribe-arabic-07-2026-MLX-6bit
Automatic Speech Recognition • 0.6B • Updated • 40 -
majentik/cohere-transcribe-arabic-07-2026-MLX-8bit
Automatic Speech Recognition • 0.8B • Updated • 115 -
majentik/MERaLiON-3-3B-ASR-MLX-8bit
Automatic Speech Recognition • 1B • Updated • 33
MERaLiON — Singapore speech models on MLX
MERaLiON speech and ASR model packs (MLX, GGUF, quantized variants) for Apple Silicon.
-
majentik/MERaLiON-3-10B-MLX
Automatic Speech Recognition • Updated • 101 -
majentik/MERaLiON-3-10B-MLX-4bit
Automatic Speech Recognition • 2B • Updated • 17 -
majentik/MERaLiON-3-10B-MLX-8bit
Automatic Speech Recognition • 3B • Updated • 18 -
majentik/MERaLiON-3-10B-RotorQuant-MLX-2bit
Automatic Speech Recognition • Updated • 24
Qwen 3.5 — quantized
Quantized GGUF and MLX packs of the Qwen 3.5 model family, including the 397B-A17B MoE.
-
majentik/Qwen3.5-27B-RotorQuant-GGUF-IQ4_XS
Text Generation • 27B • Updated • 101 -
majentik/Qwen3.5-27B-RotorQuant-GGUF-Q2_K
Text Generation • 27B • Updated • 133 • 2 -
majentik/Qwen3.5-27B-RotorQuant-GGUF-Q3_K_M
Text Generation • 27B • Updated • 120 -
majentik/Qwen3.5-27B-RotorQuant-GGUF-Q4_K_M
Text Generation • 27B • Updated • 102
Qwen 3.6 — quantized
Quantized GGUF, MLX, and FP8/FP4 packs of the Qwen 3.6 model family.
Gemma 4 — quantized (GGUF + MLX)
Quantized GGUF and MLX packs of the Gemma 4 family (dense and MoE variants) for local inference.
-
majentik/gemma-4-26B-A4B-it-RotorQuant-GGUF-IQ4_XS
Image-Text-to-Text • 25B • Updated • 157 -
majentik/gemma-4-26B-A4B-it-RotorQuant-GGUF-Q2_K
Image-Text-to-Text • 25B • Updated • 394 -
majentik/gemma-4-26B-A4B-it-RotorQuant-GGUF-Q3_K_M
Image-Text-to-Text • 25B • Updated • 126 -
majentik/gemma-4-26B-A4B-it-RotorQuant-GGUF-Q4_K_M
Image-Text-to-Text • 25B • Updated • 390 • 1