File size: 1,781 Bytes
d66eadd
 
 
 
 
 
 
 
 
 
ac785ec
 
d66eadd
 
 
 
 
 
 
 
cbc4b1e
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
# Altronis

**Less chaos. More progress.** Altronis is a Singapore AI consultancy that designs, builds and runs practical AI for organisations whose data cannot leave their walls: private LLMs, on-prem and edge AI, agentic workflows, and the governance around them.

We publish models we actually run in production, with honestly labeled evaluations - sample sizes and who scored them stated on every card.

## What's here

**Judgment QC gate** - a family of fine-tunes that distill one operator's engineering judgment into a small local reviewer: an always-on quality gate that judges outbound work before it ships, returning the verdict a careful engineer would give (or "OK"). Trained with Unsloth on a desk-side AMD Strix Halo machine (128GB unified memory, ROCm). The GGUF build is the exact artifact serving in our production pipeline today.

Evaluated on a 58-example held-out set the models never trained on: the served Qwen3-4B gate holds its output contract 58/58, with zero false blocks on clean work and 33 of 34 real violations caught. A Gemma-4-E4B sibling trained on the identical split matches it on judgment and loses on format - we publish the comparison, not only the winner.

Full build story and recipe: https://altronis.sg/blog/fine-tune-llm-unsloth-amd-strix-halo

## Links

- Web: https://altronis.sg
- Private and on-prem LLM services: https://altronis.sg/private-llm-sg
- Contact: hello@altronis.sg

**Built on AMD.** Our models are trained and served on AMD hardware (Strix Halo / Ryzen AI Max, 128GB unified memory, ROCm) - the judgment-gate family on this page was fine-tuned, evaluated, and deployed entirely on a single desk-side AMD machine. We publish our AMD setup and contribute to the AMD open-source AI ecosystem.

Also NVIDIA Inception member.