How to use from
OpenClaw
Start the llama.cpp server
# Install llama.cpp:
brew install llama.cpp
# Start a local OpenAI-compatible server:
llama serve -hf datadab/writeamp-models:Q4_K_M
Configure OpenClaw
# Install OpenClaw:
npm install -g openclaw@latest
# Register the local server and set it as the default model:
openclaw onboard --non-interactive --mode local \
  --auth-choice custom-api-key \
  --custom-base-url http://127.0.0.1:8080/v1 \
  --custom-model-id "datadab/writeamp-models:Q4_K_M" \
  --custom-provider-id llama-cpp \
  --custom-compatibility openai \
  --custom-text-input \
  --accept-risk \
  --skip-health
Run OpenClaw
openclaw agent --local --agent main --message "Hello from Hugging Face"
Quick Links

WriteAmp's on-device model catalog

This repository hosts WriteAmp's downloadable model artifacts. The catalog has three tiers (mini, midi, max); two of them are mirrors of upstream Q4_K_M quantizations, one is a custom Hinglish fine-tune.

Tier File Size SHA-256 Source
mini mini/SmolLM2-135M.Q4_K_M.gguf 105,454,016 B 127c52bae983b2d0de09c60d6292743e3239e44239bbe8ccb8cf3241513ff5a0 Mirror of mradermacher/SmolLM2-135M-GGUF (Apache 2.0)
midi midi/SmolLM2-360M.Q4_K_M.gguf 270,590,464 B 4ac7aa712c43fe7d07ed910298b5eaab932292692954abd8e0703307724e4e04 Mirror of mradermacher/SmolLM2-360M-GGUF (Apache 2.0)
max WriteAmp-Qwen3-0.6B-Hinglish-v3.Q4_K_M.gguf 396,704,544 B 98df7dd1daaf65602fa002733d5c5a32527acd174730b591235f0ed170f17185 Custom QLoRA fine-tune of Qwen/Qwen3-0.6B-Base (Apache 2.0)

Each tier's mirror card (under mini/ and midi/) carries provenance and license details. The max tier's full description, training data, and known limitations are in WriteAmp-Qwen3-0.6B-Hinglish-v3.Q4_K_M.gguf's card above.

Mirroring policy

mini and midi are byte-identical mirrors of mradermacher's community Q4_K_M quantizations. The WriteAmp app's catalog keeps mradermacher's URLs as the primary download path (the community attribution is preserved), and uses this repo as a cold-start fallback when the primary becomes unreachable or returns HTTP 404. See ADR-245 in the WriteAmp decision log for the full rationale.

The max tier is hosted here as the primary (no third-party mirror). It is also the source of truth for the max catalog slot.

License

All artifacts under this repo are distributed under the Apache License 2.0, consistent with their upstream base models (HuggingFaceTB/SmolLM2-135M, HuggingFaceTB/SmolLM2-360M, Qwen/Qwen3-0.6B-Base) and mradermacher's Q4_K_M quantizations of those bases.

Citation

@misc{writeamp_models_2026,
  title  = {WriteAmp on-device model catalog: SmolLM2 mirrors + Qwen3-0.6B Hinglish fine-tune},
  author = {DataDab LLP},
  year   = {2026},
  url    = {https://huggingface.co/datadab/writeamp-models}
}
Downloads last month
62
GGUF
Model size
0.6B params
Architecture
qwen3
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support