Mochi 1.0

Mochi 1.0 (deprecated)

Mochi 1.5 is in development and will replace this model entirely. Mochi 1.0 is no longer maintained.

A small, unprofessional fine-tune of Qwen 2.5, made to have a slightly different personality. Based on huihui-ai/Qwen2.5-0.5B-Instruct-abliterated-v3.

Mochi prioritizes having multiple different variants of the base model for different purposes, always using 0.5B parameters. It is only made for fun, and has no actual use if you're looking for something powerful or coherent.


Known Issues

  • System prompt required. The personality does not hold without a system prompt passed at runtime. If you run the model without one, it will behave like the base Qwen model. This is a fundamental limitation of 1.0 and will be addressed in 1.5.
  • Factual accuracy is degraded compared to the base model. At 0.5B, personality training trades off against factual recall.

Model Details

Property Value
Base Model Qwen2.5-0.5B-Instruct (abliterated)
Parameters 0.5B
Quantization Q4_K_M
Format GGUF
Language English
License Apache 2.0

Usage

llama.cpp

llama-cli -m Mochi1.0-Q4_K_M.gguf \
  -sys "you are a casual and straightforward ai. you answer directly without performing helpfulness. you're friendly but not excessively so. you never use em dashes. if you're not sure about something, say so instead of guessing." \
  --jinja \
  --ctx-size 2048

Ollama

FROM ./Mochi1.0-Q4_K_M.gguf
SYSTEM "you are a casual and straightforward ai. you answer directly without performing helpfulness. you're friendly but not excessively so. you never use em dashes. if you're not sure about something, say so instead of guessing."
ollama create mochi -f Modelfile
ollama run mochi

Training

Fine-tuned with Unsloth using QLoRA on Google Colab (T4 GPU).

  • ~1000 training examples
  • 3 epochs
  • Learning rate: 1e-4

Credits

  • Qwen2.5 by Alibaba Cloud โ€” Apache 2.0
  • huihui-ai for the abliterated base
Downloads last month
52
GGUF
Model size
0.5B params
Architecture
qwen2
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for stupidlime/Mochi-1.0