h4xx0r-v1-2B

A fine-tuned version of unsloth/Qwen3.5-2B trained on theprint h4xx0r data using Auto-SFT — an automated hyperparameter search and supervised fine-tuning pipeline.

The base model was adapted to follow the style and content of the theprint h4xx0r dataset. Expect improved performance on tasks similar to those represented in the training data.

Model Details

Property Value
Base model unsloth/Qwen3.5-2B
Training data theprint/h4xx0r
Fine-tuning epochs 2
Fine-tuning date 2026-07-11
Fine-tuning method LoRA (merged to full 16-bit)

Training Hyperparameters

LoRA

Parameter Value
r 8
alpha 32
dropout 0.05
target_modules ['q_proj', 'v_proj', 'k_proj', 'o_proj']

Training

Parameter Value
learning_rate 0.0005
batch_size 2
gradient_accumulation_steps 1
warmup_ratio 0.0
max_seq_length 1024
quantization none

Usage

from transformers import AutoModelForCausalLM, AutoTokenizer

model     = AutoModelForCausalLM.from_pretrained("theprint/h4xx0r-v1-2B")
tokenizer = AutoTokenizer.from_pretrained("theprint/h4xx0r-v1-2B")

Generated by Auto-SFT

Downloads last month
649
Safetensors
Model size
2B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for theprint/h4xx0r-v1-2B

Finetuned
Qwen/Qwen3.5-2B
Adapter
(51)
this model
Adapters
2 models

Dataset used to train theprint/h4xx0r-v1-2B