You need to agree to share your contact information to access this model

This repository is publicly accessible, but you have to accept the conditions to access its files and content.

Log in or Sign Up to review the conditions and access this model content.

KumaKuma-Qwen3.5-9B-r09

Archived copy of the exported Hugging Face checkpoint from KumaKuma core round r09:

07-24-2026-1-kumakuma9b-core-r09-global_step_50-hf

Training provenance

  • Architecture: Qwen3_5ForConditionalGeneration
  • Precision: bfloat16
  • Round: 50 GRPO steps
  • Previous checkpoint: KumaKuma core r08, step 50
  • Curriculum: 3,024 deterministic multi-turn rows in 378 batches
  • Per-batch pool mix: 4 target-stage + 2 other-stage + 2 end-to-end
  • Target stage: S4
  • Source: 126 approved clean rows
  • Source SHA-256: 087b4a108a487686a8d3c67f0aab35d1b0191972922e65d1fd59bdaceb679b2f
  • Materialized train file SHA-256: da60027744d9df15067c29327cf73c225af06349ffb96bd7f47c792a1d183564

Export integrity

model.safetensors SHA-256:

fd13812def69bf7bba734b53ab82c83f018a24d05c0d9c2661c89e4fc3916b28

Evaluation note

The round completed all 50 finite training steps. Its seven-task gate measured overall 0.585914 and task score 0.489229, but the gate decision was hold because one track-regression check failed. This is an archived research checkpoint, not a promoted production release.

Downloads last month
4
Safetensors
Model size
9B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for hab-swe/KumaKuma-Qwen3.5-9B-R09

Finetuned
Qwen/Qwen3.5-9B
Finetuned
(991)
this model