laya-nli-conflict-v5 β€” research archive (round 5, arm A), NOT delivered

⚠️ Research archive β€” NOT a delivered model. This checkpoint failed its round's acceptance gates and was never shipped. The current production head is slow-stack/laya-nli-memory-conflict (v4). Uploaded 2026-10-02 for provenance/backup while round 11 (multi-run verdict protocol) waits for Kaggle GPU quota.

Round 5 (2026-09-28) of the laya NLI memory-conflict head program by modusensus ran two arms on the same corpus: arm A = the v4 recipe (RL-style term + CE), this checkpoint; arm B = pure CE (see laya-nli-conflict-v5-ce).

Headline results (frozen 1000-pair main val unless noted)

  • main val 0.893 (gate β‰₯0.896, βœ—), old-20 acceptance 19/20 (βœ—), negation 4/5 (βœ—), new-10 8/10 (βœ—)
  • passed: polarity diagnostics, val_soft 1 error/300
  • Key finding: the RL term clashes with graded soft targets β€” arm B (pure CE) overtook arm A on every axis β‡’ from v6 on, the main arm is pure CE

Artifacts

file value
model.safetensors SHA256 8ac971b1…40f413 (full hash in archive_sha256_manifest.txt)
rl_agent_config.json Ο„(noul) = 1.1384; encoder jhu-clsp/mmBERT-base; bf16
metrics.json val_accuracy 0.893, val_ece 0.0321, n_val 1000, no_rl false
val_probs.json frozen-val probability dump (calibration analyses)

Provenance

  • Training: Kaggle GPU kernel daphnelaurent/laya-nli-memory-conflict-fine-tune v10, dataset daphnelaurent/nli-conflict-pairs v10
  • Round record & full gate table: kaggle_eval/HANDOFF_NLI_V5.md
  • Base: fine-tuned from convaiinnovations/laya-multilingual lineage; bf16 weights; checkpoint_latest/ (optimizer/step state) intentionally not uploaded
Downloads last month

-

Downloads are not tracked for this model. How to track
Safetensors
Model size
0.3B params
Tensor type
F16
Β·
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for slow-stack/laya-nli-conflict-v5

Finetuned
(66)
this model

Dataset used to train slow-stack/laya-nli-conflict-v5