Llama-3-Ko-OpenOrca

Model Details

Model Description

Original model: beomi/Llama-3-Open-Ko-8B (2024.04.24 버전)

Dataset: kyujinpy/OpenOrca-KO

Training details

Training: Axolotl을 μ΄μš©ν•΄ LoRA-8bit둜 4epoch ν•™μŠ΅ μ‹œμΌ°μŠ΅λ‹ˆλ‹€.

  • sequence_len: 4096
  • bf16

ν•™μŠ΅ μ‹œκ°„: A6000x2, 6μ‹œκ°„

Evaluation

  • 0 shot kobest
Tasks n-shot Metric Value Stderr
kobest_boolq 0 acc 0.5021 Β± 0.0133
kobest_copa 0 acc 0.6920 Β± 0.0146
kobest_hellaswag 0 acc 0.4520 Β± 0.0223
kobest_sentineg 0 acc 0.7330 Β± 0.0222
kobest_wic 0 acc 0.4881 Β± 0.0141
  • 5 shot kobest
Tasks n-shot Metric Value Stderr
kobest_boolq 5 acc 0.7123 Β± 0.0121
kobest_copa 5 acc 0.7620 Β± 0.0135
kobest_hellaswag 5 acc 0.4780 Β± 0.0224
kobest_sentineg 5 acc 0.9446 Β± 0.0115
kobest_wic 5 acc 0.6103 Β± 0.0137

License:

https://llama.meta.com/llama3/license

Downloads last month
10
Safetensors
Model size
8B params
Tensor type
BF16
Β·
Inference Providers NEW

Model tree for werty1248/Llama-3-Ko-8B-OpenOrca

Finetuned
(20)
this model
Quantizations
4 models

Dataset used to train werty1248/Llama-3-Ko-8B-OpenOrca

Spaces using werty1248/Llama-3-Ko-8B-OpenOrca 9