Adding Evaluation Results

This is an automated PR created with https://huggingface.co/spaces/Weyaxi/open-llm-leaderboard-results-pr

The purpose of this PR is to add evaluation results from the Open LLM Leaderboard to your model card.

If you encounter any issues, please report them to https://huggingface.co/spaces/Weyaxi/open-llm-leaderboard-results-pr/discussions

Files changed (1) hide show

README.md +117 -1

README.md CHANGED Viewed

@@ -1,5 +1,108 @@
 ---
 license: mit
 ---
 This is a fine-tuned 13B chat model
@@ -74,4 +177,17 @@ This incident marked a turning point in Ava's crusade for environmental protecti
 Today, Yosemite National Park remains a testament to the power of individual actions combined with collective effort. Its pristine landscapes continue inspiring countless visitors each year, reminding us all that we have a responsibility towards safeguarding our planet for future generations. And amidst these stunning vistas stands Ava, proudly carrying forth John Muir's legacy, ensuring that his dream of preserving nature lives on forever.</s>
-```

 ---
 license: mit
+model-index:
+- name: Mixtral_13B_Chat
+  results:
+  - task:
+      type: text-generation
+      name: Text Generation
+    dataset:
+      name: AI2 Reasoning Challenge (25-Shot)
+      type: ai2_arc
+      config: ARC-Challenge
+      split: test
+      args:
+        num_few_shot: 25
+    metrics:
+    - type: acc_norm
+      value: 67.41
+      name: normalized accuracy
+    source:
+      url: https://huggingface.co/spaces/HuggingFaceH4/open_llm_leaderboard?query=cloudyu/Mixtral_13B_Chat
+      name: Open LLM Leaderboard
+  - task:
+      type: text-generation
+      name: Text Generation
+    dataset:
+      name: HellaSwag (10-Shot)
+      type: hellaswag
+      split: validation
+      args:
+        num_few_shot: 10
+    metrics:
+    - type: acc_norm
+      value: 85.87
+      name: normalized accuracy
+    source:
+      url: https://huggingface.co/spaces/HuggingFaceH4/open_llm_leaderboard?query=cloudyu/Mixtral_13B_Chat
+      name: Open LLM Leaderboard
+  - task:
+      type: text-generation
+      name: Text Generation
+    dataset:
+      name: MMLU (5-Shot)
+      type: cais/mmlu
+      config: all
+      split: test
+      args:
+        num_few_shot: 5
+    metrics:
+    - type: acc
+      value: 64.54
+      name: accuracy
+    source:
+      url: https://huggingface.co/spaces/HuggingFaceH4/open_llm_leaderboard?query=cloudyu/Mixtral_13B_Chat
+      name: Open LLM Leaderboard
+  - task:
+      type: text-generation
+      name: Text Generation
+    dataset:
+      name: TruthfulQA (0-shot)
+      type: truthful_qa
+      config: multiple_choice
+      split: validation
+      args:
+        num_few_shot: 0
+    metrics:
+    - type: mc2
+      value: 58.98
+    source:
+      url: https://huggingface.co/spaces/HuggingFaceH4/open_llm_leaderboard?query=cloudyu/Mixtral_13B_Chat
+      name: Open LLM Leaderboard
+  - task:
+      type: text-generation
+      name: Text Generation
+    dataset:
+      name: Winogrande (5-shot)
+      type: winogrande
+      config: winogrande_xl
+      split: validation
+      args:
+        num_few_shot: 5
+    metrics:
+    - type: acc
+      value: 80.43
+      name: accuracy
+    source:
+      url: https://huggingface.co/spaces/HuggingFaceH4/open_llm_leaderboard?query=cloudyu/Mixtral_13B_Chat
+      name: Open LLM Leaderboard
+  - task:
+      type: text-generation
+      name: Text Generation
+    dataset:
+      name: GSM8k (5-shot)
+      type: gsm8k
+      config: main
+      split: test
+      args:
+        num_few_shot: 5
+    metrics:
+    - type: acc
+      value: 56.63
+      name: accuracy
+    source:
+      url: https://huggingface.co/spaces/HuggingFaceH4/open_llm_leaderboard?query=cloudyu/Mixtral_13B_Chat
+      name: Open LLM Leaderboard
 ---
 This is a fine-tuned 13B chat model
 Today, Yosemite National Park remains a testament to the power of individual actions combined with collective effort. Its pristine landscapes continue inspiring countless visitors each year, reminding us all that we have a responsibility towards safeguarding our planet for future generations. And amidst these stunning vistas stands Ava, proudly carrying forth John Muir's legacy, ensuring that his dream of preserving nature lives on forever.</s>
+```
+# [Open LLM Leaderboard Evaluation Results](https://huggingface.co/spaces/HuggingFaceH4/open_llm_leaderboard)
+Detailed results can be found [here](https://huggingface.co/datasets/open-llm-leaderboard/details_cloudyu__Mixtral_13B_Chat)
+|             Metric              |Value|
+|---------------------------------|----:|
+|Avg.                             |68.98|
+|AI2 Reasoning Challenge (25-Shot)|67.41|
+|HellaSwag (10-Shot)              |85.87|
+|MMLU (5-Shot)                    |64.54|
+|TruthfulQA (0-shot)              |58.98|
+|Winogrande (5-shot)              |80.43|
+|GSM8k (5-shot)                   |56.63|