danish-foundation-models
/

dfm-decoder-open-v0-7b-pt

Text Generation

Model card Files Files and versions

peter-sk commited on Nov 1, 2025

Commit

0d1a559

·

1 Parent(s): c672506

updated model card

Files changed (1) hide show

README.md +6 -0

README.md CHANGED Viewed

@@ -10,6 +10,12 @@ base_model:
 - common-pile/comma-v0.1-2t
 pipeline_tag: text-generation
 ---
 | Stage | Batch size | Steps | HF path | Data mix | Comments |
 |-|-|-|-|-|-|

 - common-pile/comma-v0.1-2t
 pipeline_tag: text-generation
 ---
+# Munin-7B-Open-pt
+Munin-7B-open-pt is a 7 billion parameter language model continually pre-trained from [Comma v0.1-2T](https://huggingface.co/common-pile/comma-v0.1-2t/) using 30B tokens using a mix of the [Dynaword](https://huggingface.co/datasets/danish-foundation-models/danish-dynaword) and [the Comma v0.1 dataset](https://huggingface.co/datasets/common-pile/comma_v0.1_training_dataset), comprising only public domain and openly licensed data.
+Munin-7B-open-pt is a base model that can be used a the starting point for fine-tuning and post-training. It has not been instruction-tuned and cannot directly be expected to function as a chat model.
+Munin-7B-open-pt has been trained using the [maester](https://github.com/rlrs/maester) framework developed as part of the [Danish Foundation Models project](https://foundationmodels.dk/). The three pre-training stages are detailed in the following table:
 | Stage | Batch size | Steps | HF path | Data mix | Comments |
 |-|-|-|-|-|-|