jnjj commited on
Commit
a4a16a6
·
verified ·
1 Parent(s): c640811

Update README.md via script

Browse files
Files changed (1) hide show
  1. README.md +6 -6
README.md CHANGED
@@ -23,14 +23,14 @@ The fully merged model weights and tokenizer are updated periodically at the roo
23
  - **Dynamic Dataset Source:** The script iterates through a wide array of Hugging Face Hub datasets.
24
  - **Rapid Iteration Strategy:** Training per dataset configuration is brief (`max_steps=1`), prioritizing breadth of exposure over depth on any single dataset.
25
  ## Training Progress
26
- - **Datasets Processed (Successfully trained on at least one config):** 8
27
- - **Text Examples Streamed (Total):** 48
28
- - **Tokens Processed (Total):** 24576
29
- - **Last Successful Model Update:** 2025-05-08 15:48:23 UTC
30
  ### Evaluation Snapshot (Approximate)
31
 
32
- - **Current Perplexity (wikitext Subset):** 284.44
33
- - **Perplexity Change:** `-0.02` ⬇️ (vs previous cycle's perplexity)
34
 
35
  #### Generated Examples (Qualitative Assessment)
36
 
 
23
  - **Dynamic Dataset Source:** The script iterates through a wide array of Hugging Face Hub datasets.
24
  - **Rapid Iteration Strategy:** Training per dataset configuration is brief (`max_steps=1`), prioritizing breadth of exposure over depth on any single dataset.
25
  ## Training Progress
26
+ - **Datasets Processed (Successfully trained on at least one config):** 9
27
+ - **Text Examples Streamed (Total):** 54
28
+ - **Tokens Processed (Total):** 27648
29
+ - **Last Successful Model Update:** 2025-05-08 15:50:02 UTC
30
  ### Evaluation Snapshot (Approximate)
31
 
32
+ - **Current Perplexity (wikitext Subset):** 284.24
33
+ - **Perplexity Change:** `-0.20` ⬇️ (vs previous cycle's perplexity)
34
 
35
  #### Generated Examples (Qualitative Assessment)
36