Upload README.md with huggingface_hub
Browse files
README.md
ADDED
|
@@ -0,0 +1,56 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 1 |
+
---
|
| 2 |
+
title: Leotsha
|
| 3 |
+
emoji: 🎮
|
| 4 |
+
colorFrom: blue
|
| 5 |
+
colorTo: purple
|
| 6 |
+
sdk: docker
|
| 7 |
+
pinned: false
|
| 8 |
+
tags:
|
| 9 |
+
- sepedi
|
| 10 |
+
- african-language
|
| 11 |
+
- sediba
|
| 12 |
+
- datafication
|
| 13 |
+
- edu
|
| 14 |
+
- game
|
| 15 |
+
- cdi
|
| 16 |
+
---
|
| 17 |
+
|
| 18 |
+
# Leotsha — 7-Mode CDI Datafication Surface
|
| 19 |
+
|
| 20 |
+
**Leotsha** (Leotsa la Setshaba) is the public-facing **7-mode datafication game surface** of the Sediba AI Sepedi platform — the interactive application layer where people play the 7 mode games that the CDI (Content/Domain Intelligence) layer produced.
|
| 21 |
+
|
| 22 |
+
Leotsha sits on top of the **Leotsa la Setshaba 7-model CDI/d-acidification layer** (the Sepedi-xlvi Content/Domain Intelligence layer held in the `leotsha_project` repo). It gives end users a Sepedi-native way to interact with and play the datafication modes that the CDI layer produced.
|
| 23 |
+
|
| 24 |
+
## What Leotsha does
|
| 25 |
+
|
| 26 |
+
- **7 mode games** — the interactive datafication surface of the Sediba Sepedi platform. Each mode is a game built on CDI-layer content/domain intelligence produced by the 7-model CDI layer.
|
| 27 |
+
- **Sepedi-first** — the whole surface is Sepedi-native; it is built for Sepedi speakers and Sepedi contexts first.
|
| 28 |
+
- **An entry point for the broader Sediba AI Sepedi stack** — datasets, models (sediba-XLM-R, Zabantu Nso sentiment), and the CDI training pipeline feed into the 7 modes.
|
| 29 |
+
|
| 30 |
+
## The 7 modes
|
| 31 |
+
|
| 32 |
+
The Space exposes the **7 modes** of the CDI datafication layer as interactive games. These 7 modes are the product of the 7-model CDI/d-acidification layer held in `leotsha_project`; Leotsha is the surface that makes them playable on the internet.
|
| 33 |
+
|
| 34 |
+
## The Sediba stack
|
| 35 |
+
|
| 36 |
+
| Layer | Component | Where |
|
| 37 |
+
|---|---|---|
|
| 38 |
+
| Application / game layer | **Leotsha** (this Space) — 7-mode datafication surface | `huggingface.co/spaces/Sediba-AI/leotsha` |
|
| 39 |
+
| CDI / model layer | Leotsa la Setshaba (conversational Sepedi) + sediba-XLM-R (MLM) + 7-model CDI layer | held in `leotsha_project` repo + Sediba vaults |
|
| 40 |
+
| Data layer | Sepedi training corpora + Sepedi-xlvi datasets | `huggingface.co/datasets/Sediba-AI/sepedi-training-v1` |
|
| 41 |
+
|
| 42 |
+
## Model family context
|
| 43 |
+
|
| 44 |
+
| Model | Role | Status | Where |
|
| 45 |
+
|---|---|---|--|
|
| 46 |
+
| **sediba-XLM-R** | Masked-language Sepedi NLU anchor | **Shipped** | `Sediba-AI/sediba-XLM-R` on HF |
|
| 47 |
+
| **Leotsa la Setshaba** | Conversational Sepedi (RLHF tier) | **Pending — publication holds until 1B+ base tier** | Sediba vaults; HF publication pending |
|
| 48 |
+
| **SedibaLM V7** | QLoRA experiment: Qwen2.5-1.5B + custom Sepedi vocab | **In training (Kaggle trail)** | Kaggle: `sedibaai/sedibalm-v7-sepedi-qlora` |
|
| 49 |
+
|
| 50 |
+
Leotsa la Setshaba and SedibaLM V7 are **not** published under this Space's model references. They are tracked separately in Sediba's vaults and training pipeline.
|
| 51 |
+
|
| 52 |
+
## Status
|
| 53 |
+
|
| 54 |
+
This Space is the **7-mode CDI datafication surface** in public — the application-layer game surface that Leotsha (the 7-model CDI layer) owns. The underlying models are at different stages: sediba-XLM-R is shipped; Leotsa la Setshaba publication holds until the 1B+ base tier is reached; SedibaLM V7 is in training on the Kaggle trail.
|
| 55 |
+
|
| 56 |
+
Built by Sediba AI | Mankweng, Limpopo.
|