jamesdumay commited on
Commit
eb2f570
·
verified ·
1 Parent(s): 3caa668

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +128 -0
README.md ADDED
@@ -0,0 +1,128 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ library_name: mesh-llm
3
+ base_model:
4
+ - "unsloth/Trinity-Large-Preview-GGUF"
5
+ pipeline_tag: text-generation
6
+ tags:
7
+ - gguf
8
+ - mesh-llm
9
+ - layer-package
10
+ - skippy
11
+ - distributed-inference
12
+ - local-inference
13
+ - openai-compatible
14
+ ---
15
+
16
+ <div align="center">
17
+ <a href="https://www.meshllm.cloud">
18
+ <img src="https://github.com/Mesh-LLM/mesh-llm/raw/main/docs/mesh-llm-logo.svg" alt="Mesh LLM" width="220">
19
+ </a>
20
+
21
+ <h1>Trinity-Large-Preview-UD-Q4_K_XL</h1>
22
+
23
+ <p>
24
+ <strong>Distributed GGUF inference package for Mesh LLM</strong>
25
+ </p>
26
+
27
+ <p>
28
+ <a href="https://www.meshllm.cloud"><img alt="Website" src="https://img.shields.io/badge/Website-meshllm.cloud-111111?style=for-the-badge"></a>
29
+ <a href="https://github.com/Mesh-LLM/mesh-llm"><img alt="GitHub" src="https://img.shields.io/badge/GitHub-Mesh--LLM-24292f?style=for-the-badge&logo=github"></a>
30
+ <a href="https://discord.gg/rs6fmc63eN"><img alt="Discord" src="https://img.shields.io/badge/Discord-Join-5865F2?style=for-the-badge&logo=discord&logoColor=white"></a>
31
+ </p>
32
+ </div>
33
+
34
+ GGUF layer package for running **Trinity-Large-Preview-UD-Q4_K_XL** across a local Mesh LLM cluster.
35
+
36
+ This package is derived from [unsloth/Trinity-Large-Preview-GGUF](https://huggingface.co/unsloth/Trinity-Large-Preview-GGUF) and keeps the original GGUF distribution split into per-layer artifacts for distributed inference.
37
+
38
+ ## Highlights
39
+
40
+ | Run locally | Pool multiple machines | OpenAI-compatible | Package variant |
41
+ |---|---|---|---|
42
+ | Private inference on your hardware | Split layers across peers | Serve `/v1/chat/completions` locally | `UD-Q4_K_XL` layer package |
43
+
44
+ ## Model Overview
45
+
46
+ | Property | Value |
47
+ |---|---|
48
+ | **Source model** | [unsloth/Trinity-Large-Preview-GGUF](https://huggingface.co/unsloth/Trinity-Large-Preview-GGUF) |
49
+ | **Model id** | `unsloth/Trinity-Large-Preview-GGUF:UD-Q4_K_XL` |
50
+ | **Family** | Trinity |
51
+ | **Parameter scale** | not recorded |
52
+ | **Quantization** | `UD-Q4_K_XL` |
53
+ | **Layer count** | 60 |
54
+ | **Activation width** | 3072 |
55
+ | **Package size** | 230.8 GB |
56
+ | **Source file** | `UD-Q4_K_XL/Trinity-Large-Preview-UD-Q4_K_XL-00001-of-00005.gguf` |
57
+ | **Package repo** | [meshllm/Trinity-Large-Preview-UD-Q4_K_XL-layers](https://huggingface.co/meshllm/Trinity-Large-Preview-UD-Q4_K_XL-layers) |
58
+
59
+ ## Recommended Use
60
+
61
+ - Local and private inference with Mesh LLM.
62
+ - Multi-machine serving when the full GGUF is too large for one host.
63
+ - OpenAI-compatible chat/completions workflows through Mesh LLM's local API.
64
+
65
+ For upstream architecture details, chat template guidance, sampling recommendations, license terms, and benchmark notes, see the source model card: [unsloth/Trinity-Large-Preview-GGUF](https://huggingface.co/unsloth/Trinity-Large-Preview-GGUF).
66
+
67
+ ## Quickstart
68
+
69
+ ```bash
70
+ # Run this on each machine that should contribute memory/compute.
71
+ mesh-llm serve --model "meshllm/Trinity-Large-Preview-UD-Q4_K_XL-layers" --split
72
+ ```
73
+
74
+ ```bash
75
+ # Check the mesh and discover the OpenAI-compatible model name.
76
+ curl -s http://localhost:3131/api/status
77
+ curl -s http://localhost:3131/v1/models
78
+ ```
79
+
80
+ ```bash
81
+ # Send an OpenAI-compatible chat request.
82
+ curl -s http://localhost:3131/v1/chat/completions \
83
+ -H "Content-Type: application/json" \
84
+ -d '{
85
+ "model": "unsloth/Trinity-Large-Preview-GGUF:UD-Q4_K_XL",
86
+ "messages": [{"role": "user", "content": "Write a tiny hello-world function in Rust."}],
87
+ "max_tokens": 128
88
+ }'
89
+ ```
90
+
91
+ ## Package Variant
92
+
93
+ | Property | Value |
94
+ |---|---|
95
+ | **Format** | `layer-package` |
96
+ | **Canonical source ref** | `unsloth/Trinity-Large-Preview-GGUF@main/UD-Q4_K_XL/Trinity-Large-Preview-UD-Q4_K_XL-00001-of-00005.gguf` |
97
+ | **Source revision** | `main` |
98
+ | **Source SHA-256** | `13632564d2e8a57be6a4bcde297cbecb97f5e9a40ffc06b8057877f3301b694d` |
99
+ | **Skippy ABI** | `0.1.22` |
100
+ | **Package manifest SHA-256** | `47046ac1365fee8e7797e23fb9d72449edea78b39429fdd15645673e24cbdc5c` |
101
+
102
+ ## What Is Included
103
+
104
+ | Artifact | Path | Contents | SHA-256 |
105
+ |---|---|---|---|
106
+ | Manifest | `model-package.json` | Package schema, source identity, checksums | `47046ac1365fee8e7797e23fb9d72449edea78b39429fdd15645673e24cbdc5c` |
107
+ | Metadata | `shared/metadata.gguf` | 0 tensors, 7.0 MB | `9ee98b39241a826751b809f398a339e1ce248c882be1aa69dd25987349a018da` |
108
+ | Embeddings | `shared/embeddings.gguf` | 1 tensors, 336.9 MB | `0a090b2ce10bf3f42027d6cfe98fa725d2c275e518cc1de0464fe12c10709b68` |
109
+ | Output head | `shared/output.gguf` | 2 tensors, 488.1 MB | `d9789c03eab663ee059faf1dfd1a6f2b36476f8407be5e7beece8572d8d5efe3` |
110
+ | Transformer layers | `layers/layer-*.gguf` | 60 layer artifacts, 1110 tensors, 230.0 GB | `see model-package.json` |
111
+
112
+ ## Validation
113
+
114
+ Generated by the Mesh LLM HF Jobs splitter from `mesh-llm` ref `main`.
115
+ Each artifact is checksummed as it is written, uploaded to this repository, and removed from the job workspace before the next artifact is produced.
116
+
117
+ ```bash
118
+ skippy-model-package write-package "/source/UD-Q4_K_XL/Trinity-Large-Preview-UD-Q4_K_XL-00001-of-00005.gguf" --out-dir "/tmp/meshllm-layer-job-meshllm_Trinity-Large-Preview-UD-Q4_K_XL-layers-198/package"
119
+ ```
120
+
121
+ ## Links
122
+
123
+ - Source model: [unsloth/Trinity-Large-Preview-GGUF](https://huggingface.co/unsloth/Trinity-Large-Preview-GGUF)
124
+ - Mesh LLM website: [meshllm.cloud](https://www.meshllm.cloud)
125
+ - Mesh LLM: [github.com/Mesh-LLM/mesh-llm](https://github.com/Mesh-LLM/mesh-llm)
126
+ - Discord: [discord.gg/rs6fmc63eN](https://discord.gg/rs6fmc63eN)
127
+ - Package catalog: [meshllm/catalog](https://huggingface.co/datasets/meshllm/catalog)
128
+ - Package format: [layer-package-repos.md](https://github.com/Mesh-LLM/mesh-llm/blob/main/docs/specs/layer-package-repos.md)