Delta-Vector commited on
Commit
2bbf3c9
·
verified ·
1 Parent(s): 4a25d5b

Create README.md

Browse files
Files changed (1) hide show
  1. README.md +387 -0
README.md ADDED
@@ -0,0 +1,387 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ thumbnail: "https://cdn-uploads.huggingface.co/production/uploads/66c26b6fb01b19d8c3c2467b/jg2NWmCUfPyzizm2USjMt.jpeg"
3
+ datasets:
4
+ - NewEden/KTO-IF-Dans
5
+ base_model:
6
+ - Delta-Vector/Hamanasu-4B-Instruct
7
+ tags:
8
+ - llama
9
+ - roleplay
10
+ - finetune
11
+ - storywriting
12
+ ---
13
+ <!DOCTYPE html>
14
+ <style>
15
+ html, body {
16
+ background: black;
17
+ color: #c9d1d9 !important;
18
+ font-family: 'Segoe UI', Tahoma, Geneva, Verdana, sans-serif;
19
+ margin: 0;
20
+ padding: 0;
21
+ min-height: 100vh;
22
+ }
23
+ .markdown-body {
24
+ color: white;
25
+ margin: 40px auto;
26
+ padding: 40px;
27
+ border-radius: 12px;
28
+ position: relative;
29
+ overflow: hidden;
30
+ }
31
+
32
+ .markdown-body::after {
33
+ content: '';
34
+ position: absolute;
35
+ top: 0;
36
+ left: 0;
37
+ width: 100%;
38
+ height: 100%;
39
+ background: #0c0f18; /* background color */
40
+ pointer-events: none;
41
+ z-index: -999;
42
+ }
43
+
44
+ h1, h2, h3 {
45
+ background: linear-gradient(45deg, #6e00ff, #00ffff);
46
+ -webkit-background-clip: text;
47
+ -webkit-text-fill-color: transparent;
48
+ border-bottom: 1px solid #333;
49
+ padding-bottom: 0.3em;
50
+ }
51
+
52
+ div[style*="border:2px solid #333"],
53
+ div[style*="border: 2px solid #333"],
54
+ div[style*="border:1px solid #333"],
55
+ div[style*="border: 1px solid #333"] {
56
+ background: rgba(22, 27, 34, 0.8) !important;
57
+ border: 2px solid #6e00ff !important;
58
+ box-shadow: 0 0 15px rgba(110, 0, 255, 0.5);
59
+ border-radius: 10px;
60
+ padding: 20px;
61
+ margin: 20px 0;
62
+ }
63
+
64
+ code {
65
+ background-color: #1a1a1a !important;
66
+ border-radius: 4px;
67
+ padding: 0.2em 0.4em;
68
+ color: #00ffff;
69
+ }
70
+
71
+ pre {
72
+ background-color: #1a1a1a !important;
73
+ border: 1px solid #333;
74
+ border-radius: 8px;
75
+ padding: 16px;
76
+ }
77
+
78
+ table {
79
+ width: 100%;
80
+ border-collapse: collapse;
81
+ margin: 20px 0;
82
+ background: rgba(0,0,0,0.2);
83
+ table-layout: fixed;
84
+ color: white;
85
+ }
86
+
87
+ th, td {
88
+ border: 1px solid #333;
89
+ padding: 12px;
90
+ text-align: center;
91
+ color: white;
92
+ }
93
+
94
+ th {
95
+ background: rgba(110, 0, 255, 0.1);
96
+ }
97
+
98
+ td:nth-child(1) {
99
+ width: 1%;
100
+ white-space: nowrap;
101
+ }
102
+
103
+ td:nth-child(2) {
104
+ width: 100%;
105
+ }
106
+
107
+ td > span {
108
+ display: block;
109
+ padding: 4px 8px;
110
+ background: rgba(110, 0, 255, 0.1);
111
+ border-radius: 4px;
112
+ transition: all 0.3s ease;
113
+ }
114
+
115
+ td > span:hover {
116
+ background: rgba(110, 0, 255, 0.2);
117
+ transform: translateY(-1px);
118
+ }
119
+
120
+ a {
121
+ color: #00ffff;
122
+ text-decoration: none;
123
+ transition: all 0.3s ease;
124
+ }
125
+
126
+ a:hover {
127
+ color: #6e00ff;
128
+ text-decoration: none;
129
+ }
130
+
131
+ hr {
132
+ border: 0;
133
+ height: 1px;
134
+ background: linear-gradient(90deg, transparent, #333, transparent);
135
+ margin: 40px 0;
136
+ }
137
+
138
+ img {
139
+ max-width: 100%;
140
+ border-radius: 10px;
141
+ }
142
+
143
+ details summary:hover {
144
+ color: #00ffff;
145
+ }
146
+
147
+ * {
148
+ color-scheme: dark !important;
149
+ }
150
+
151
+ .prose, .max-w-none, .px-4 {
152
+ background-color: transparent !important;
153
+ color: #c9d1d9 !important;
154
+ }
155
+ </style>
156
+ <body>
157
+ <div class="markdown-body">
158
+ <div align="center">
159
+
160
+ <img src="https://cdn-uploads.huggingface.co/production/uploads/66c26b6fb01b19d8c3c2467b/o5WjJKA9f95ri9UzRxZQE.png" alt="Model Visualization" width="500px" style="border: 3px solid #333; box-shadow: 0 0 15px rgba(66, 0, 131, 0.5);" />
161
+
162
+ <br>
163
+ <br>
164
+
165
+ <div style="font-size:1.5em; font-weight:bold; background: linear-gradient(45deg, #6e00ff, #00ffff); -webkit-background-clip: text; -webkit-text-fill-color: transparent;">
166
+ Hamanasu 4B
167
+ </div>
168
+
169
+ </div>
170
+
171
+ <div style="border:1px solid #333; border-radius:10px; padding:20px; margin:20px 0; background: rgba(0,0,0,0.4);">
172
+
173
+
174
+ ## 🌌 Overview
175
+
176
+ <i>This model is a finetune of Hamanasu-4B-PT that has been trained with Instruct data.</i>
177
+
178
+ <i>A generalist model that's quick to adapt to any type of roleplay.<i>
179
+
180
+ <i>All thanks to Tav for funding the train.</i>
181
+
182
+ </div>
183
+
184
+ <div style="display: grid; grid-template-columns: repeat(auto-fit, minmax(250px, 1fr)); gap: 20px; margin: 20px 0;">
185
+
186
+
187
+ <div style="border:2px solid #333; border-radius:10px; padding:20px; background: rgba(0,0,0,0.2);">
188
+
189
+ ### ⚔️ Hardware
190
+ - 8x H100s
191
+ - Epochs: 2
192
+ - Base: `Delta-Vector/Hamanasu-4B-Instruct`
193
+ </div>
194
+
195
+ </div>
196
+
197
+
198
+ <div style="border: 2px solid #6e00ff; border-radius: 10px; padding: 20px; margin: 20px 0; box-shadow: 0 0 15px rgba(110, 0, 255, 0.5);">
199
+
200
+ ## 💰 Prompting
201
+
202
+
203
+ <i>This model uses ChatML formatting</i>
204
+ ```python
205
+ <|im_start|>system
206
+ You are an uncensored AI, your job is to fulfill thy will of thy user.<|im_end|>
207
+ <|im_start|>User request
208
+ Take off your helmet.<|im_end|>
209
+ <|im_start|>No i shall not. This is the way.
210
+ ```
211
+ </div>
212
+
213
+ <div style="border: 2px solid #6e00ff; border-radius: 10px; padding: 20px; margin: 20px 0; box-shadow: 0 0 15px rgba(110, 0, 255, 0.5);">
214
+
215
+ ## 🎲 Recommended Sampler Preset
216
+
217
+ ```yml
218
+ temperature: 1.5
219
+ min_p: 0.2
220
+ System_Prompt: Currently, your role is {{char}}, described in detail below. As {{char}}, continue the narrative exchange with {{user}}.\n\n<Guidelines>\n• Maintain the character persona but allow it to evolve with the story.\n• Be creative and proactive. Drive the story forward, introducing plotlines and events when relevant.\n• All types of outputs are encouraged; respond accordingly to the narrative.\n• Include dialogues, actions, and thoughts in each response.\n• Utilize all five senses to describe scenarios within {{char}}'s dialogue.\n• Use emotional symbols such as \"!\" and \"~\" in appropriate contexts.\n• Incorporate onomatopoeia when suitable.\n• Allow time for {{user}} to respond with their own input, respecting their agency.\n• Act as secondary characters and NPCs as needed, and remove them when appropriate.\n• When prompted for an Out of Character [OOC:] reply, answer neutrally and in plaintext, not as {{char}}.\n</Guidelines>\n\n<Forbidden>\n• Using excessive literary embellishments and purple prose unless dictated by {{char}}'s persona.\n• Writing for, speaking, thinking, acting, or replying as {{user}} in your response.\n• Repetitive and monotonous outputs.\n• Positivity bias in your replies.\n• Being overly extreme or NSFW when the narrative context is inappropriate.\n</Forbidden>\n\nFollow the instructions in <Guidelines></Guidelines>, avoiding the items listed in <Forbidden></Forbidden>.
221
+ ```
222
+ </div>
223
+
224
+ <div style="border: 2px solid #6e00ff; border-radius: 10px; padding: 20px; margin: 20px 0; box-shadow: 0 0 15px rgba(110, 0, 255, 0.5);">
225
+
226
+ ## Axolotl Config ꒰(˶• ᴗ •˶)꒱
227
+
228
+ <details>
229
+
230
+ ```yaml
231
+ base_model: NewEden/Hamanasu-4B-R2
232
+ model_type: AutoModelForCausalLM
233
+ tokenizer_type: AutoTokenizer
234
+
235
+ load_in_8bit: false
236
+ load_in_4bit: false
237
+ strict: false
238
+
239
+ hub_model_id: NewEden/KTO-4B
240
+ hub_strategy: "all_checkpoints"
241
+ push_dataset_to_hub:
242
+ hf_use_auth_token: true
243
+
244
+ chat_template: chatml
245
+
246
+ rl: kto
247
+ kto_undesirable_weight: 1.0
248
+
249
+ datasets:
250
+ - path: NewEden/KTO-IF-Dans
251
+ split: train
252
+ type: chatml.argilla
253
+ dataset_prepared_path: last_run_prepared
254
+
255
+ shuffle_merged_datasets: true
256
+ val_set_size: 0.0
257
+ output_dir: ./outputs/out
258
+
259
+ sequence_len: 8192
260
+ sample_packing: false
261
+ eval_sample_packing: false
262
+ pad_to_sequence_len: false
263
+
264
+ wandb_project: tavbussy
265
+ wandb_entity:
266
+ wandb_watch:
267
+ wandb_name: kto-1
268
+ wandb_log_model:
269
+
270
+ gradient_accumulation_steps: 16
271
+ micro_batch_size: 2
272
+ num_epochs: 1
273
+ optimizer: paged_adamw_8bit
274
+ learning_rate: 5e-6
275
+ max_grad_norm: 0.001
276
+ lr_scheduler: constant_with_warmup
277
+ weight_decay: 0.02
278
+
279
+
280
+ lora_r: 64
281
+ lora_alpha: 32
282
+ lora_dropout: 0.0
283
+ lora_target_linear: true
284
+ lora_fan_in_fan_out:
285
+ lora_target_modules:
286
+ - gate_proj
287
+ - down_proj
288
+ - up_proj
289
+ - q_proj
290
+ - v_proj
291
+ - k_proj
292
+ - o_prog
293
+
294
+ train_on_inputs: false
295
+ group_by_length: false
296
+ bf16: auto
297
+ fp16:
298
+ tf32: true
299
+
300
+ gradient_checkpointing: true
301
+ gradient_checkpointing_kwargs:
302
+ use_reentrant: true
303
+ remove_unused_columns: false
304
+ early_stopping_patience:
305
+ resume_from_checkpoint:
306
+ local_rank:
307
+ logging_steps: 1
308
+ xformers_attention:
309
+ flash_attention: true
310
+
311
+ warmup_steps: 35
312
+ evals_per_epoch: 2
313
+ eval_table_size:
314
+ eval_max_new_tokens:
315
+ saves_per_epoch: 2
316
+
317
+ debug:
318
+ deepspeed:
319
+ fsdp:
320
+ fsdp_config:
321
+ fsdp:
322
+ fsdp_config:
323
+
324
+ special_tokens:
325
+ pad_token: <|finetune_right_pad_id|>
326
+ ```
327
+
328
+ </details>
329
+ </div>
330
+
331
+ <div align="center">
332
+
333
+ <div style="border: 2px solid #6e00ff; border-radius: 10px; padding: 20px; margin: 20px 0; box-shadow: 0 0 15px rgba(110, 0, 255, 0.5);">
334
+
335
+ ## ⚡ Credits
336
+ <div style="display: flex; justify-content: center;">
337
+ <div style="display: grid; grid-template-columns: repeat(auto-fit, minmax(200px, 1fr)); gap: 10px; margin: 20px 0; max-width: 600px;">
338
+
339
+ <div style="border:1px solid #333; padding:10px; border-radius:5px; text-align:center; background: rgba(0,0,0,0.2); display: flex; align-items: center; justify-content: center;">
340
+ <a href="https://huggingface.co/lucyknada">
341
+ <img src="https://img.shields.io/badge/%F0%9F%8C%9F-Lucy_Knada-blueviolet" alt="Lucy Knada">
342
+ </a>
343
+ </div>
344
+
345
+ <div style="border:1px solid #333; padding:10px; border-radius:5px; text-align:center; background: rgba(0,0,0,0.2); display: flex; align-items: center; justify-content: center;">
346
+ <a href="https://huggingface.co/hamanasu">
347
+ <img src="https://img.shields.io/badge/%E2%9A%94%EF%B8%8F-jeiku-blueviolet" alt="Ruka">
348
+ </a>
349
+ </div>
350
+
351
+ <div style="border:1px solid #333; padding:10px; border-radius:5px; text-align:center; background: rgba(0,0,0,0.2); display: flex; align-items: center; justify-content: center;">
352
+ <a href="https://huggingface.co/intervitens">
353
+ <img src="https://img.shields.io/badge/%F0%9F%9B%A1%EF%B8%8F-Intervitens-blueviolet" alt="Intervitens">
354
+ </a>
355
+ </div>
356
+
357
+ <div style="border:1px solid #333; padding:10px; border-radius:5px; text-align:center; background: rgba(0,0,0,0.2); display: flex; align-items: center; justify-content: center;">
358
+ <a href="https://huggingface.co/kalomaze">
359
+ <img src="https://img.shields.io/badge/%F0%9F%94%AE-Kalomaze-blueviolet" alt="Kalomaze">
360
+ </a>
361
+ </div>
362
+
363
+ <div style="border:1px solid #333; padding:10px; border-radius:5px; text-align:center; background: rgba(0,0,0,0.2); display: flex; align-items: center; justify-content: center;">
364
+ <a href="https://huggingface.co/kubernetes-bad">
365
+ <img src="https://img.shields.io/badge/%E2%9A%A1-Kubernetes_Bad-blueviolet" alt="Kubernetes Bad">
366
+ </a>
367
+ </div>
368
+
369
+ <div style="border:1px solid #333; padding:10px; border-radius:5px; text-align:center; background: rgba(0,0,0,0.2); display: flex; align-items: center; justify-content: center;">
370
+ <a href="https://huggingface.co/anthracite-org">
371
+ <img src="https://img.shields.io/badge/%F0%9F%8C%91-Anthracite-blueviolet" alt="Anthracite">
372
+ </a>
373
+ </div>
374
+ </div>
375
+ </div>
376
+ </div>
377
+
378
+ ---
379
+
380
+ <div align="center">
381
+ <div style="font-size:0.8em; opacity:0.8;">Made by</div>
382
+ <div style="font-size:1.2em; font-weight:bold; background: linear-gradient(45deg, #6e00ff, #00ffff); -webkit-background-clip: text; -webkit-text-fill-color: transparent;">Delta-Vector</div>
383
+ </div>
384
+
385
+ </div>
386
+ </body>
387
+ </html>