Maxilicious20 commited on
Commit
fc1ea39
·
verified ·
1 Parent(s): c3f9730

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +12 -16
README.md CHANGED
@@ -5,7 +5,6 @@ tags:
5
  - llama.cpp
6
  - lm-studio
7
  - aether
8
- - aether-2.5
9
  - german
10
  - english
11
  - text-generation
@@ -15,15 +14,15 @@ language:
15
  - en
16
  ---
17
 
18
- # Aether 2.5 Pro - GGUF
19
 
20
- Pre-quantized GGUF binaries for **Aether 2.5 Pro**.
21
 
22
- Aether 2.5 Pro is a fine-tuned version of Qwen2.5-3B-Instruct, trained with SFT (Supervised Fine-Tuning) and PEFT (LoRA). It delivers improved reasoning, stronger instruction following, and better multilingual performance in German and English while remaining efficient for local use.
23
 
24
- > 🔗 **Looking for the Base / LoRA Adapter?**
25
- > If you want to use the Hugging Face Transformers PEFT adapter instead, check out the main repository:
26
- > 👉 **[Maxilicious20/Aether-2.5-Pro](https://huggingface.co/Maxilicious20/Aether-2.5-Pro)**
27
 
28
  ---
29
 
@@ -33,11 +32,9 @@ Choose the right file depending on your system's VRAM/RAM and performance needs:
33
 
34
  | Filename | Quantization | Quality | Size | Description / Recommendation |
35
  | :--- | :--- | :--- | :--- | :--- |
36
- | `aether-2.5-pro-f16.gguf` | FP16 / F16 | Maximum | ~6.0 GB | Uncompressed full precision. Best quality. |
37
- | `aether-2.5-pro-q8_0.gguf` | Q8_0 | Very High | ~3.2 GB | Near-lossless quantization. Excellent quality. |
38
- | `aether-2.5-pro-q5_k_m.gguf` | Q5_K_M | High | ~2.3 GB | High quality with good performance. |
39
- | `aether-2.5-pro-q4_k_m.gguf` | Q4_K_M | Balanced | ~1.9 GB | **Recommended.** Best balance of quality, speed and VRAM usage. |
40
- | `aether-2.5-pro-q3_k_m.gguf` | Q3_K_M | Good | ~1.5 GB | Lower VRAM usage, still usable quality. |
41
 
42
  ---
43
 
@@ -45,12 +42,11 @@ Choose the right file depending on your system's VRAM/RAM and performance needs:
45
 
46
  ### 1. LM Studio
47
  1. Open LM Studio.
48
- 2. Search for `Maxilicious20/Aether-2.5-Pro-GGUF` or paste the repository ID.
49
- 3. Download your preferred quantization (recommended: `aether-2.5-pro-q4_k_m.gguf`).
50
  4. Load the model and start chatting!
51
 
52
  ### 2. Ollama / llama.cpp
53
  You can run the GGUF file directly using `llama.cpp`:
54
-
55
  ```bash
56
- ./llama-cli -m aether-2.5-pro-q4_k_m.gguf -p "Hello Aether 2.5 Pro!" -n 256
 
5
  - llama.cpp
6
  - lm-studio
7
  - aether
 
8
  - german
9
  - english
10
  - text-generation
 
14
  - en
15
  ---
16
 
17
+ # Aether 2.2 Pro - GGUF
18
 
19
+ Pre-quantized GGUF binaries for **Aether 2.2 Pro**.
20
 
21
+ Trained with SFT (Supervised Fine-Tuning) and PEFT (LoRA) on a custom dataset using local NVIDIA RTX GPU acceleration, Aether 2.2 Pro delivers optimized performance, strong conversational capabilities, and reliable multilingual responses in German and English.
22
 
23
+ > 🔗 **Looking for the Base / LoRA Adapter?**
24
+ > If you want to use the Hugging Face Transformers PEFT adapter instead, check out the main repository:
25
+ > 👉 **[Maxilicious20/Aether-2.2-Pro](https://huggingface.co/Maxilicious20/Aether-2.2-Pro)**
26
 
27
  ---
28
 
 
32
 
33
  | Filename | Quantization | Quality | Size | Description / Recommendation |
34
  | :--- | :--- | :--- | :--- | :--- |
35
+ | `aether_2_2_pro_f16.gguf` | FP16 / F16 | Maximum | ~2.88 GB | Uncompressed full precision. Best quality. |
36
+ | `aether_2_2_pro_q8_0.gguf` | Q8_0 | Very High | ~1.53 GB | Near-lossless quantization. Excellent balance of precision and speed. |
37
+ | `aether_2_2_pro_q4_k_m.gguf` | Q4_K_M | Balanced | ~940 MB | **Recommended.** Lightweight, fast, and optimized for low VRAM/RAM setups. |
 
 
38
 
39
  ---
40
 
 
42
 
43
  ### 1. LM Studio
44
  1. Open LM Studio.
45
+ 2. Search for `Maxilicious20/Aether-2.2-Pro-GGUF` or paste the repository ID.
46
+ 3. Download your preferred quantization (e.g., `aether_2_2_pro_q4_k_m.gguf`).
47
  4. Load the model and start chatting!
48
 
49
  ### 2. Ollama / llama.cpp
50
  You can run the GGUF file directly using `llama.cpp`:
 
51
  ```bash
52
+ ./llama-cli -m aether_2_2_pro_q4_k_m.gguf -p "Hello Aether Pro!" -n 256