VisualArchiveSystem commited on
Commit
2d28770
·
verified ·
1 Parent(s): 255708b

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +39 -0
README.md ADDED
@@ -0,0 +1,39 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: apache-2.0
3
+ tags:
4
+ - vas
5
+ - llama.cpp
6
+ - gguf
7
+ - photography
8
+ - desktop-ai
9
+ ---
10
+
11
+ # VAS Pro — AI Models
12
+
13
+ Pre-quantized GGUF models for VAS Pro (Visual Archive System) local AI assistant.
14
+
15
+ ## Models
16
+
17
+ | Model | File | Size | Purpose | Original |
18
+ |-------|------|------|---------|----------|
19
+ | Phi-4 Mini | `phi-4-mini-Q4_K_M.gguf` | ~2.5 GB | Fast responses, greetings | [microsoft/Phi-4-mini-instruct](https://huggingface.co/microsoft/Phi-4-mini-instruct) |
20
+ | Gemma 3 4B | `gemma-4-4b-it-Q4_K_M.gguf` | ~2.5 GB | Standard tasks, tool calling | [google/gemma-3-4b-it](https://huggingface.co/google/gemma-3-4b-it) |
21
+ | Qwen2.5-VL 7B | `qwen3.5-9b-vision-Q4_K_M.gguf` | ~4.7 GB | Vision, OCR, image analysis | [Qwen/Qwen2.5-VL-7B-Instruct](https://huggingface.co/Qwen/Qwen2.5-VL-7B-Instruct) |
22
+ | MxBAI Embed Large | `mxbai-embed-large-v1-f16.gguf` | ~670 MB | Semantic search embeddings | [mixedbread-ai/mxbai-embed-large-v1](https://huggingface.co/mixedbread-ai/mxbai-embed-large-v1) |
23
+
24
+ ## Usage
25
+
26
+ These models are automatically downloaded by VAS Pro on first run. No manual setup required.
27
+
28
+ ## Quantization
29
+
30
+ - Text models use **Q4_K_M** quantization (best quality/size ratio for 4-bit)
31
+ - Embedding model uses **F16** (full precision for maximum retrieval accuracy)
32
+
33
+ ## License
34
+
35
+ Models retain their original licenses:
36
+ - Phi-4 Mini: MIT License
37
+ - Gemma 3: [Gemma Terms of Use](https://ai.google.dev/gemma/terms)
38
+ - Qwen2.5-VL: Apache 2.0
39
+ - MxBAI Embed: Apache 2.0