Cubex11 commited on
Commit
e73d146
·
verified ·
1 Parent(s): 014d91c

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +2 -0
README.md CHANGED
@@ -17,6 +17,8 @@ datasets:
17
  - HuggingFaceH4/rlaif-v_formatted
18
  ---
19
 
 
 
20
  # Solari: Hallucination-Reduced Vision Language Model
21
 
22
  Solari is a 500M parameter vision-language model fine-tuned for **reduced hallucination** on real-world images. Built on [SmolVLM2-500M-Video-Instruct](https://huggingface.co/HuggingFaceTB/SmolVLM2-500M-Video-Instruct), Solari uses **QLoRA + Direct Preference Optimization (DPO)** on the [RLAIF-V](https://huggingface.co/datasets/HuggingFaceH4/rlaif-v_formatted) dataset to align the model toward more faithful visual descriptions.
 
17
  - HuggingFaceH4/rlaif-v_formatted
18
  ---
19
 
20
+ ![Model Logo](thumbnail.png)
21
+
22
  # Solari: Hallucination-Reduced Vision Language Model
23
 
24
  Solari is a 500M parameter vision-language model fine-tuned for **reduced hallucination** on real-world images. Built on [SmolVLM2-500M-Video-Instruct](https://huggingface.co/HuggingFaceTB/SmolVLM2-500M-Video-Instruct), Solari uses **QLoRA + Direct Preference Optimization (DPO)** on the [RLAIF-V](https://huggingface.co/datasets/HuggingFaceH4/rlaif-v_formatted) dataset to align the model toward more faithful visual descriptions.