Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
gaaaaaaaaaaa
/
MultimodalReasoning3B
Like
0
Safetensors
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
MultimodalReasoning3B
/
data
/
ReasoningDPO
500 kB
Ctrl+K
Ctrl+K
1 contributor
History:
2 commits
gaaaaaaaaaaa
SFT Finetuned - Pre-DPO
0f7217b
verified
5 months ago
actual_yap_rejected_dpo.jsonl
Safe
189 kB
SFT Finetuned - Pre-DPO
5 months ago
format_cleanup_dpo.jsonl
Safe
49 kB
SFT Finetuned - Pre-DPO
5 months ago
logical_reasoning_dpo.jsonl
Safe
60 kB
SFT Finetuned - Pre-DPO
5 months ago
math_reasoning_dpo.jsonl
Safe
129 kB
SFT Finetuned - Pre-DPO
5 months ago
no_thinking_dpo.jsonl
Safe
33.7 kB
SFT Finetuned - Pre-DPO
5 months ago
uncertainties_dpo.jsonl
Safe
38.4 kB
SFT Finetuned - Pre-DPO
5 months ago