-
tencent/HunyuanOCR
Image-Text-to-Text • 1B • Updated • 703k • 825 -
opendatalab/MinerU2.5-2509-1.2B
Image-Text-to-Text • 1B • Updated • 9.24k • 374 -
PaddlePaddle/PaddleOCR-VL-1.5
Image-Text-to-Text • 1.0B • Updated • 16.6k • 665 -
PaddlePaddle/PaddleOCR-VL
Image-Text-to-Text • 1.0B • Updated • 8.24k • 1.66k
valdanito
valdanito
AI & ML interests
None yet
Recent Activity
liked a model 2 days ago
HoppouAI/Breeze-TTS-2.cpp liked a model 2 days ago
BreezeBlue/Breeze-TTS-2 liked a model 3 days ago
Aero-Ex/Qwen-Image2.1_Normal2RGBOrganizations
None yet
rmbg
background removal model
data structuring
vlm
medical
-
Intelligent-Internet/II-Medical-8B-1706
Text Generation • 8B • Updated • 466 • • 141 -
Intelligent-Internet/II-Medical-8B
Text Generation • 8B • Updated • 915 • • 214 -
lingshu-medical-mllm/Lingshu-7B
Image-Text-to-Text • 8B • Updated • 8.86k • 82 -
unsloth/medgemma-4b-it-bnb-4bit
Image-Text-to-Text • 4B • Updated • 506 • 5
asr
-
FireRedTeam/FireRedASR-AED-L
Automatic Speech Recognition • Updated • 330 • 72 -
microsoft/Phi-4-multimodal-instruct
Automatic Speech Recognition • 6B • Updated • 267k • 1.62k -
Qwen/Qwen3-ASR-1.7B
Automatic Speech Recognition • 2B • Updated • 1.92M • • 1.12k -
Qwen/Qwen3-ASR-0.6B
Automatic Speech Recognition • 0.9B • Updated • 553k • • 365
tts
llm
retrieval
-
thenlper/gte-large-zh
Sentence Similarity • 0.3B • Updated • 6.41k • • 121 -
richinfoai/ritrieve_zh_v1
Sentence Similarity • 0.3B • Updated • 1.29k • • 48 -
boboliu/Qwen3-Embedding-4B-W4A16-G128
Feature Extraction • 4B • Updated • 414k • 5 -
boboliu/Qwen3-Embedding-0.6B-W4A16-G128
Feature Extraction • 0.6B • Updated • 86 • 4
ocr
-
tencent/HunyuanOCR
Image-Text-to-Text • 1B • Updated • 703k • 825 -
opendatalab/MinerU2.5-2509-1.2B
Image-Text-to-Text • 1B • Updated • 9.24k • 374 -
PaddlePaddle/PaddleOCR-VL-1.5
Image-Text-to-Text • 1.0B • Updated • 16.6k • 665 -
PaddlePaddle/PaddleOCR-VL
Image-Text-to-Text • 1.0B • Updated • 8.24k • 1.66k
asr
-
FireRedTeam/FireRedASR-AED-L
Automatic Speech Recognition • Updated • 330 • 72 -
microsoft/Phi-4-multimodal-instruct
Automatic Speech Recognition • 6B • Updated • 267k • 1.62k -
Qwen/Qwen3-ASR-1.7B
Automatic Speech Recognition • 2B • Updated • 1.92M • • 1.12k -
Qwen/Qwen3-ASR-0.6B
Automatic Speech Recognition • 0.9B • Updated • 553k • • 365
rmbg
background removal model
tts
data structuring
llm
vlm
retrieval
-
thenlper/gte-large-zh
Sentence Similarity • 0.3B • Updated • 6.41k • • 121 -
richinfoai/ritrieve_zh_v1
Sentence Similarity • 0.3B • Updated • 1.29k • • 48 -
boboliu/Qwen3-Embedding-4B-W4A16-G128
Feature Extraction • 4B • Updated • 414k • 5 -
boboliu/Qwen3-Embedding-0.6B-W4A16-G128
Feature Extraction • 0.6B • Updated • 86 • 4
medical
-
Intelligent-Internet/II-Medical-8B-1706
Text Generation • 8B • Updated • 466 • • 141 -
Intelligent-Internet/II-Medical-8B
Text Generation • 8B • Updated • 915 • • 214 -
lingshu-medical-mllm/Lingshu-7B
Image-Text-to-Text • 8B • Updated • 8.86k • 82 -
unsloth/medgemma-4b-it-bnb-4bit
Image-Text-to-Text • 4B • Updated • 506 • 5