PsOCR: Benchmarking Large Multimodal Models for Optical Character Recognition in Low-resource Pashto Language Paper • 2505.10055 • Published May 15, 2025 • 3
Qwen2.5-VL Collection Vision-language model series based on Qwen2.5 • 10 items • Updated 30 days ago • 559
Qwen2-VL Collection Vision-language model series based on Qwen2 • 15 items • Updated 30 days ago • 231