VLM + OCR
-
Image-Text-to-Text • 0.9B • Updated • 708 • 82 -
erax-ai/EraX-VL-7B-V1.0
Image-Text-to-Text • 8B • Updated • 53 • 44 -
granite-docling-258M demo
📝282Convert and query documents from images with AI
-
datalab-to/chandra
Image-Text-to-Text • 9B • Updated • 66.4k • 532 -
deepseek-ai/DeepSeek-OCR
Image-Text-to-Text • 3B • Updated • 2.31M • 3.34k -
Multimodal OCR3
🌖69Chandra-OCR / Nanonets-OCR2 / olmOCR-2 / Dots.OCR
-
lightonai/LightOnOCR-2-1B
Image-Text-to-Text • 1B • Updated • 588k • 795 -
HuggingFaceFW/finepdfs
Viewer • Updated • 476M • 53.3k • 916
baidu/Qianfan-OCR
Image-Text-to-Text • 5B • Updated • 112k • 1.2kNote 4B direct image-to-Markdown conversion and supports a broad range of prompt-driven tasks — from structured document parsing and table extraction to chart understanding, document question answering, and key information extraction
tinixai/ocr_annual_financials
Viewer • Updated • 18.2k • 2k • 25Note báo cáo tài chính 10 năm vào dataset tinixai/ocr_annual_financials trên Hugging Face. Hiện tại dataset bao gồm: • 18.231 báo cáo tài chính • 1.491 mã chứng khoán • Dữ liệu từ 2015–2025 • ~26 triệu rows Parquet • ~194GB PDF + OCR text • OCR accuracy ~95% với số liệu và bảng biểu Đây có thể xem là một trong những bộ dữ liệu nguồn mở lớn nhất Việt Nam về: ✔ Financial AI ✔ OCR tiếng Việt ✔ Document AI ✔ Financial RAG ✔ Vietnamese LLM Dataset chứa: • Báo cáo tài chính hợp nhất • Báo cáo công ty mẹ •