PP-OCRv6 Collection From 1.5M to 34.5M Parameters, Surpassing Billion-Scale VLMs on OCR Tasks • 20 items • Updated Aug 14 • 114
PaddleOCR-VL-1.6 Collection Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Training • 7 items • Updated Jul 8 • 19
ColPali: Efficient Document Retrieval with Vision Language Models Paper • 2407.01449 • Published Jun 27, 2024 • 51
Spotting LLMs With Binoculars: Zero-Shot Detection of Machine-Generated Text Paper • 2401.12070 • Published Jan 22, 2024 • 45