저가형 GPU에서 더 빠르게 문서를 파싱할 수 있는 jina-ocr-v1 모델 소개.
jina-ocr-v1은 스캔 문서와 PDF의 텍스트, 표, 수식을 읽어 Markdown으로 변환하는 34억 파라미터 OCR 모델로, DeepSeek-OCR을 기반으로 정확도와 처리 속도를 개선했습니다. FastMTP 기술을 통해 출력 결과를 변경하지 않고도 생성 시간을 줄여 저가형 GPU에서도 높은 성능을 발휘합니다.
Introduction of jina-ocr-v1 model for faster document parsing on low-cost GPUs.
jina-ocr-v1 is a 3.4 billion parameter OCR model that reads text, tables, and formulas from scanned documents and PDFs, converting them into Markdown. It improves accuracy and processing speed based on DeepSeek-OCR. With FastMTP technology, it reduces generation time without altering output results, enabling high performance even on low-cost GPUs.