Repository: localaiLicense: apache-2.0

OvisOCR2 is ATH-MaaS's compact 0.8B vision-language model for page-level document parsing. It converts document images into Markdown while preserving formulas as LaTeX, tables as HTML, and the original reading order. This default entry uses the Q4_K_M GGUF with the F16 vision projector.
Links
Tags

OvisOCR2 in the higher-fidelity Q8_0 GGUF format with the F16 vision projector.
Links
Tags