Model Gallery

2 models from 1 repositories

Filter by type:

Filter by tags:

ovisocr2
OvisOCR2 is ATH-MaaS's compact 0.8B vision-language model for page-level document parsing. It converts document images into Markdown while preserving formulas as LaTeX, tables as HTML, and the original reading order. This default entry uses the Q4_K_M GGUF with the F16 vision projector.

Repository: localaiLicense: apache-2.0

ovisocr2-q8
OvisOCR2 in the higher-fidelity Q8_0 GGUF format with the F16 vision projector.

Repository: localaiLicense: apache-2.0