

A vision-language model that outperforms GPT-5 on Arabic OCR
Baseer is a vision-language model built specifically for Arabic documents it outperforms GPT-5, Gemini 2.5 Pro, and Azure Document Intelligence on OCR benchmarks. Extracts text, tables (HTML), and equations (LaTeX) while preserving structure. API, on-prem, or web.
No comments yet. Be the first!
Real conversations about Baseer on X
Post on X