Hi everyone,
Based on your experience, which trained VLM performs best for extracting text from handwritten Arabic forms (e.g., applications, questionnaires, administrative documents)?
I’m mainly looking for the highest possible accuracy in:
-
Arabic handwritten text recognition
-
Structured form understanding (fields, tables, checkboxes)
-
Mixed printed and handwritten Arabic content
-
Ideally offline/self-hosted inference Have you tested any models such as Qwen-VL, Florence, PaliGemma, or Arabic-specific OCR/VLM models on this type of document? If possible, I’d appreciate benchmark results, model recommendations, and fine-tuning advice.
Thank you!