Skip to content

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

CrowdCue: Specialist-Cue Conditioning for Vision-Language Crowd Counting

Generative vision-language models (VLMs) offer a counting paradigm in which one model produces both a count and a natural-language account of the scene, yet their raw counting accuracy sits in the range of sub-million-parameter specialist regressors. The open question is whether auxiliary guidance from a pretrained spe...

M. Farazi, B. Ciftler, Abdulhalim Dandoush et al. · 0 citations
Preprint Aug 2026

When Do VLMs Help Arabic Manuscript OCR? A Cross-Dataset Study

Vision-language models (VLMs) are increasingly being used for document understanding, yet their role in Arabic and Islamic manuscript recognition remains underexplored. To address such a gap in this paper, we evaluate traditional OCR, general-purpose VLMs, Arabic-specialized VLMs, and OCR-conditioned VLM correction acr...

M. Farazi, Firoj Alam, A. Maaradji et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.