Skip to content

Author

Sajjad Ghiasvand

We have 6 of 15 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

Kernel-Based Steering of CLIP with Vision-Language Model Preferences

Large vision-language models (VLMs) can judge visual similarity, but their judgments are not directly available as compact image embeddings for efficient comparison. We study how to transfer these preferences into CLIP while retaining its image--text capabilities. We introduce ASK, a kernel-based steering method that l...

Sajjad Ghiasvand, Haniyeh Ehsani Oskouie, Sina Mansouri et al. · 0 citations
Preprint Sep 2026

Low-Rank Prompt Learning for Vision-Language Models with Fixed-Token Bases

Across seven few-shot benchmarks and two CLIP backbones, low-rank prompts match or improve dense CoOp at far fewer parameters, with the clearest gains on low-shot base-to-new generalization.

Tanvir Muntakim Tonoy, Sajjad Ghiasvand, Mahnoosh Alizadeh et al. · 0 citations
Preprint Aug 2026

ZOMP: Zeroth-Order Multi-Modal Prompt Tuning for Vision-Language Models

This work proposes ZOMP (Zeroth-Order Multimodal Prompt tuning), a query-efficient, fully forward-only method that tunes deep prompts in both the vision and text branches of a frozen CLIP model using simultaneous perturbation stochastic approximation.

Sajjad Ghiasvand, Yifan Yang, Mahnoosh Alizadeh et al. · 1 citation

Can MLLMs Critique Like Humans? Evaluating Open-Ended Aesthetic Reasoning in Multimodal Large Language Models

It is argued that reference-based similarity rewards a fluent, comprehensive critique style rather than the selectivity and specificity of human critique, and that reference-based similarity gives a misleading picture.

Sajjad Ghiasvand, Maryam Amirizaniani, Haniyeh Ehsani Oskouie et al. · 1 citation

REALM: Reliable Expertise-Aware Language Model Fine-Tuning from Noisy Annotations

REALM is proposed, which jointly learns the model parameters and a scalar expertise value for each annotator, entirely unsupervised and requiring nothing beyond annotator identity, and extends to multiple tasks via a learned expertise matrix.

Sajjad Ghiasvand, M. Beliaev, Mahnoosh Alizadeh et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.