Preprint
Jul 2026
Diarization-Guided Qwen-ASR Adaptation for Multilingual Two-Speaker Conversational Speech
Results show that supervised fine-tuning provides the largest gain, while synthetic-speech LoRA adaptation and reinforcement learning further improve robustness, while synthetic-speech LoRA adaptation and reinforcement learning further improve robustness.
Hao Wu, Rong-Qi Han, Zhen Wang et al.
· 0 citations