Open access
Aug 2026
Deep learning-based multi-speaker separation and speech enhancement for forensic audio analysis
A hybrid deep learning framework for multi-speaker separation and speech enhancement by integrating Gated Convolutional Neural Networks (GCNNs) for speech source separation with Long Short-Term Memory (LSTM) networks for temporal speech enhancement is proposed.
O. Ojerinde, R. Mudasiru, Ramatu Abubakar et al.
· Nature Journal of Emerging S... · 0 citations