Deep learning-based multi-speaker separation and speech enhancement for forensic audio analysis
A hybrid deep learning framework for multi-speaker separation and speech enhancement by integrating Gated Convolutional Neural Networks (GCNNs) for speech source separation with Long Short-Term Memory (LSTM) networks for temporal speech enhancement is proposed.