Co-occurring keywords
Papers
Multimodal Emotion Recognition Using Cross-Modal Attention and 1D Convolutional Neural Networks
INTERSPEECH 2020
Real-Time Single-Channel Deep Neural Network-Based Speech Enhancement on Edge Devices
INTERSPEECH 2020
Detection of Voicing and Place of Articulation of Fricatives with Deep Learning in a Virtual Speech and Language Therapy Tutor
INTERSPEECH 2020
Low-Latency Single Channel Speech Dereverberation Using U-Net Convolutional Neural Networks
INTERSPEECH 2020
Metadata-Aware End-to-End Keyword Spotting
INTERSPEECH 2020
Improved RawNet with Feature Map Scaling for Text-Independent Speaker Verification Using Raw Waveforms
INTERSPEECH 2020
Channel-Wise Subband Input for Better Voice and Accompaniment Separation on High Resolution Music
INTERSPEECH 2020
ATReSN-Net: Capturing Attentive Temporal Relations in Semantic Neighborhood for Acoustic Scene Classification
INTERSPEECH 2020
On Front-End Gain Invariant Modeling for Wake Word Spotting
INTERSPEECH 2020
AutoSpeech: Neural Architecture Search for Speaker Recognition
INTERSPEECH 2020
Exploring Deep Hybrid Tensor-to-Vector Network Architectures for Regression Based Speech Enhancement
INTERSPEECH 2020
Compressive sensing with un-trained neural networks: Gradient descent finds a smooth approximation
ICML 2020
Conformer: Convolution-augmented Transformer for Speech Recognition
INTERSPEECH 2020
Calibrating CNNs for Lifelong Learning
NIPS 2020