Co-occurring keywords
Papers
ESPnet-SE++: Speech Enhancement for Robust Speech Recognition, Translation, and Understanding
INTERSPEECH 2022
Improving Data Driven Inverse Text Normalization using Data Augmentation and Machine Translation
INTERSPEECH 2022
End-to-end model for named entity recognition from speech without paired training data
INTERSPEECH 2022
Evaluating User Perception of Speech Recognition System Quality with Semantic Distance Metric
INTERSPEECH 2022
When Is TTS Augmentation Through a Pivot Language Useful?
INTERSPEECH 2022
Mitigating bias against non-native accents
INTERSPEECH 2022
Leveraging Simultaneous Translation for Enhancing Transcription of Low-resource Language via Cross Attention Mechanism
INTERSPEECH 2022
ASR-Robust Natural Language Understanding on ASR-GLUE dataset
INTERSPEECH 2022
Leveraging Acoustic Contextual Representation by Audio-textual Cross-modal Learning for Conversational ASR
INTERSPEECH 2022
SKYE: More than a conversational AI
INTERSPEECH 2022
Improving Rare Word Recognition with LM-aware MWER Training
INTERSPEECH 2022
Listen, Adapt, Better WER: Source-free Single-utterance Test-time Adaptation for Automatic Speech Recognition
INTERSPEECH 2022
On-the-fly ASR Corrections with Audio Exemplars
INTERSPEECH 2022