Co-occurring keywords
Papers
On the Prediction Network Architecture in RNN-T for ASR
INTERSPEECH 2022
DEFORMER: Coupling Deformed Localized Patterns with Global Context for Robust End-to-end Speech Recognition
INTERSPEECH 2022
On joint training with interfaces for spoken language understanding
INTERSPEECH 2022
DAVIS: Driver’s Audio-Visual Speech recognition
INTERSPEECH 2022
Towards Efficiently Learning Monotonic Alignments for Attention-based End-to-End Speech Recognition
INTERSPEECH 2022
Analysis of Self-Attention Head Diversity for Conformer-based Automatic Speech Recognition
INTERSPEECH 2022
End-to-End Spontaneous Speech Recognition Using Disfluency Labeling
INTERSPEECH 2022
PM-MMUT: Boosted Phone-mask Data Augmentation using Multi-Modeling Unit Training for Phonetic-Reduction-Robust E2E Speech Recognition
INTERSPEECH 2022
Voice2Alliance: Automatic Speaker Diarization and Quality Assurance of Conversational Alignment
INTERSPEECH 2022
Recent improvements of ASR models in the face of adversarial attacks
INTERSPEECH 2022
Cross-lingual Self-Supervised Speech Representations for Improved Dysarthric Speech Recognition
INTERSPEECH 2022
Investigating Self-supervised Pretraining Frameworks for Pathological Speech Recognition
INTERSPEECH 2022
End-to-End Multi-Loss Training for Low Delay Packet Loss Concealment
INTERSPEECH 2022
UserLibri: A Dataset for ASR Personalization Using Only Text
INTERSPEECH 2022
Automatic Learning of Subword Dependent Model Scales
INTERSPEECH 2022