Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Djamel Addou

Transfer Learning-Based Deep Residual Learning for Speech Recognition in Clean and Noisy Environments

May 02, 2025

Noussaiba Djeffal, Djamel Addou, Hamza Kheddar, Sid Ahmed Selouani

Figure 1 for Transfer Learning-Based Deep Residual Learning for Speech Recognition in Clean and Noisy Environments

Figure 2 for Transfer Learning-Based Deep Residual Learning for Speech Recognition in Clean and Noisy Environments

Figure 3 for Transfer Learning-Based Deep Residual Learning for Speech Recognition in Clean and Noisy Environments

Figure 4 for Transfer Learning-Based Deep Residual Learning for Speech Recognition in Clean and Noisy Environments

Abstract:Addressing the detrimental impact of non-stationary environmental noise on automatic speech recognition (ASR) has been a persistent and significant research focus. Despite advancements, this challenge continues to be a major concern. Recently, data-driven supervised approaches, such as deep neural networks, have emerged as promising alternatives to traditional unsupervised methods. With extensive training, these approaches have the potential to overcome the challenges posed by diverse real-life acoustic environments. In this light, this paper introduces a novel neural framework that incorporates a robust frontend into ASR systems in both clean and noisy environments. Utilizing the Aurora-2 speech database, the authors evaluate the effectiveness of an acoustic feature set for Mel-frequency, employing the approach of transfer learning based on Residual neural network (ResNet). The experimental results demonstrate a significant improvement in recognition accuracy compared to convolutional neural networks (CNN) and long short-term memory (LSTM) networks. They achieved accuracies of 98.94% in clean and 91.21% in noisy mode.

* 2024 International Conference on Telecommunications and Intelligent Systems (ICTIS)

Via

Access Paper or Ask Questions

Automatic Speech Recognition with BERT and CTC Transformers: A Review

Oct 12, 2024

Noussaiba Djeffal, Hamza Kheddar, Djamel Addou, Ahmed Cherif Mazari, Yassine Himeur

Figure 1 for Automatic Speech Recognition with BERT and CTC Transformers: A Review

Figure 2 for Automatic Speech Recognition with BERT and CTC Transformers: A Review

Figure 3 for Automatic Speech Recognition with BERT and CTC Transformers: A Review

Figure 4 for Automatic Speech Recognition with BERT and CTC Transformers: A Review

Abstract:This review paper provides a comprehensive analysis of recent advances in automatic speech recognition (ASR) with bidirectional encoder representations from transformers BERT and connectionist temporal classification (CTC) transformers. The paper first introduces the fundamental concepts of ASR and discusses the challenges associated with it. It then explains the architecture of BERT and CTC transformers and their potential applications in ASR. The paper reviews several studies that have used these models for speech recognition tasks and discusses the results obtained. Additionally, the paper highlights the limitations of these models and outlines potential areas for further research. All in all, this review provides valuable insights for researchers and practitioners who are interested in ASR with BERT and CTC transformers.

* 2023 2nd International Conference on Electronics, Energy and Measurement (IC2EM)

Via

Access Paper or Ask Questions