错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Enhanced Sound Recognition and Classification Through Spectrogram Analysis, MEMS Sensors, and PyTorch: A Comprehensive Approach

  • Alexandros Spournias,
  • Nikolaos Nanos,
  • Evanthia Faliagka,
  • Christos Antonopoulos,
  • Nikolaos Voros,
  • Giorgos Keramidas

摘要

The importance of sound recognition and classification systems in various fields has led researchers to seek innovative methods to address these challenges. In this paper, the authors propose a concise yet effective approach for sound recognition and classification by combining spectrogram analysis, Micro-Electro-Mechanical Systems (MEMS) sensors, and the Pytorch deep learning framework. This method utilizes the rich information in audio signals to develop a robust and accurate sound recognition and classification system. The authors outline a three-stage process: data acquisition, feature extraction, and classification. MEMS sensors are employed for data acquisition, offering advantages such as reduced noise, low power consumption, and enhanced sensitivity compared to traditional microphones. The acquired audio signals are then preprocessed and converted into spectrograms, visually representing the audio data’s frequency, amplitude, and temporal attributes. During feature extraction, the spectrograms are analyzed to extract significant features conducive to sound recognition and classification. The classification task is performed using a custom deep learning model in Pytorch, leveraging modern neural networks’ pattern recognition capabilities. The model is trained and validated on a diverse dataset of audio samples, ensuring its proficiency in recognizing and classifying various sound types. The experimental results demonstrate the effectiveness of the proposed method, surpassing existing techniques in sound recognition and classification performance. By integrating spectrogram analysis, MEMS sensors, and Pytorch, the authors present a compact yet powerful sound recognition system with potential applications in numerous domains, such as predictive maintenance, environmental monitoring, and personalized voice-controlled devices.