Search
12 articles for “MFCC”
-
Automatic Baby Cry Detector with sleep music player (ABCD)
Abstract: In today’s world, our lives have become moredependent on technology. One problem which caught our eyesis when parents have to leave their wards (aged between 3months to 2 years) alone due to some essential tasks for a shortduration, the baby goes unmonitored. Normally, parents dothis by leaving their child asleep. In many cases, the childwakes up and starts to cry. In absence of loved ones, babiesneed immediate calmness relief. A …
Published in Journal of Mechatronics and Automation · Vol. 10, Issue 1, 2023 · pp. 21–29 Read article
-
Optimized Residual Neural Network for Audio Spoof Detection in Speaker Verification Systems
Abstract: The Automatic speaker verification system is a type of biometric technology that utilizes speech to determine if a person is an authentic user or not. Unfortunately, such systems can be susceptible to audio-spoofing attacks. The proposed work deals with this problem of Audio Spoofing by employing Residual Networks to determine if a voice signal is bonafide or not. A comparative study of Mel-frequency Cepstral coefficients (MFCC), Constant Q Cepstral Coefficients …
Published in Current Trends in Signal Processing · Vol. 12, Issue 2, 2022 · pp. 19–32 Read article
-
Exploiting Phase of Speech Signal for Speaker Recognition
Abstract: Abstract—Performance of speaker recognition system with feature based on temporal phase is presented in this paper. In the state of the art spectral feature, Mel Frequency Cepstal Coefficient (MFCC) only magnitude of the Fourier Transform of speech signal is considered while phase is ignored. The cepstral coefficients extracted from temporal phase (MFTPC) of speech signal are used as features for speaker recognition system. The performance of MFTPC feature is evaluated …
Published in Journal of Communication Engineering & Systems · Vol. 9, Issue 2, 2019 · pp. 81–87 Read article
-
Classifying Abnormalities in Heartbeat Sound
Abstract: Heartbeat sounds play a major role in the detection of various diseases such as heart disease, hyperthyroidism, and high blood pressure in their early stages. In the proposed method, various abnormal and healthy heartbeat audio signals are given as input and the features are extracted using MFCC (mel-frequency cepstral coefficients). Then, a deep learning approach is applied in which the MFCC audio signals are sent to the CNN (convolutional neural …
Published in Research & Reviews: A Journal of Embedded System & Applications · Vol. 12, Issue 1, 2024 · pp. 24–31 Read article
-
Development of Polymer-Based Sensors for Speech Emotion Recognition
Abstract: Traditional SER research often utilizes microphones with polymer components like Diaphragms and Membranes. Within some microphone designs, polymer membranes which plays a crucial role in converting sound pressure into electrical signals. The paper highlights the application (speech emotion recognition) and have tried to find polymer-based sensors. This work further delves deeper, investigating the performance of the CatBoost algorithm for emotion recognition in voice assistants designed for Indian languages. The research …
Published in Journal of Polymer & Composites · Vol. 12, Issue 5, 2024 · pp. 268–274 Read article
-
Principal Component Analysis Based Dominant Features Selection Method for Speaker Identification
Abstract: Cepstrum based features are mostly used in speaker identification. Mel-frequency cepstrum coefficients (MFCCs) and their statistical properties (skewness, kurtosis and standard deviation) are used in this paper for text-dependent speaker identification. Principal component analysis (PCA) is employed to select the dominant feature vector representing the speaker characteristics. Multi-layer neural network is used as the classification engine. There occurs the inter-speaker variation of speech length uttering the same word. The feature …
Published in Current Trends in Signal Processing · Vol. 1, Issue 1-3, 2025 · pp. 33–45 Read article
-
Emotion Recognition based on Human voice tones
Abstract: In this period of technology and automation, human interaction with machines has become an unavoidable occurrence. Machines are making our lives much easier, stretching their hands wherever it is possible for smoother operation of various tasks. Emotion recognition system is one of the system that can help humans for a much easier and meaningful interaction with machines. It can be used in various branches such as artificial intelligence, health care, …
Published in Current Trends in Information Technology · Vol. 10, Issue 2, 2020 · pp. 1–5 Read article
-
Bias Detection and Accuracy Enhancement in Voice-based Banking Authentication Using Deep Learning
Abstract: Biometric systems have become an integral part of how many people access banking services today, and voice verification systems can be a secure and easy-to-use source of banking authentication that does not require any physical contact with the bank or any other person. From the security perspective, these systems would normally provide an effective means of identifying an individual but frequently exhibit bias with respect to demographics such as the …
Published in International Journal of Information Security Engineering · Vol. 4, Issue 2, 2026 Read article
-
Fuzzy Variable Frame Analysis for Speech Recognition
Abstract: AbstractRecent works in machine learning has focused on models such as support vector machine (SVM), artificial neural network (ANN) and long short-term memory (LSTM), for automatically controlling the generalization and parameterization of the optimization process. This paper presents a fuzzy interpretation frame analysis procedure using LSTM classifier for noisy speech at word level using thresholding and local maxima procedure at framing level for the recognition process. Front end MFCC procedure …
Published in Current Trends in Signal Processing · Vol. 9, Issue 3, 2019 · pp. 9–18 Read article
-
Performance Analysis of Speech Performance Analysis of Speech Recognition using Advanced Algorithm
Abstract: In this paper, the authors have worked on a speech processing technique wherein they took audio speech as input signal from user and by recognition process, they perceived recorded word. Specifically, the authors concentrated on the amplitude and zero cross of speech signal, and two of the most important and useful terms, namely MFCC and Mahalanobis distance, for the purpose of getting an efficient outcome. The authors evaluated the proposed …
Published in Current Trends in Signal Processing · Vol. 3, Issue 1, 2013 · pp. 20–25 Read article
-
Extraction of speech Emotion Features Using MLP Classifier
Abstract: Speech Emotion Recognition is a thriving research topic. Speech emotion recognition use MLP classifier to categorize the emotions from the speech. This Speech is also used as the medium, through which one can express their feelings and mind state in human to machine interaction. The case is very easy where two humans communicate along with their emotions as by nature, they can recognize each other’s emotions. But for computer, if …
Published in Journal of Instrumentation Technology & Innovations · Vol. 12, Issue 3, 2022 · pp. 11–19 Read article
-
Identification of English Dialects and Emotions using Spectral and Prosodic Features of Speech Signal Processing
Abstract: AbstractIn this paper, the authors have explored speech features to identify English dialects and emotions. A dialect is any distinguishable variety of a language spoken by a group of people. Emotions provide naturalness to speech. Speech database considered for dialect identification task consists of spontaneous speech spoken by male and female speakers. The emotions considered in this study are anger, disgust, fear, happy, neutral and sad. Prosodic and spectral features …
Published in Journal of Computer Technology & Applications · Vol. 4, Issue 2, 2013 · pp. 10–17 Read article