Search
21 articles for “speech signal processing”
-
Identification of English Dialects and Emotions using Spectral and Prosodic Features of Speech Signal Processing
Abstract: AbstractIn this paper, the authors have explored speech features to identify English dialects and emotions. A dialect is any distinguishable variety of a language spoken by a group of people. Emotions provide naturalness to speech. Speech database considered for dialect identification task consists of spontaneous speech spoken by male and female speakers. The emotions considered in this study are anger, disgust, fear, happy, neutral and sad. Prosodic and spectral features …
Published in Journal of Computer Technology & Applications · Vol. 4, Issue 2, 2013 · pp. 10–17 Read article
-
A Review Of Deep Learning Applications For Speech Processing Improvement
Abstract: Improve the quality of the spoken word is a common goal for many audio and speech signal processing applications. A noisy voice signal's quality and understandability may be improved via speech augmentation. Speech augmentation is critical in a wide range of fields, including hearing aids, ASR, and mobile communication. DNN-based architectures for speech recognition and augmentation have shown to be quite effective in recent years, according to a new study. …
Published in Journal of Telecommunication, Switching Systems and Networks · Vol. 9, Issue 2, 2022 · pp. 14–19 Read article
-
Speech Recognition and Analysis using Energy of the Signal
Abstract: AbstractSpeech plays an important role in day to day communication. It is a natural way to transfer thoughts from one to another. If any information needs to deliver in between human’s speech is the best way of delivery. With the help of the electronic system, we can extract the information from the original speech signal by doing signals processing. Here, in this research work, we formulate the system that recognizes …
Published in Current Trends in Signal Processing · Vol. 8, Issue 2, 2018 · pp. 25–33 Read article
-
Genesis of English Language with its Cybernetics Origin
Abstract: AbstractThis paper is looking into the genesis of the English language with its Cybernetics origins and hence looking into the possibilities of developing new techniques to be used in speech processing applications. As we all know, English is the third-most-spoken language in the world by a number of native speakers, which is used to create the present world, which is also a West Germanic language that was first spoken early …
Published in Current Trends in Signal Processing · Vol. 10, Issue 1, 2020 · pp. 19–28 Read article
-
An Efficient Recursive Least Square (ERLS) Algorithm for Spectral Estimation with the Aid of Wavelet and Artificial Intelligence
Abstract: The spectral estimation technique is used for time frequency signal analysis, speech processing, and other signal processing applications. Some drawbacks of RLS algorithm are that it requires high computational power and the output obtained is numerically instability. So, the spectral efficiency of the signal is affected and a power error occurs in the estimator. In this paper, an efficient recursive least square (ERLS) algorithm is proposed for improving the power …
Published in Current Trends in Signal Processing · Vol. 2, Issue 1-3, 2012 · pp. 1–10 Read article
-
Performance Analysis of Speech Performance Analysis of Speech Recognition using Advanced Algorithm
Abstract: In this paper, the authors have worked on a speech processing technique wherein they took audio speech as input signal from user and by recognition process, they perceived recorded word. Specifically, the authors concentrated on the amplitude and zero cross of speech signal, and two of the most important and useful terms, namely MFCC and Mahalanobis distance, for the purpose of getting an efficient outcome. The authors evaluated the proposed …
Published in Current Trends in Signal Processing · Vol. 3, Issue 1, 2013 · pp. 20–25 Read article
-
AI Voice Detection Tool
Abstract: In today’s digital era, distinguishing between AI-generated and human voices is more important than ever. This project introduces an AI-based voice detection system designed to accurately identify synthetic voices, ensuring security and authenticity across various applications like cybersecurity, media verification, and fraud prevention.Our system works by analyzing incoming audio samples and comparing them against a diverse database of both AI-generated and real human voices. Using advanced machine learning and signal …
Published in Journal of Instrumentation Technology & Innovations · Vol. 16, Issue 1, 2026 · pp. 1–8 Read article
-
Fuzzy Variable Frame Analysis for Speech Recognition
Abstract: AbstractRecent works in machine learning has focused on models such as support vector machine (SVM), artificial neural network (ANN) and long short-term memory (LSTM), for automatically controlling the generalization and parameterization of the optimization process. This paper presents a fuzzy interpretation frame analysis procedure using LSTM classifier for noisy speech at word level using thresholding and local maxima procedure at framing level for the recognition process. Front end MFCC procedure …
Published in Current Trends in Signal Processing · Vol. 9, Issue 3, 2019 · pp. 9–18 Read article
-
Evaluation of Speech Recognition System for Security Purposes using Comparative Correlation Method
Abstract: Humans rely heavily on language for communication and most often, speech (oral) is one of the commonest means of communication. The speech recognition system is a smart system which grants access to users by recognizing the speech of the authorized user. Speech recognition is smart and precise in terms of authentication and validation. The problem with most security systems is imbalance pitch estimation, random noise and increment in theft due …
Published in Current Trends in Signal Processing · Vol. 12, Issue 3, 2022 · pp. 1–9 Read article
-
ChatGPT Based Voice Assistant for Blind People
Abstract: The proposed system for converting speech input into text format to facilitate interaction with ChatGPT is a sophisticated integration of hardware and cloud-based services. Utilizing state-of-the-art technologies, it facilitates seamless communication between users and the AI model. At the outset, the microphone serves as the input device, capturing audio signals from the user's speech. These signals are then amplified to ensure clarity and fidelity before being transmitted to the ESP32 …
Published in Journal of Operating Systems Development & Trends · Vol. 11, Issue 2, 2024 · pp. 23–31 Read article
-
Comparative Analysis of MCNN and RCNN for Speech Emotion Recognition Using Gender Information
Abstract: Speech emotion recognition is a speech processing task and a computer-based approach designed to identify and classify the emotions conveyed in audio signals. The aim of this system is to evaluate a speaker's emotional state, such as happiness, anger, sadness, or frustration, by analyzing their speech patterns, which include prosodic features like pitch, frequency, and rhythm. Speech emotion recognition is used in various real-life scenarios that include Customer Service, Healthcare, …
Published in Journal of Communication Engineering & Systems · Vol. 15, Issue 1, 2025 · pp. 1–10 Read article
-
Perceptual Features Based Continuous Speech Recognition in Additive Noise Environment Using Various Modelling Techniques
Abstract: The main objective of this paper is to discuss the effectiveness of Mel frequency perceptual features and the noise reduction technique in evaluating the performance of multi speaker independent continuous speech recognition system in additive noise environment by using various modelling techniques. The proposed perceptual features are captured and trained using clustering technique, GMM, continuous density HMM and back propagation neural networks. Speech recognition system is evaluated on clean and …
Published in Current Trends in Signal Processing · Vol. 2, Issue 1-3, 2012 · pp. 67–81 Read article
-
A Review Paper on The Mathematical Foundations of Artificial Intelligence
Abstract: Artificial Intelligence (AI) is deeply rooted in various branches of mathematics, which provide the theoretical foundation and practical tools for developing intelligent systems. This paper explores the crucial role of mathematics in AI, focusing on key areas such as Linear Algebra, Probability and Statistics, Optimization Techniques, Calculus, Graph Theory, and Fourier and Wavelet Transforms. Linear Algebra is fundamental for representing and manipulating data, with applications in dimensionality reduction and neural …
Published in Research & Reviews: Discrete Mathematical Structures · Vol. 12, Issue 3, 2025 · pp. 7–14 Read article
-
Deep Learning Approach to Produce Artificial Speech (Text-To-Audio)
Abstract: This program utilizes key features of the .NET framework to facilitate smooth text-to-speech conversion and audio playback. Upon execution, users are prompted to input text via a graphical user interface (GUI), which the program converts into speech using the ‘SpeechSynthesizer’ class from the ‘System. Speech.Synthesis’ namespace. The audio that has been synthesized is handled and stored as a WAV file called ‘output.wav’ by utilizing the ‘FileStream’ class, allowing for future …
Published in Journal of Artificial Intelligence Research & Advances · Vol. 12, Issue 1, 2025 · pp. 28–33 Read article
-
Voice Recognized D.C. Motor Drive
Abstract: Speed controlling of a DC motor is a very important as DC motor is very useful in many industrial applications. The advancement in solid state devices has made DC motor speed regulation easier and smoother. The variation in speed can be easily obtained by the help of PWM technique where the gate pulse monitors the speed. In recent times voice automated systems are becoming very common in our day to …
Published in Journal of Electronic Design Technology · Vol. 14, Issue 1, 2023 · pp. 29–36 Read article
-
Bias Detection and Accuracy Enhancement in Voice-based Banking Authentication Using Deep Learning
Abstract: Biometric systems have become an integral part of how many people access banking services today, and voice verification systems can be a secure and easy-to-use source of banking authentication that does not require any physical contact with the bank or any other person. From the security perspective, these systems would normally provide an effective means of identifying an individual but frequently exhibit bias with respect to demographics such as the …
Published in International Journal of Information Security Engineering · Vol. 4, Issue 2, 2026 Read article
-
HMM-Based Text-to-Speech Synthesis & Stressed Speech Processing
Abstract: According to this paper, the new system produces synthetic speech that is noticeably higher-quality thanspeaker-dependent systems when actual speech data sets are used, and it can compete with speaker-dependent approacheseven in situations when substantial speech data sets are available. This excitation signal, the glottal source, has naturallypiqued the interest of speech synthesis, and a variety of techniques have been developed to mimic the glottal source ofspontaneous speech.The use of artificial …
Published in Recent Trends in Electronics Communication Systems · Vol. 10, Issue 3, 2023 · pp. 7–14 Read article
-
Mind-Machine Synergy: The Evolution and Future of Brain-Computer Interfaces
Abstract: Brain-Computer Interfaces (BCIs) represent a transformative technology that enables the direct communication between the human brain and external devices, bypassing the traditional output mechanisms, such as speech or physical movement. BCIs hold the potential to revolutionize fields, such as healthcare, neuroscience, and human-computer interaction by providing new ways to restore lost functions, enhance cognitive abilities, enable seamless communication, and create novel user experiences across various platforms and environments. This article …
Published in International Journal of Brain Sciences · Vol. 2, Issue 2, 2025 · pp. 19–31 Read article
-
Effective FPGA Implementation of Comb Filters to Improve Perception of Sensorineural Hearing Impaired
Abstract: AbstractIn this paper, effective method for implementation of comb filter on FPGA platform has been proposed. This work is intended for hearing aid to be used for people suffering from sensorineural hearing loss. The algorithmic implementation on FPGA is intricate during the design process. Here we propose effective and simplified approach for implementation of comb filter with 512 coefficients using Spartan-6 FPGA for dichotic presentation. The design of comb filter …
Published in Current Trends in Signal Processing · Vol. 7, Issue 3, 2017 · pp. 15–29 Read article
-
Speaker Segmentation Using Non Linear Energy Operator Based Variance Spectral Flux
Abstract: AbstractSpeaker segmentation is considered to be a process that attempts to find speaker segment boundaries in a given audio stream. It can be used in various applications of speaker diarization, speaker indexing, word count etc. Classification of speech and non-speech can be obtained by traditional method of variance spectral flux (VSF). In this paper, we investigate the new techniques to perform the speaker segmentation task for multiple speakers in one …
Published in Current Trends in Signal Processing · Vol. 7, Issue 3, 2017 · pp. 1–6 Read article