Search
18 articles for “speaker recognition”
-
Exploiting Phase of Speech Signal for Speaker Recognition
Abstract: Abstract—Performance of speaker recognition system with feature based on temporal phase is presented in this paper. In the state of the art spectral feature, Mel Frequency Cepstal Coefficient (MFCC) only magnitude of the Fourier Transform of speech signal is considered while phase is ignored. The cepstral coefficients extracted from temporal phase (MFTPC) of speech signal are used as features for speaker recognition system. The performance of MFTPC feature is evaluated …
Published in Journal of Communication Engineering & Systems · Vol. 9, Issue 2, 2019 · pp. 81–87 Read article
-
Speaker Recognition System Through Matlab
Abstract: This paper presents the work on the Speaker Recognition System through the MATLAB software. Speaker recognition means recognizing the person who is actually speaking rather than knowing what has been said. This system is widely used in banking, security systems, voice mail, etc. This paper presents three stages of the system, namely, the voice acquisition stage in which data (speech signal) is acquired from the speaker and is stored for …
Published in Current Trends in Signal Processing · Vol. 3, Issue 2, 2013 · pp. 26–29 Read article
-
Speaker Variations and Vocal Disguise
Abstract: The process of recognizing the speaker based on parameters like pitch, loudness and other acoustic attributes is called speaker recognition. Speaker recognition is considered to be challenging. The voice changing apps facilitating vocal disguise have further made this process even difficult. The voice changing apps can induce variations, some of these variations may be predominantly different, even though these apps can disguise a person’s voice, certain parameters may stay real …
Published in Research and Reviews: A Journal of Health Professions · Vol. 13, Issue 2, 2023 · pp. 15–19 Read article
-
Perceptual Features Based Continuous Speech Recognition in Additive Noise Environment Using Various Modelling Techniques
Abstract: The main objective of this paper is to discuss the effectiveness of Mel frequency perceptual features and the noise reduction technique in evaluating the performance of multi speaker independent continuous speech recognition system in additive noise environment by using various modelling techniques. The proposed perceptual features are captured and trained using clustering technique, GMM, continuous density HMM and back propagation neural networks. Speech recognition system is evaluated on clean and …
Published in Current Trends in Signal Processing · Vol. 2, Issue 1-3, 2012 · pp. 67–81 Read article
-
Lip Reading: Transforming Speech to Text
Abstract: Lip reading, the ability to interpret spoken language by observing lip movements, is a valuable skill that can aid in various applications, particularly in enhancing speech recognition systems. This project explores the implementation of a deep learning-based lip-reading model to improve the accuracy and robustness of speech recognition in challenging environments, such as noisy or audio-limited settings. The proposed lip-reading system leverages Convolutional Neural Networks (CNNs) and Recurrent Neural Networks …
Published in Current Trends in Signal Processing · Vol. 14, Issue 1, 2024 · pp. 23–33 Read article
-
Speech Recognition and Analysis using Energy of the Signal
Abstract: AbstractSpeech plays an important role in day to day communication. It is a natural way to transfer thoughts from one to another. If any information needs to deliver in between human’s speech is the best way of delivery. With the help of the electronic system, we can extract the information from the original speech signal by doing signals processing. Here, in this research work, we formulate the system that recognizes …
Published in Current Trends in Signal Processing · Vol. 8, Issue 2, 2018 · pp. 25–33 Read article
-
Interactive Smart Mirror for Personal Information Management
Abstract: AbstractThe future of Personal Information Management depends on the Internet of things (IoT). Though the applications of IoT are diverse, the one that concerns the common man is how it can be used to make life easier and faster. This is where Personal Information Management using IoT comes into the picture. Interactive Smart Mirror represents the design and development of a futuristic system for the personal information management for commercial …
Published in Journal of Electronic Design Technology · Vol. 10, Issue 2, 2019 · pp. 16–21 Read article
-
Comparative Analysis of MCNN and RCNN for Speech Emotion Recognition Using Gender Information
Abstract: Speech emotion recognition is a speech processing task and a computer-based approach designed to identify and classify the emotions conveyed in audio signals. The aim of this system is to evaluate a speaker's emotional state, such as happiness, anger, sadness, or frustration, by analyzing their speech patterns, which include prosodic features like pitch, frequency, and rhythm. Speech emotion recognition is used in various real-life scenarios that include Customer Service, Healthcare, …
Published in Journal of Communication Engineering & Systems · Vol. 15, Issue 1, 2025 · pp. 1–10 Read article
-
ChatGPT Based Voice Assistant for Blind People
Abstract: The proposed system for converting speech input into text format to facilitate interaction with ChatGPT is a sophisticated integration of hardware and cloud-based services. Utilizing state-of-the-art technologies, it facilitates seamless communication between users and the AI model. At the outset, the microphone serves as the input device, capturing audio signals from the user's speech. These signals are then amplified to ensure clarity and fidelity before being transmitted to the ESP32 …
Published in Journal of Operating Systems Development & Trends · Vol. 11, Issue 2, 2024 · pp. 23–31 Read article
-
Native Language Identification from Spoken Indian English
Abstract: Automatic speech recognition (ASR) systems that facilitate voice based search and information retrieval have gained importance recently. While the performance of ASR systems for Indian languages have improved in the recent past. They have yet to gain wide acceptability as much as the ASR systems for English spoken by Indians. Almost all Indians learn English as a second or third language. So, the phoneme set and the prosody of native …
Published in Trends in Electrical Engineering · Vol. 9, Issue 2, 2019 · pp. 1–6 Read article
-
Control the Robot via Both Voice & Moblie Application in Departmental Information Robot
Abstract: The concept and development of an autonomous, interactive humanoid robot created to deliver college information via voice and mobile application control are presented in this work. To facilitate easy user interaction with the robot, the suggested system combines intelligent control, wireless connection, and speech recognition. The robot is equipped with a 10-inch monitor that provides instructors and students with information about college departments, facilities, and services. Oral responses from a …
Published in Recent Trends in Electronics Communication Systems · Vol. 13, Issue 2, 2026 Read article
-
Keyless: LQ Based On IOT
Abstract: In recent years, there has been a growing interest in smart home technologies, with smart door locks being one of the focal points due to their potential to enhance security and convenience. Smart door lock systems are transforming access control by offering a keyless alternative for both residential and commercial settings. These electronic locks replace traditional keys with secure methods such as emergency alarms, privacy modes, battery backup, cameras, voice …
Published in Journal Of Network security Read article
-
Voice Control Music System Using Raspberry Pi
Abstract: This paper presents a voice-controlled music system using Raspberry Pi, designed for hands-free music playback through voice commands. The system is implemented using Raspberry Pi 4B, a microphone, a speaker, and a 32GB SD card. The software is developed in Python using Thonny IDE, leveraging libraries such as SpeechRecognition, PyAudio, and gTTS for speech processing and audio playback. The system recognizes user commands like play, pause, next, and stop, converting …
Published in Journal of VLSI Design Tools and Technology · Vol. 15, Issue 2, 2025 · pp. 34–39 Read article
-
Keyless Lock Based on Internet of Things
Abstract: In recent years, there has been a growing interest in smart home technologies, with smart door locks being one of the focal points due to their potential to enhance security and convenience. Smart door lock systems are transforming access control by offering a keyless alternative for both residential and commercial settings. These electronic locks replace traditional keys with secure methods such as emergency alarms, privacy modes, battery backup, cameras, voice …
Published in Journal Of Network security · Vol. 12, Issue 2, 2024 · pp. 18–21 Read article
-
Speech-Text-Speech Translator: A Generative AI Framework for Real-Time, Identity-Preserving S2S Translation
Abstract: Different languages have been proved a great obstacle to global communication despite the internet's role in allowing information sharing all over the world. While presenting an extensive number of current solutions, traditional Machine Translation (MT) systems are unable to convey complex contextual information and dialects including "Hinglish". Above all, the voice of the interlocutor is lost and is replaced with an artificial one, programmed to mimic the voice of the …
Published in Journal of Mechatronics and Automation · Vol. 13, Issue 2, 2026 · pp. 47–57 Read article
-
Development of Polymer-Based Sensors for Speech Emotion Recognition
Abstract: Traditional SER research often utilizes microphones with polymer components like Diaphragms and Membranes. Within some microphone designs, polymer membranes which plays a crucial role in converting sound pressure into electrical signals. The paper highlights the application (speech emotion recognition) and have tried to find polymer-based sensors. This work further delves deeper, investigating the performance of the CatBoost algorithm for emotion recognition in voice assistants designed for Indian languages. The research …
Published in Journal of Polymer & Composites · Vol. 12, Issue 5, 2024 · pp. 268–274 Read article
-
Identification of English Dialects and Emotions using Spectral and Prosodic Features of Speech Signal Processing
Abstract: AbstractIn this paper, the authors have explored speech features to identify English dialects and emotions. A dialect is any distinguishable variety of a language spoken by a group of people. Emotions provide naturalness to speech. Speech database considered for dialect identification task consists of spontaneous speech spoken by male and female speakers. The emotions considered in this study are anger, disgust, fear, happy, neutral and sad. Prosodic and spectral features …
Published in Journal of Computer Technology & Applications · Vol. 4, Issue 2, 2013 · pp. 10–17 Read article
-
Genesis of English Language with its Cybernetics Origin
Abstract: AbstractThis paper is looking into the genesis of the English language with its Cybernetics origins and hence looking into the possibilities of developing new techniques to be used in speech processing applications. As we all know, English is the third-most-spoken language in the world by a number of native speakers, which is used to create the present world, which is also a West Germanic language that was first spoken early …
Published in Current Trends in Signal Processing · Vol. 10, Issue 1, 2020 · pp. 19–28 Read article