Search
30 articles for “Speech Feature Extraction”
-
Speaker Identification using Hopfield Neural Network based Classifier
Abstract: The aim of this work is to enhance the performance of speaker identification using Hopfield neural network algorithm. Speech signals are collected from VALID Audio-Visual dataset and some speech signal pre-processing techniques are applied to process the speech to feed the Hopfield neural network algorithm. Filtering technique is applied to remove the noises from the speech signals. MFCC based standard speech feature extraction technique is used to extract the speech …
Published in Journal Of Network security · Vol. 3, Issue 2, 2015 · pp. 32–35 Read article
-
Extraction of speech Emotion Features Using MLP Classifier
Abstract: Speech Emotion Recognition is a thriving research topic. Speech emotion recognition use MLP classifier to categorize the emotions from the speech. This Speech is also used as the medium, through which one can express their feelings and mind state in human to machine interaction. The case is very easy where two humans communicate along with their emotions as by nature, they can recognize each other’s emotions. But for computer, if …
Published in Journal of Instrumentation Technology & Innovations · Vol. 12, Issue 3, 2022 · pp. 11–19 Read article
-
Comparative Analysis Between Librosa and OpenSMILE
Abstract: This research work focuses on comparative study of Librosa, a python-based library, and openSMILE, a C++ toolkit, with python bindings used in audio speech analysis. Librosa is ideal for beginners due to its simple structure and flexibility with strong integration with machine learning frameworks like TensorFlow and PyTorch. On the other hand, OpenSMILE is ideal for speech-centric tasks like speech-emotion recognition or paralinguistic studies, offering a wide range of pre-defined …
Published in Journal of Open Source Developments · Vol. 12, Issue 3, 2025 · pp. 06–10 Read article
-
An Overview and Comparative Study of Speech Recognition Techniques
Abstract: The easy way of communication between machine and human being is automatic speech recognition. Automatic speech recognition in different languages is one major research areas in signal processing field. Accuracy in SR is an important challenge in the research work. This paper discusses researches in the SR field in the past 60 years and gives an overview on SR process, types of speech, approaches, classifiers such as dynamic time warping …
Published in Research & Reviews: A Journal of Embedded System & Applications · Vol. 4, Issue 1, 2016 · pp. 7–12 Read article
-
An Approach of Multi-modal Biometric Iris and Speech based Person Recognition System with Decision Fusion Technique
Abstract: This paper presents a unique approach of multi-modal iris and speech feature based person recognition system. Iris images and speech signals are taken from CASIA iris database and NOIZEUS speech database respectively. Iris features are extracted after applying iris images noise removing and image pre-processing techniques. On the other hand, speech signal noise removing, start-end points detection algorithm, silence parts removal, windowing and feature extraction techniques are applied to extract …
Published in Journal of Multimedia Technology & Recent Advancements · Vol. 2, Issue 2, 2015 · pp. 5–10 Read article
-
Identification of English Dialects and Emotions using Spectral and Prosodic Features of Speech Signal Processing
Abstract: AbstractIn this paper, the authors have explored speech features to identify English dialects and emotions. A dialect is any distinguishable variety of a language spoken by a group of people. Emotions provide naturalness to speech. Speech database considered for dialect identification task consists of spontaneous speech spoken by male and female speakers. The emotions considered in this study are anger, disgust, fear, happy, neutral and sad. Prosodic and spectral features …
Published in Journal of Computer Technology & Applications · Vol. 4, Issue 2, 2013 · pp. 10–17 Read article
-
Development of Polymer-Based Sensors for Speech Emotion Recognition
Abstract: Traditional SER research often utilizes microphones with polymer components like Diaphragms and Membranes. Within some microphone designs, polymer membranes which plays a crucial role in converting sound pressure into electrical signals. The paper highlights the application (speech emotion recognition) and have tried to find polymer-based sensors. This work further delves deeper, investigating the performance of the CatBoost algorithm for emotion recognition in voice assistants designed for Indian languages. The research …
Published in Journal of Polymer & Composites · Vol. 12, Issue 5, 2024 · pp. 268–274 Read article
-
Bandwidth Extension of Speech Signal Artificially & Transmission & Coding for Wireless Communication
Abstract: AbstractIn today’s next generation wireless communication, decoded speech quality at recipient side initiate quiet and slim; mainly recognized to intrinsic band limitation (300-3400 Hz). so as to increase clearness and genuineness of improved speech signal. Narrowband speech coders be supposed to promote to wideband coders supports bandwidth of 50-7000Hz. extensive time duration has been left for advancement from narrowband to fully wideband well-suited systems. Many smart devices at the present …
Published in Journal of Communication Engineering & Systems · Vol. 9, Issue 2, 2019 · pp. 111–116 Read article
-
Automatic annotation of reading using Speech Recognition: A pilot study
Abstract: Speech is a time-varying continuous stimulus. Analyzing a speech sample is generally done by transcribing it which can be considered to be a time-consuming process. A time annotated transcription is required for a detailed evaluation of the speech sample to study the language and prosodic features. The focus of this paper is to develop a tool to automatically segment audio recordings into silences and chunks of speech corpus and further …
Published in Research & Reviews: A Journal of Bioinformatics · Vol. 5, Issue 2, 2018 · pp. 25–29 Read article
-
Lip Reading: Transforming Speech to Text
Abstract: Lip reading, the ability to interpret spoken language by observing lip movements, is a valuable skill that can aid in various applications, particularly in enhancing speech recognition systems. This project explores the implementation of a deep learning-based lip-reading model to improve the accuracy and robustness of speech recognition in challenging environments, such as noisy or audio-limited settings. The proposed lip-reading system leverages Convolutional Neural Networks (CNNs) and Recurrent Neural Networks …
Published in Current Trends in Signal Processing · Vol. 14, Issue 1, 2024 · pp. 23–33 Read article
-
Design and development of an IP core for speech detection in varying noise environments
Abstract: Speech detection is a process of separating speech and silence regions from noisy speech. It is an important part in speech processing. One of the important and difficult tasks in speech detection is the detection of speech in a very low signal to noise ratio. The application of speech detector includes speech coding, where the accurate speech coding can reduce bit number low down the bit rates, and in speech …
Published in Journal of Multimedia Technology & Recent Advancements · Vol. 7, Issue 1, 2020 · pp. 19–27 Read article
-
Exploiting Phase of Speech Signal for Speaker Recognition
Abstract: Abstract—Performance of speaker recognition system with feature based on temporal phase is presented in this paper. In the state of the art spectral feature, Mel Frequency Cepstal Coefficient (MFCC) only magnitude of the Fourier Transform of speech signal is considered while phase is ignored. The cepstral coefficients extracted from temporal phase (MFTPC) of speech signal are used as features for speaker recognition system. The performance of MFTPC feature is evaluated …
Published in Journal of Communication Engineering & Systems · Vol. 9, Issue 2, 2019 · pp. 81–87 Read article
-
AI Voice Detection Tool
Abstract: In today’s digital era, distinguishing between AI-generated and human voices is more important than ever. This project introduces an AI-based voice detection system designed to accurately identify synthetic voices, ensuring security and authenticity across various applications like cybersecurity, media verification, and fraud prevention.Our system works by analyzing incoming audio samples and comparing them against a diverse database of both AI-generated and real human voices. Using advanced machine learning and signal …
Published in Journal of Instrumentation Technology & Innovations · Vol. 16, Issue 1, 2026 · pp. 1–8 Read article
-
Feature Selection and Weighted Based Optimized Weight based Multi-Tier Stacked Ensemble (WMTSE) Classification for Twitter Sentiment Analysis
Abstract: AbstractThis research work concentrates on both feature selection and classification methods for utilizing twitter data. A new classifier is introduced for classifying “tweets” into positive, negative and neutral sentiment. The system contains four steps: Preprocessing by Tokenization, Text Cleaning, Part of Speech (PoS) Tagging, Stemming and Stop Words Removal, Feature Extraction by Bag-of-words (BoW), Lexicon-based features and Term Frequency- Inverse Document Frequency(TF-IDF), Feature Selection by Binary Swallow Swarm Optimization (BSSO) …
Published in Journal Of Network security · Vol. 8, Issue 3, 2020 · pp. 20–23 Read article
-
Audio-Only Speaker Identification using Principal Component Analysis based Back-Propagation Learning Neural Network in Noisy Environment
Abstract: This paper introduces text dependent speaker identification system on Principal Component Analysis based Back-Propagation learning neural network which deals with detecting a particular speaker from a known populations under noisy environment. For audio pre-processing, ends point detection, silence parts removal, frame segmentation and windowing techniques have been used and wiener filter has been applied to remove the background noise from the audio speech utterances. To reduce the dimension of the …
Published in Current Trends in Signal Processing · Vol. 3, Issue 3, 2013 · pp. 1–10 Read article
-
Emotion Recognition based on Human voice tones
Abstract: In this period of technology and automation, human interaction with machines has become an unavoidable occurrence. Machines are making our lives much easier, stretching their hands wherever it is possible for smoother operation of various tasks. Emotion recognition system is one of the system that can help humans for a much easier and meaningful interaction with machines. It can be used in various branches such as artificial intelligence, health care, …
Published in Current Trends in Information Technology · Vol. 10, Issue 2, 2020 · pp. 1–5 Read article
-
A Review Paper on The Mathematical Foundations of Artificial Intelligence
Abstract: Artificial Intelligence (AI) is deeply rooted in various branches of mathematics, which provide the theoretical foundation and practical tools for developing intelligent systems. This paper explores the crucial role of mathematics in AI, focusing on key areas such as Linear Algebra, Probability and Statistics, Optimization Techniques, Calculus, Graph Theory, and Fourier and Wavelet Transforms. Linear Algebra is fundamental for representing and manipulating data, with applications in dimensionality reduction and neural …
Published in Research & Reviews: Discrete Mathematical Structures · Vol. 12, Issue 3, 2025 · pp. 7–14 Read article
-
Bias Detection and Accuracy Enhancement in Voice-based Banking Authentication Using Deep Learning
Abstract: Biometric systems have become an integral part of how many people access banking services today, and voice verification systems can be a secure and easy-to-use source of banking authentication that does not require any physical contact with the bank or any other person. From the security perspective, these systems would normally provide an effective means of identifying an individual but frequently exhibit bias with respect to demographics such as the …
Published in International Journal of Information Security Engineering · Vol. 4, Issue 2, 2026 Read article
-
Oculight: AI Based System for Visual Assistance to Blind and Visually Impaired People
Abstract: This research introduces a new captioning model that utilizes both image and caption models to produce textual descriptions of images. By combining a convolutional neural network (CNN) and a long short-term memory (LSTM) recurrent neural network (RNN), this deep learning architecture offers a promising solution for improving the quality of life and independence of blind people through AIbased systems. The system's implementation on an Android device ensures that it is …
Published in Journal of Artificial Intelligence Research & Advances · Vol. 10, Issue 1, 2023 · pp. 22–29 Read article
-
Summarizer: A Single Document text summarization using Hybrid Approach
Abstract: Summarization task is a process of diminishing the size of the input text, without altering its overall meaning, by selecting only the significant parts of the input text as an output. Summary of a text can be obtained either by an abstractive approach or an extractive approach. Our system is a hybrid system (that is the combination of abstractive and extractive approach). The first stage of this system is to …
Published in Journal of Advancements in Robotics · Vol. 5, Issue 2, 2018 · pp. 15–25 Read article