Search
58 articles for “audio”
-
Deep Learning Approach to Produce Artificial Speech (Text-To-Audio)
Abstract: This program utilizes key features of the .NET framework to facilitate smooth text-to-speech conversion and audio playback. Upon execution, users are prompted to input text via a graphical user interface (GUI), which the program converts into speech using the ‘SpeechSynthesizer’ class from the ‘System. Speech.Synthesis’ namespace. The audio that has been synthesized is handled and stored as a WAV file called ‘output.wav’ by utilizing the ‘FileStream’ class, allowing for future …
Published in Journal of Artificial Intelligence Research & Advances · Vol. 12, Issue 1, 2025 · pp. 28–33 Read article
-
Audio Summarization of Podcasts
Abstract: Podcasts have emerged as a significant medium for disseminating information, sharing stories, and providing entertainment. As their popularity continues to soar, the sheer volume of available content poses a challenge for listeners seeking to efficiently consume information. In this context, creating and deploying an audio summarizer for podcasts becomes highly significant. This research paper delves into the motivation for creating such a tool, emphasizing the increasing need for concise and …
Published in Journal of Software Engineering Tools & Technology Trends · Vol. 11, Issue 3, 2024 · pp. 18–26 Read article
-
Reliable And Secure Audio Transmission in Underwater Communication Using Li-Fi
Abstract: Ensuring the security of audio transmission becomes crucial as underwater communication systems become more and more important in a variety of areas, including defense, marine research, and offshore organizations. For this reason, Li-Fi technology is an innovative solution that uses its immunity to electromagnetic interference to get around the restrictions of conventional radio frequencies. Li-Fi allows for high-speed data transfer while reducing interference and signal degradation by encoding audio onto …
Published in Trends in Opto-electro & Optical Communication · Vol. 15, Issue 1, 2025 · pp. 23–29 Read article
-
Real-Time Multilanguage Platform with Encryption
Abstract: The proposed system, titled “Real-Time Multilanguage Platform With Encryption”, is designed to enable fast, reliable, and secure communication across various languages without relying on any third-party APIs. It utilizes a built-in audio-to-text conversion mechanism that processes spoken input using Python-based libraries like Speech Recognition or OpenAI’s Whisper, ensuring high accuracy in transcriptions. Once the audio is converted to text, the platform employs an offline translation engine to convert the transcribed …
Published in Journal of Communication Engineering & Systems · Vol. 15, Issue 2, 2025 · pp. 1–7 Read article
-
Li-Fi based Underwater Audio and Data Transmission System
Abstract: Visible Light Communication, or Li-Fi (Light Fidelity), is a wireless optical networking technique used for communication (VLC). Due to its lack of RF communication's drawbacks, such as water absorption and scattering, Li-Fi has the potential to completely transform underwater communication. This project suggests utilizing two Arduinos, a solar panel, a speaker, a display, an amplifier, a laser, a battery, and other components to create a Li-Fi based underwater audio and …
Published in Recent Trends in Fluid Mechanics · Vol. 11, Issue 2, 2024 · pp. 23–29 Read article
-
Music Reactive led using Arduino
Abstract: In recent years, the integration of technology into interactive systems has garnered significant attention. Among the many applications, music-reactive LED systems have become popular, offering dynamic visualizations that respond to audio inputs. This paper explores the development and implementation of a music-reactive LED system using Arduino, focusing on real-time audio signal processing and LED control based on the frequency spectrum of the sound. By utilizing Fast Fourier Transform (FFT) algorithms …
Published in Journal of Microelectronics and Solid State Devices · Vol. 13, Issue 1, 2026 · pp. 27–33 Read article
-
Quick Reach Alert System
Abstract: This project presents the design and implementation of a compact and low-cost GPS tracking system using the A9G module. The system is capable of tracking real-time location, sending emergency alerts, and providing audio surveillance. It includes an SOS button that, when pressed, immediately sends the user’s current GPS location via SMS to a predefined phone number. Additionally, the system allows remote audio monitoring through a microphone connected to the A9G …
Published in Journal of Communication Engineering & Systems · Vol. 15, Issue 3, 2025 · pp. 34–40 Read article
-
Crypto Talk Voice Shield: Secure Speech Communication System Using Arduino
Abstract: In the rapidly evolving landscape of communication security, this study presents a system designed around Arduino Uno technology, specifically engineered for the secure encoding, transmission, and decoding of speech data. By integrating advanced encryption algorithms, the system ensures that speech data is transmitted in segmented bit chunks, each enveloped in multiple layers of security to prevent unauthorized access or interception. This multi-tiered encryption approach establishes a highly secure communication channel, …
Published in Research & Reviews: A Journal of Embedded System & Applications · Vol. 13, Issue 1, 2025 · pp. 23–30 Read article
-
Advancements in AI-Driven Sound Spectrogram Analysis: From Deep Learning to Quantum and Neuromorphic Processing
Abstract: The rapid advancement of artificial intelligence (AI) has significantly reshaped the field of audio signal processing, with sound spectrogram analysis emerging as a central research focus. Spectrograms provide a rich time–frequency representation of audio signals, making them particularly suitable for data-driven learning approaches. This paper presents an in-depth and original review of modern AI-based techniques applied to spectrogram analysis, highlighting their growing impact across critical application areas such as healthcare …
Published in Journal of Multimedia Technology & Recent Advancements · Vol. 13, Issue 1, 2026 · pp. 01–06 Read article
-
Voice Control Music System Using Raspberry Pi
Abstract: This paper presents a voice-controlled music system using Raspberry Pi, designed for hands-free music playback through voice commands. The system is implemented using Raspberry Pi 4B, a microphone, a speaker, and a 32GB SD card. The software is developed in Python using Thonny IDE, leveraging libraries such as SpeechRecognition, PyAudio, and gTTS for speech processing and audio playback. The system recognizes user commands like play, pause, next, and stop, converting …
Published in Journal of VLSI Design Tools and Technology · Vol. 15, Issue 2, 2025 · pp. 34–39 Read article
-
Classifying Abnormalities in Heartbeat Sound
Abstract: Heartbeat sounds play a major role in the detection of various diseases such as heart disease, hyperthyroidism, and high blood pressure in their early stages. In the proposed method, various abnormal and healthy heartbeat audio signals are given as input and the features are extracted using MFCC (mel-frequency cepstral coefficients). Then, a deep learning approach is applied in which the MFCC audio signals are sent to the CNN (convolutional neural …
Published in Research & Reviews: A Journal of Embedded System & Applications · Vol. 12, Issue 1, 2024 · pp. 24–31 Read article
-
Lip Reading: Transforming Speech to Text
Abstract: Lip reading, the ability to interpret spoken language by observing lip movements, is a valuable skill that can aid in various applications, particularly in enhancing speech recognition systems. This project explores the implementation of a deep learning-based lip-reading model to improve the accuracy and robustness of speech recognition in challenging environments, such as noisy or audio-limited settings. The proposed lip-reading system leverages Convolutional Neural Networks (CNNs) and Recurrent Neural Networks …
Published in Current Trends in Signal Processing · Vol. 14, Issue 1, 2024 · pp. 23–33 Read article
-
Acoustic Sensing for City Flow: Quasi-Supervised Recognition of Sirens and Traffic for Urban Mobility Intelligence
Abstract: This paper frames environmental audio as a mobility telemetry source, extending a benchmark urban-sound corpus with transportation-critical classes—ambulance, firetruck, police, and traffic—and training spectrogram-based models under a quasi-supervised regime to support real-time city operations; leveraging 10-fold protocols, class-weighted objectives, and audiospecific augmentations (time stretch, pitch shift, SpecAugment, PatchAugment), the system benchmarks multiple CNN backbones combined with self-supervised learning paradigms enable the extraction of rich, discriminative acoustic representations, achieving strong multi-class …
Published in Trends in Electrical Engineering · Vol. 16, Issue 1, 2025 · pp. 42–50 Read article
-
AI Voice Detection Tool
Abstract: In today’s digital era, distinguishing between AI-generated and human voices is more important than ever. This project introduces an AI-based voice detection system designed to accurately identify synthetic voices, ensuring security and authenticity across various applications like cybersecurity, media verification, and fraud prevention.Our system works by analyzing incoming audio samples and comparing them against a diverse database of both AI-generated and real human voices. Using advanced machine learning and signal …
Published in Journal of Instrumentation Technology & Innovations · Vol. 16, Issue 1, 2026 · pp. 1–8 Read article
-
IoT-Based Smart Stick for the Visually Impaired: Real-Time Obstacle Detection and Auditory Feedback System
Abstract: The smart stick is designed to assist visually impaired individuals by helping them detect obstacles in their surroundings without relying solely on physical contact. The device is equipped with sensors that continuously scan the path ahead of the user. When an obstacle is detected, the sensors convert the information into real-time audio signals that are delivered through headphones. These audio alerts allow users to understand their environment more clearly and …
Published in Journal of Communication Engineering & Systems · Vol. 16, Issue 1, 2026 · pp. 29–35 Read article
-
The Prospects of Multimedia in 2025: Emerging Patterns to Monitor
Abstract: The term multimedia describes the computer-aided combination of text, illustrations, graphics, audio, animation, still and moving images (videos), and any other medium that allows for the digital expression, storing, processing, and communication of any kind of information. The word of the decade is multimedia. It is a buzzword that has been used in a variety of contexts. It appears on the covers of movies, CDs, literature, newspapers, and handheld games. …
Published in Journal of Multimedia Technology & Recent Advancements · Vol. 11, Issue 3, 2024 · pp. 25–31 Read article
-
Stealth Listening Device and Global Positioning System Tracker
Abstract: The Audio Spy and GPS tracker device functions as a covert listening tool, enabling users to secretly overhear conversations through a cellular network. Additionally, it serves as an SOS (Save Our Souls) button, allowing users to send their current location via SMS with a simple press of a button. This feature also initiates a call to a designated SOS number, while the device can be tracked by sending a single …
Published in Research & Reviews: A Journal of Embedded System & Applications · Vol. 12, Issue 3, 2024 · pp. 1–6 Read article
-
Enhancing Analog Radio Broadcasting Using DDRB Algorithm
Abstract: With the rapid advancement of digital technology and the global transition toward modern radio broadcasting, Nepal has a significant opportunity to upgrade its radio broadcasting system. Digital radio services offer numerous benefits, including superior audio quality, broader coverage, interactive features, and enhanced capabilities. This thesis focuses on improving Nepal's analog radio broadcasting through the implementation of DDRB (Digital Radio Mondiale [DRM] Digital Radio Broadcasting) technology. The study begins by analyzing …
Published in Journal of Telecommunication, Switching Systems and Networks · Vol. 11, Issue 3, 2024 · pp. 34–43 Read article
-
Garbage Classifier Using Arduino
Abstract: In today's era, where prioritizing sustainable methods of waste disposal is crucial, a pioneering initiative titled "Garbage Collection using Arduino Nano " stands out as a revolutionary approach to refining the traditional, labor-intensive methods of household waste collection. Utilizing the Arduino Nano microcontroller, this project introduces a sophisticated waste management system equipped with diverse sensors and mechanisms. Through the automation of waste collection, this avant-garde system not only enhances efficiency …
Published in Journal of Mechatronics and Automation · Vol. 11, Issue 1, 2024 · pp. 18–23 Read article
-
Hearing, Speech, and Language Characteristics in a Case with Hemoglobinopathy Secondary to Beta Thalassemia Intermedia
Abstract: Background: Thalassemia is an inherited disorder characterized by a reduced amount or absence of hemoglobin, the oxygen-carrying protein inside the red blood cell. Thalassemia has its types called alpha and beta-thalassemia. Beta thalassemia is a condition in which there is a reduction or deficit in the synthesis of the beta-globin chain of hemoglobin molecules caused by a mutation in chromosome eleven. Beta thalassemia can be broadly categorized into three main …
Published in Research and Reviews : A Journal of Medical Science and Technology · Vol. 14, Issue 1, 2024 · pp. 1–5 Read article