Search
8 articles for “text-to-speech synthesizer”
-
HMM-Based Text-to-Speech Synthesis & Stressed Speech Processing
Abstract: According to this paper, the new system produces synthetic speech that is noticeably higher-quality thanspeaker-dependent systems when actual speech data sets are used, and it can compete with speaker-dependent approacheseven in situations when substantial speech data sets are available. This excitation signal, the glottal source, has naturallypiqued the interest of speech synthesis, and a variety of techniques have been developed to mimic the glottal source ofspontaneous speech.The use of artificial …
Published in Recent Trends in Electronics Communication Systems · Vol. 10, Issue 3, 2023 · pp. 7–14 Read article
-
Optical/Handwritten Data Recognition
Abstract: For many years, optical character recognition (OCR) has been a hot topic. It's the process of breaking down a document image into its individual characters. Despite decades of intensive research, producing OCR with human-like skills is still a work in progress. The industries have long used localization and recognition of written characters for a number of applications. When working in a sterile environment, such as scanners or static settings, most …
Published in Trends in Opto-electro & Optical Communication · Vol. 11, Issue 2, 2021 · pp. 22–31 Read article
-
Real-time Voice-over Translation
Abstract: The real-time voice-over translation project stands as a pioneering endeavor poised to reshape the landscape of cross-cultural communication. In response to the escalating demand for fluid interaction amidst linguistic diversity, this initiative sets out to engineer a transformative solution. Its core objective lies in the seamless mediation of language barriers during real-time exchanges. Through the fusion of cutting-edge speech recognition and machine translation technologies, the project endeavors to fabricate a …
Published in International Journal of Computer Science Languages · Vol. 2, Issue 1, 2024 · pp. 24–32 Read article
-
ChatGPT Based Voice Assistant for Blind People
Abstract: The proposed system for converting speech input into text format to facilitate interaction with ChatGPT is a sophisticated integration of hardware and cloud-based services. Utilizing state-of-the-art technologies, it facilitates seamless communication between users and the AI model. At the outset, the microphone serves as the input device, capturing audio signals from the user's speech. These signals are then amplified to ensure clarity and fidelity before being transmitted to the ESP32 …
Published in Journal of Operating Systems Development & Trends · Vol. 11, Issue 2, 2024 · pp. 23–31 Read article
-
Empowering Accessibility: A Review of Text-to-Speech Systems for the Visually Impaired with Raspberry Pi
Abstract: The automatic text reader for people who are blind is presented in the paper. built on a Raspberry Pi. It makes use of computer programming and image sensing tools to identify printed characters employing optical character recognition technology. It creates machine-encoded text from an image using text that has been printed, typed, or written by hand. In the present work, text-to-speech synthesis and OCR are used to convert the image …
Published in Trends in Opto-electro & Optical Communication · Vol. 13, Issue 1, 2023 · pp. 16–21 Read article
-
Text-To-Speech (TTS) Conversion System for Gujarati Language
Abstract: Text-to-speech (TTS) conversion system enables user to enter text in Gujarati language and as an output it generates the equivalent sound. This type of system will be greatly useful for illiterate and vision-impaired people to hear and understand the content. TTS systems are still suffering from the problem of producing emotional speech like that of human being. Scientists are trying to give emotions and feelings to it. This shows that …
Published in Journal of Electronic Design Technology · Vol. 4, Issue 3, 2013 · pp. 19–23 Read article
-
Deep Learning Approach to Produce Artificial Speech (Text-To-Audio)
Abstract: This program utilizes key features of the .NET framework to facilitate smooth text-to-speech conversion and audio playback. Upon execution, users are prompted to input text via a graphical user interface (GUI), which the program converts into speech using the ‘SpeechSynthesizer’ class from the ‘System. Speech.Synthesis’ namespace. The audio that has been synthesized is handled and stored as a WAV file called ‘output.wav’ by utilizing the ‘FileStream’ class, allowing for future …
Published in Journal of Artificial Intelligence Research & Advances · Vol. 12, Issue 1, 2025 · pp. 28–33 Read article
-
Speech-Text-Speech Translator: A Generative AI Framework for Real-Time, Identity-Preserving S2S Translation
Abstract: Different languages have been proved a great obstacle to global communication despite the internet's role in allowing information sharing all over the world. While presenting an extensive number of current solutions, traditional Machine Translation (MT) systems are unable to convey complex contextual information and dialects including "Hinglish". Above all, the voice of the interlocutor is lost and is replaced with an artificial one, programmed to mimic the voice of the …
Published in Journal of Mechatronics and Automation · Vol. 13, Issue 2, 2026 · pp. 47–57 Read article