Search
91 articles for “multi-modal”
-
Multi-Modal Iris and Retina Recognition System using Discrete Hidden Markov Model based Score Fusion Technique
Abstract: The contribution of this work is to propose a model of multi-modal iris and retina recognition system where Hidden Markov Model based score fusion technique has been used for decision fusion approach. Two different uni-modal techniques such as iris recognition and retina recognition outputs are combined by using baseline reliability ratio based score fusion technique. CASIA iris dataset and DRIVE retinal dataset have been used to acquire the iris and …
Published in Current Trends in Information Technology · Vol. 5, Issue 2, 2015 · pp. 7–12 Read article
-
An Approach of Multi-modal Biometric Iris and Speech based Person Recognition System with Decision Fusion Technique
Abstract: This paper presents a unique approach of multi-modal iris and speech feature based person recognition system. Iris images and speech signals are taken from CASIA iris database and NOIZEUS speech database respectively. Iris features are extracted after applying iris images noise removing and image pre-processing techniques. On the other hand, speech signal noise removing, start-end points detection algorithm, silence parts removal, windowing and feature extraction techniques are applied to extract …
Published in Journal of Multimedia Technology & Recent Advancements · Vol. 2, Issue 2, 2015 · pp. 5–10 Read article
-
AI- based Prediction of Misinformation Virality Before Wide Dissemination using Attention-based Multi-modal
Abstract: Misinformation on social media has emerged as a critical global challenge, impacting public health, democratic institutions, and societal trust. While existing research has largely concentrated on detecting misinformation after it begins circulating, predicting its virality before wide dissemination remains an underexplored area, limited work addresses predicting its virality before wide dissemination. This paper presents a conceptual framework using attention-based multi-modal deep learning models to estimate the virality of misinformation posts …
Published in Journal of Mobile Computing, Communications & Mobile Networks · Vol. 12, Issue 3, 2025 Read article
-
Cyclist Safety Enhancement: A Multi-Modal Hazard Detection System
Abstract: This study presents a multi-modal hazard detection system to enhance cyclist safety in urban environments. Lever- aging a combination of computer vision, object tracking, and predictive modeling, the system offers a comprehensive approach to identifying and mitigating potential risks. Key contributions include improved depth estimation through object size priors, multi-class tracking utilizing KCF and Brisk, and a novel recurrent neural network architecture for predicting bicycle movement. The system’s collision detection …
Published in International Journal of Machine Systems and Manufacturing Technology · Vol. 1, Issue 2, 2023 · pp. 35–83 Read article
-
Design And Implementation Of A Multi-Modal Mobile Application Safety Analytics Utilizing Nlp
Abstract: This study suggests a multi-modal mobile app safety analytics platform that uses natural language processing (NLP) to handle voice, text, and SOS messages. For effective intent recognition and decision-making, the platform processes all messages in a standard text or SOS flag format. Tokenization and normalization are used to process text communications, whereas noise reduction and text conversion are used to handle voice messages. For a quicker response, the SOS messages …
Published in Journal of Mobile Computing, Communications & Mobile Networks · Vol. 13, Issue 2, 2026 Read article
-
VeriSci—AI-Based Multi-Modal Research Assistant
Abstract: The exponential growth of scientific literature is a major bottleneck for academic researchers who want to efficiently discover, assess, and synthesize relevant scholarly knowledge. Traditional methods of literature review are heavily reliant on manual keyword searching, human screening, and subjective data extraction, making them time-consuming, susceptible to cognitive bias, and less effective as the tidal wave of information continues to grow. To overcome these limitations, this study introduces VeriSci, an …
Published in Journal of Advanced Database Management & Systems · Vol. 13, Issue 2, 2026 · pp. 20–27 Read article
-
Multi-Modal (Hybrid) More Efficient and Highly Secured Enhanced Hand Geometry Authentication using 3 Steps Authentication by Means of Biometrics, Steganography and Encryption
Abstract: “Biometrics” is the science, which is used to verify the identity of the persons through either behavioral traits or physical characteristics. This area has gained great importance in maintaining the security of the places that require high security accuracy. Hand geometry is considered one of the biometrics, which derived its reputation from the ease of use and the acceptance of many people to use it. Hand geometry based biometric systems …
Published in Journal of Image Processing & Pattern Recognition Progress · Vol. 1, Issue 2, 2014 · pp. 22–33 Read article
-
Forecasting Climate-Driven Healthcare Demand in Agricultural Regions: A Multi-Modal AI Approach
Abstract: The rapidly increasing instability of world climatic regimes has made past meteorological thresholds irrelevant, especially in the agricultural areas where monetary stability and well-being of humans are closely intertwined with an environmental situation. The more the frequency of 1 in every 1000-year events, i.e., heatwaves and catastrophic flooding increase, the greater the rural healthcare systems are in crisis, i.e., unable to predict a surge in demand because of data scarcity, …
Published in International Journal of Climate Conditions · Vol. 2, Issue 2, 2025 · pp. 28–38 Read article
-
Fake News Detection Using Multimodal Framework
Abstract: AbstractThe explosive growth in fake news and its erosion to democracy, justice, and public trust has increased the demand for fake news analysis, detection and intervention. The explosive growth of fake news and its erosion to democracy, justice, and public trust increased the demand for fake news detection. As an interdisciplinary topic, the study of fake news encourages a concerted effort of experts in computer and information science, political science, …
Published in Journal of Multimedia Technology & Recent Advancements · Vol. 7, Issue 3, 2020 · pp. 14–20 Read article
-
Role of Machine Vision in Autonomous Vehicles: A Review
Abstract: The integration of machine vision in autonomous vehicles (AVs) is a critical advancement in the field of intelligent transportation systems. Machine vision systems enable AVs to perceive their environment, understand road conditions, detect obstacles, and make real-time decisions necessary for safe navigation. These systems rely heavily on image processing techniques, which have evolved significantly over the past decade, leading to improved performance in complex driving scenarios. These developments are largely …
Published in Trends in Machine design · Vol. 12, Issue 1, 2025 · pp. 38–43 Read article
-
Transfer Learning in Deep Learning Models for Medical Imaging: Utilizing Pretrained Models to Improve Performance in Medical Image Analysis
Abstract: Transfer learning is now a trending technique in deep learning, especially in medical imaging. This technique solves landmark problems by utilizing the pre-trained models, including the limited availability of the annotated medical data and the time-consuming computational costs of training deep learning models from scratch. The generalizability of deep models could increase diagnostic precision for specific medical tasks, require fewer samples to train, and take less time to train due …
Published in Journal of Image Processing & Pattern Recognition Progress · Vol. 12, Issue 1, 2025 · pp. 67–85 Read article
-
Sensor Technologies in Robotics: A Review of Vision, Tactile, and Proximity Sensing Systems
Abstract: Robotics has undergone remarkable advancements in recent decades, largely driven by the integration of cutting-edge sensor technologies. Sensors serve as crucial for allowing robots to precisely logic, interpret, and react to the world around them. Among the most essential sensor types used in robotics are vision sensors, tactile sensors, and proximity sensors. These technologies strengthen a robot’s capacity for successful navigation, for example, object manipulation, and contact with people and …
Published in International Journal of Robotics and Automation in Mechanics · Vol. 3, Issue 1, 2025 · pp. 31–37 Read article
-
AI Chatbot for Expressing Visual Content
Abstract: Recently, the artificial intelligence (AI) chatbot for expressing visual content has shown remarkable multi-modal capabilities. It can recognize funny features in photos and create webpages straight from handwritten text. These characteristics are uncommon in earlier vision language models. We think the use of a more sophisticated large language model (LLM) is the main factor behind vision verbalizer's superior multi-modal generating capabilities. We introduce vision verbalizer, which employs a single projection …
Published in Journal of Multimedia Technology & Recent Advancements · Vol. 11, Issue 2, 2024 · pp. 11–19 Read article
-
A Study on Site Selection Component of Model Logistics Park: A Case of Uttar Pradesh, India
Abstract: Uttar Pradesh, India’s 4th largest state and 3rd largest economy eyeing being a One trillion-dollar economy. Among the top 5 manufacturing states of India, home to the second-highest number of Micro, Small, and Medium Enterprises (organized and unorganized) in India. The National Highways Logistics Management Limited under the Ministry of Road Transport and Highways (MoRTH) and the National Highways Authority of India proposed Multi-Modal Logistics Parks (MMLPs) as an initiative …
Published in Trends in Transport Engineering and Applications Read article
-
Amalgamation of Medical Image using Wavelet Theory
Abstract: Image fusion has become a common term used within medical diagnostics and treatment. The term is used when multiple patient images are registered and overlaid or merged to provide additional information. Fused images may be created from multiple images from the same imaging modality, or by combining information from multiple modalities, such as magnetic resonance image (MRI), computed tomography (CT) etc. In radiology and radiation these images serve different purposes. …
Published in Research & Reviews: Discrete Mathematical Structures · Vol. 2, Issue 1, 2015 · pp. 9–14 Read article
-
Challenges in the Field of Soft Robotics.
Abstract: One of the interesting topics in robotics research of recent years has been Soft Robotics. Theuse of unconventional materials in robotic systems is an idea that is revolutionizing roboticsand automation. Robotic systems are being designed so that they can mimic biologicalsystems. The detailed study of interactions between a physical system and the environment iscrucial in developing this technology. Novel robotic components are being designed using softmaterials but the complexity of …
Published in Trends in Electrical Engineering · Vol. 10, Issue 3, 2020 · pp. 46–52 Read article
-
Integrating Digital Twins, Smart Materials, and Human Machine Collaboration for Sustainable Smart Manufacturing: Smart CNC & Industry 4.0 Applications
Abstract: The rapid evolution of Industry 4.0 and the emerging transition toward Industry 5.0 have been catalyzed by the convergence of intelligent digital technologies such as digital twins, cyber–physical systems (CPS), artificial intelligence (AI), the Internet of Things (IoT), and human-in-the-loop (HITL) frameworks. These technologies have transformed traditional manufacturing into adaptive, data-centric ecosystems capable of real-time optimization and predictive decision-making. In recent years, the fusion of computer numerical control (CNC) machines, …
Published in International Journal of Manufacturing and Production Engineering · Vol. 3, Issue 2, 2025 · pp. 1–8 Read article
-
Channel Estimated Modulation Techniques for Wireless Communication Systems: Review
Abstract: The high spectrum and energy capabilities of Multiple Input Multiple Output systems make them a propitious technology for 5th-generation wireless communication systems. The acquisition of channel information is crucial for utilizing the potential gains of the multi-modal systems. Numerous studies have established channel estimation techniques and are still in research which faced several challenges with downlink-uplink overheads, complexities, and pilot contamination. This work discusses the insights obtained from the reviewed …
Published in Trends in Opto-electro & Optical Communication · Vol. 15, Issue 2, 2025 · pp. 1–10 Read article
-
AirDraw: Gesture-Driven Sketching for Immersive Design Interfaces
Abstract: AirDraw represents a transformative advancement in artificial intelligence, creating fluid touch-free human-computer interfaces through real-time interpretation of hand gestures as precise design control signals. This paper introduces AirDraw—a revolutionary webcam-powered virtual canvas that redefines creative workflows and educational interactions. Particularly valuable for design students prototyping ideas, senior creators exploring digital art, and users with motor impairments who find traditional peripherals challenging, AirDraw delivers truly inclusive technology access. Unlike conventional gesture …
Published in Journal of Image Processing & Pattern Recognition Progress · Vol. 13, Issue 2, 2026 · pp. 32–41 Read article
-
Recent Advances in Content-based Image Retrieval: Techniques and Applications
Abstract: Content-based image retrieval (CBIR) plays a vital role in computer vision, driven by the increasing need for fast and accurate image retrieval across fields like healthcare, e-commerce, and digital libraries. This study offers a detailed review of CBIR methodologies, charting their progression from traditional feature extraction techniques, such as Local Binary Patterns (LBP), to contemporary deep learning-driven methods. The transformative impact of convolution neural networks (CNNs) is highlighted, emphasizing their …
Published in Journal of Computer Technology & Applications · Vol. 16, Issue 1, 2025 · pp. 67–71 Read article