|
Proceedings of 13th International Conference on Microelectronics, Circuits and Systems(Micro2026).ISBN: 978-81-985770-0-9 Editors: Prof. (Dr.) Abhijit Biswas, Department of Radio Physics and Electronics, University of Calcutta, Kolkata, West Bengal, India. Prof. (Dr.) Pankaj Gupta, Department of ECE, Indira Gandhi Delhi Technical University for Women, Kashmere Gate, New Delhi, India. Dr. Priyanka Goyal, Department of ECE, Gautam Buddha University, Greater Noida, Uttar Pradesh, India. Publishing Date: December 2026 Indexed by: ACT, |
|
List of Papers:
Editorial: Editorial of this Book --------------------------------------------------------------------------------------------- AOI :10.100.234513.0201
ABSTRACT:
Epilepsy is a prevalent neurological disorder affecting millions worldwide, in which people experience frequent seizures due to abnormal electrical activity in the brain. Early detection of structural brain changes associated with seizure conditions can enhance diagnosis and treatment planning. While Electroencephalography (EEG) is widely used for real-time seizure detection, Magnetic Resonance Imaging (MRI) provides structural information that can reveal abnormalities linked to seizure disorders. This paper proposes a novel framework for analyzing seizure-associated brain structural patterns in MRI images, combining Convolutional Neural Networks (CNN) with Horizontal Visibility Graph (HVG) construction and attention-based Graph Transformer classification. The preprocessing pipeline includes grayscale conversion, noise removal, and intensity normalization. Region-of-interest (ROI) signals are derived through spatial averaging of pixel intensities. These signals are converted to HVGs, where each data point becomes a node and edges are formed using horizontal visibility criteria. Graph-theoretic features—degree, clustering coefficient, and average shortest path length—are extracted to form a spatial structural feature set. An attention-based Graph Transformer classifier then performs binary classification into abnormal (seizure) and normal cases. The core innovation lies in a hybrid CNN + HVG + Graph Transformer architecture that jointly models spatial and structural properties of MRI data. It is important to note that the dataset uses tumor-affected MRI images as proxies for seizure-associated structural changes; while this is a recognized limitation, it allows exploration of the clinical hypothesis that tumor-induced structural alterations share characteristics with seizure-related atrophy. Experiments demonstrate an overall classification accuracy of 96.08%, a seizure class (Class 1) recall of 1.00, precision of 0.94, and F1-score of 0.97, outperforming traditional machine learning, CNN-only, RNN-based, and standard Graph Neural Network baselines. These results indicate that graph-based structural representation of MRI-derived signals can support neurological diagnosis with strong computational accuracy.
AOI :10.100.234513.0202
ABSTRACT:
Replicated key-value stores sit at the core of many distributed applications, where they are expected to serve requests quickly, keep data safe, and keep working when individual machines fail. Building one is largely a question of how much consistency to trade away for availability and speed, and the answer gets harder once crashes and reconfiguration enter the picture. In this paper we describe a key-value store built around a modified chain replication protocol. It keeps data strongly consistent and durable, and stays available as long as a majority of the replicas are up. We cover the client-server protocol, how the system reacts to failures, and how state is stored on disk, and we measure the result under several workloads. Across different key and value sizes, client counts, and read-write mixes the system behaves predictably, and it stays correct when nodes go down.
AOI :10.100.234513.0203
ABSTRACT:
Phishing websites are taking the form of multi-modal threats which integrate bad URLs, spam email messages, altered images, and hacked attachments hence rendering the traditional one-layered detection a futile exercise. This paper suggests a multi-mode phishing detection system with structural validation, machine learning, and generative AI. The system validates DNS and database before the classification is done using some engineered features of the URLs, email, files, and images which is then done by the random forest. Screenshot-based phishing uses OCR to improve contextual reading and minimize false positives, whereas generative AI can be used to improve the context of phishing. The system has an overall accuracy of 97% of detection modules and is implemented with MERN architecture and with secure APIs and a monitoring dashboard.
AOI :10.100.234513.0204
ABSTRACT:
Decimal computation/calculation playing a major role and gaining attention by scientists and analysts because of their significance in various human-centric applications. Executing these operations in hardware offers better speed than software methods. This is highly beneficial for systems that need quick responses. Decimal addition is most important and basic decimal operation among all remaining operations and also one of the essential one. The Complementary Metal Oxide Semiconductor (CMOS) technology has the benefit of decreasing the propagation delay in virtual systems. Decimal adders can be developed as Ripple Carry Addition (RCA) structure or as a Carry Look-Ahead Addition (CLA) structure with more rapid and extra hardware cost. This paper discusses the design and performance analysis of a Carry Look-Ahead Decimal Adder (CLDA) implemented using CMOS technology. This work developed CLDA design to mitigate the latency limitations of conventional Ripple Carry Decimal Adder (RCDA) by bettering carry propagation efficiency. The design is executed and simulated using 45nm and 90nm technologies, and key performance metrics such as delay, power consumption, PDP, EDP are evaluated. The results shows that the proposed CLDA achieves notable performance improvement over the conventional RCDA. In 45nm technology, the CLDA provides approximately 19.9% reduction in delay and 19.8% improvement in Power Delay Product (PDP), indicating enhanced speed and energy efficiency. In 90nm technology, a delay reduction of about 13.3% is observed, while PDP remains nearly comparable. The novelty of this work lies in the effective CMOS-based execution of CLDA and its comparative analysis across different technology nodes. These results confirm that the proposed CLDA is more suitable for high-speed arithmetic applications.
AOI :10.100.234513.0205
ABSTRACT:
The threat posed by plant diseases to the global food supply is a persistent problem, causing yield and financial losses. For diagnostics, speed and accuracy are important, but traditional methods tend to be too reliant on manual inspections that are too slow, and too inconsistent to be useful for the farming communities that need them most. This paper describes an automated diagnosis method using transfer learning with the EfficientNet-B4 architecture. Training was done with the Medley Plant Disease dataset with over 100,000 labeled leaf images spanning 39 classes. Data augmentation and model generalization fine-tuning were performed for the final model to achieve a validation accuracy of 98.41% and a test accuracy of 97.02% model, demonstrating an emphasis on the robustness of the model to diverse types of diseases. This paper illustrates the use of deep learning with high accuracy to be deployed on a mobile and cloud-based with the potential for precision farming at scale and improved decision support for farmers.
AOI :10.100.234513.0206
ABSTRACT:
In this paper an optimized deep learning model has been proposed for automated Steel Surface Defect Detection
[1] to mark the disadvantages of manual inspection methods and baseline object detection models. Steel surface has so many defects which can seriously attack safety and integrity of the product if it’s not detected early. The YOLO family models offer fast and reliable object detection [2], but their performance in the market is often restricted by many limitations. This Study revolves around YOLO V12 [3]. YOLO V12 [3] is a recent growth in object detection. The effectiveness for steel Surface Defect Detection compared to earlier YOLO baseline variants has not been widely evaluated. This study aims to explore the potential of YOLO V12 in this field. When the model is tested on the NEU Surface Defect Database, without any assumptions, the proposed model shows better detection accuracy, progressive stability, increased generalization ability with a mAP50 of 0.863. The presented model gives an efficient real time solution which is needed for industrial deployment and also having quality observation systems.
AOI :10.100.234513.0207
ABSTRACT:
Aquaculture is an important process in the world food security, but fish health management is a major
challenge as bacterial diseases, viral diseases, fungal diseases, and parasitism are widespread. The conventional methods of diagnostics depend rather on the manual checking of the specialists, which leads to subjective diagnostics, slowness of the intervention, and inability to make the interventions in large or distant farms. To resolve these issues, an AI-based Fish Disease Detection System was built based on a database of fish images taken in aquaculture farms and open repositories in 2020-2024 and comprised more than 45 disease types of both freshwater and marine organisms. The preprocessing image pipeline is using the OpenCV-based resizing, normalization, bilateral filters, and contrast enhancement to standardize visual inputs and enhance features in the visual image. The identification of the disease combines a deep learning layer based on MobileNetV2 which is used to classify the disease, Gemini Vision AI which is an advanced multimodal reasoning tool, and a rule-based fallback system to remain reliable in situations of low confidence. White spots, lesions, texture variation and fin ruinous are key visual attributes that are extracted to aid in the effective learning of features. The system produces real-time, high-accurate diagnosis, which gives the disease severity, recommendations to take, as well as prevention guidelines. This intelligent diagnostic framework would boost the accuracy of detection through the combination of deep learning, computer vision, and rule-based reasoning, help decrease fish mortality, and encourage sustainable practices in the aquaculture industry.
AOI :10.100.234513.0208
ABSTRACT:
The Wireless Sensor Networks (WSNs) is a kind of
ad-hoc network of interconnected sensors to observe surrounding environment and record real-time values as the environment changes from time to time. The main source of the power supply for sensors in WSNs is power-constrained batteries built into sensors, which significantly affects the WSN lifetime. The uneven geographical structure causes multiple random paths to be followed in data collection. However, due to constrained battery power capacity and ecological variations, power utilization is a tricky issue in WSN. The work done in this paper proposes an efficient data collection mechanism in Mobile Sink WSN using reinforcement learning in order to address battery power constrain in WSN. The proposed model uses Q-Learning approach that induces automatic learning through the shortest path for data collection. The outcomes of the proposed model’s experiment show it works much better as compared to the existing model.
AOI :10.100.234513.0209
ABSTRACT:
Lithium Batteries are the powerhouses of modern electric vehicles (EVs). The measurement of the battery’s state of charge (SOC) data is crucial to planning for long-distance travel and the breaks required to charge in between. In real life, the sensor data can be prone to external noise and thus the correct SOC may not be displayed accurately. To avoid this, the current manuscript proposes a multimodal neural network and sensor-based SOC prediction for long-distance EV travel plans. The proposed model is trained from real-world battery behavior on parameters such as charging/discharging trends, distance covered and environmental factors that affect the battery SOC. The proposed model learns battery discharge patterns affected by various load scenarios, making the system robust. Furthermore, this can be used for prediction based on the sensor, environmental and load conditions. The proposed technology can be extended to other applications such as drones, delivery robots, and e-bikes for optimizing efficiency, safety and route optimization.
AOI :10.100.234513.0210
ABSTRACT:
Energy-efficient D latch design is crucial for low-power VLSI applications, as latches form the core of sequential circuits and memory elements. This paper presents the design and comparative analysis of 7T, 6T, and 5T D latches optimized for low-power operation. The designs are implemented and simulated using Cadence Virtuoso in 45 nm CMOS technology with a supply voltage of 1V. The performance is analyzed based on power dissipation, propagation delay, and Power delay product. Compared to the 6T D-latch, it is observed in 7T D-Latch that there is -99.9978% reduction in power dissipation and increase in delay by 17.704%. Similarly, when compared to the 6T D-latch, it is observed in 5T D-latch that there is -99.9973% reduction in power dissipation and increase in delay by 16.704%. Also, a 4-bit Serial-In Serial-Out (SISO) shift register is implemented using the D latch designs. The shift register performance is analyzed in terms of power, delay, and PDP. Results indicate that the 5T latch-based shift register achieves the minimum PDP of 95.3 × 10⁻¹⁹, making it highly suitable for low-power VLSI applications.
AOI :10.100.234513.0211
ABSTRACT:
Power consumption is a critical concern in contemporary Very Large-Scale Integration (VLSI) systems owing to increasing demand for energy-efficient electronic devices. Reducing power while maintaining system performance is an important objective in digital circuit design. This paper describes the design and low-power optimization of an elevator controller based on Finite State Machine (FSM) implemented in the Register Transfer Level (RTL). The controller manages elevator operations such as floor requests, movement, and emergency stop using a structured FSM architecture. The suggested design uses various low-power methods, such as power-aware state encoding with gray code and gray encoding with clock gating to lower the total amount of dynamic power used. The Xilinx Vivado suite is used to construct the suggested design in Verilog. According to empirical analysis, every method helps to lower dynamic power. In comparison to the baseline design, the improved design reduces total on-chip power by 36.69 % and dynamic power by 39.0 %. The findings show that the power efficiency of FSM-based elevator controller systems may be enhanced by combining several low-power strategies while maintaining functionality.
AOI :10.100.234513.0212
ABSTRACT:
This paper describes the detailed performance study of a low power, single ended 7T SRAM cell across various technological nodes of 45nm, 32nm, and 22nm. The best result of power dissipation is obtained at the technology node of the 32 nm process for this SRAM cell, where the powers of Write 1, Write 0, Read 1, and Read 0 operations are measured as 0.4523 (µW), 0.778 (µW), 0.259 (µW), and 0.418 (µW), respectively. Additionally, this cell shows better delay performance at the technology node of the 32 nm process, where the Write1, Write0, and Read0 delay times are measured as 18.13 ps, 14.7 ps, and 15.6 ps, respectively, to enable the memory access process faster. Moreover, the single ended 7T SRAM cell shows the highest noise margin at the technology node of the 32 nm process compared to the other two processes, where the highest noise margin value measured at the technology node of the process is 0.32807 V, as opposed to values of 0.30184 V and 0.26861 V at the technology nodes of the other two processes, respectively.
AOI :10.100.234513.0213
ABSTRACT:
Reliable perception capabilities are essential for assistive navigation systems for the visually impaired, which must operate within the constraints of computing and energy limits of edge devices. On low-power embedded platforms, the proposed study aims to implement an object identification framework that is based on edge optimization of transformers. This framework would allow for real-time navigation aid. This study proposes an architecture that efficiently captures local characteristics and global contextual information by integrating a reduced transformer encoder with a lightweight convolutional backbone. Structured pruning and low-precision inference are examples of model-level optimization that drastically cut down on memory usage and computational overhead, allowing for steady real-time performance. To improve domain relevance, training was carried out utilizing a two-stage approach that mixed a task-specific assistive navigation dataset with large-scale benchmark data (MS COCO dataset). Further enhancement in visual conditions, a multimodal sensing using ultrasonic distance estimation is needed to include. The final result of the work shows the optimized transformer framework with good comparison between the latency, detection accuracy, and resource consumption, which shows a suitable model for deployment on the continuous edge. Moreover, considering efficiency aware transformer models shows scalable assisted navigation systems which also boosts autonomous movement.
AOI :10.100.234513.0214
ABSTRACT:
In this work, a smart vacuum cleaning robot based on Arduino Uno microcontroller with an aim of achieving efficient, low cost and autonomous cleaning of floor is studied. The designed system was combination of ultrasonic sensors, which is used to detect the presence of the obstacles, the motor drivers which are used to control the navigation system, and the suction mechanism which is used to remove the dust. The algorithm on which the robot was built features a preprogrammed moving, allowing to move in a system herself without colliding, thus performs optimally in the area coverage. The hardware architecture was adjusted with caution to maintain a balance between the efficiency of the performance and the energy consumption of the hardware and the embedded C programming was used to implement the software logic in the arduino actual IDE environment. The system architecture depicted a smooth flow of interaction between sensing, processing and actuation units. The cleaning efficiency, navigation accuracy and obstacle avoidance capabilities were calculated by experimental testing in the held indoor environments. The results demonstrated that the suggested robot is reasonable outcomes in relation to improved cleaning space as well as restricted hand control compared to conventional cleaning plans. Besides this, the system has been found out to be cost effective, scalable and hence suitable in both the domestic and small scale industry. The research will have an impact on creating easy implementation automation solution and other possible bright opportunities of microcontroller robotics in daily life.
AOI :10.100.234513.0215
ABSTRACT:
This paper presents a novel hybrid image captioning architecture named DCAT (Dual CNN Attention Transformer) that eliminates major drawbacks of CNN-LSTM approaches. The method uses two distinct pre-trained CNNs, DenseNet201 and InceptionV3, working simultaneously as dual encoders. Their output feature vectors (1920-dim and 2048-dim, respectively) are combined by concatenation to create a 3968-dimensional rich fused representation. The fused vector is further divided into four learnable soft visual tokens, which are used as key-value context for a cross-attention transformer decoder. This change removes the sequential LSTM bottleneck in previous methods. A single transformer decoder block has masked multi-head self-attention, cross-attention to soft visual tokens, position-wise feed-forward network, residual connections, and layer normalisation. The training criterion is masked sparse categorical cross-entropy, ignoring padding positions. At inference time, beam search is used for decoding. When tested on the Flickr8k dataset, DCAT achieves BLEU-1: 0.5505, BLEU-2: 0.3643, BLEU-3: 0.2395, and BLEU-4:0.1504, significantly outperforming the Dense Net201+LSTM baseline by 36.7% in terms of BLEU-4 and beating all other CNN+LSTM models. Such findings illustrate that fusing complementary dual CNN encoders and a cross-attention transformer decoder yields very accurate and well-interpreted image captions.
AOI :10.100.234513.0216
ABSTRACT:
The demand to spare energy while creating efficient Very Large Scale Integration (VLSI) systems continues to increase due to the rise in use of portable devices like cell phones that access Internet of Things (IoT). This means that there is a need for designing a low-power computation unit. As a main component of digital processors, Arithmetic Logic Units (ALU) account for a major portion of system power consumption because they operate continuously throughout their operational life cycle. In addition to those reasons, this study proposed an energy efficient ALU as an example of an adaptive clock gating based on reducing switching activity that is not needed, therefore minimizing the amount of dynamic power consumption. The proposed design is developed in Verilog HDL at Register Transfer Level (RTL) and evaluated using standard Electronic Design Automation (EDA) tools via simulation. The fine grained adaptive clock gating concept allows for selective activation of ALU modules where operational requirement dictates. Performance metrics include power, delay, Power Delay Product (PDP), and energy measurements. Based on our results, the proposed adaptive ALU reduces the power from 52.30 µW to 24.10 µW for a 53.92% reduction of power, while the PDP has decreased from 109.83 pJ to 55.43 pJ. The delay increased slightly from 2.10 ns to 2.30 ns, however, the overall energy efficiency increased considerably demonstrating the effectiveness of the proposed solution.
AOI :10.100.234513.0217
ABSTRACT:
The need for optimized arithmetic units has been underscored by the escalating demand for high-speed and energy-efficient digital systems, particularly in the design of Arithmetic Logic Units (ALUs). This study presented a high-speed and low-power Vedic multiplier using an enhanced adder structure suitable for next-generation ALU applications. The proposed design is based on the Urdhva Tiryakbhyam algorithm, which allows parallel generation of partial products and thus reduces computational delay. A modular Vedic multiplier architecture is realized and merged with a hybrid enhanced adder to minimize carry propagation delay as well as switching activity. The entire design is described using Verilog HDL and synthesized in Xilinx Vivado for performance evaluation. Results show that the proposed design attains a delay of 24.8 ns at an operating power of 190 mW, which is better than conventional multipliers by about 45% in delay, 40% in power consumption, and up to 67% improvement in Power Delay Product (PDP). The enhanced hybrid adder reduces this further to 6.1 ns with lower power consumption; hence such an architecture becomes more appropriate when looking for high-performance energy-efficient ALU systems.
AOI :10.100.234513.0218
ABSTRACT:
As VLSI systems continue to grow in complexity, there is increasing demand for verification methodologies that are both efficient and scalable. Functional verification of widely used communication protocols, such as I2C and SPI, is particularly problematic because traditional verification methods often do not provide complete coverage or identify defects or errors early during the design phase. This delay in identifying defects or errors leads to an increase in the overall time and cost of the manufacturing process. This work presented a Coverage Driven Verification (CDV) and Assertion Based Verification (ABV) methodology that combined Constrained Random and Directed Testing, System Verilog assertions, and functional coverage models to provide a comprehensive verification process for VLSI systems. A reusable testbench environment based on the Universal Verification Methodology (UVM) is developed to provide verification of the functionality of the protocol, the timing constraints associated with the protocol, and corner case scenarios. The proposed CDV/ABV methodology achieved 98.75% functional coverage, 97.90% code coverage and 100% assertion coverage. Through fault injection analysis, a 100% detection rate for all fault conditions was achieved, while regression testing produced a 98.33% efficiency rating. In addition, the simulation time was reduced from 120 minutes to 85 minutes, which has improved the overall verification efficiency. The results demonstrated that the proposed methodology improves reliability, scalability and performance in the verification of VLSI systems.
AOI :10.100.234513.0219
ABSTRACT:
For the authors, the fast development of online learning requires new solutions instead of conventional Learning Management Systems (LMS). This research paper is a report about an AI-powered Progressive Web Application (PWA) that will help to optimize the online learning experience by making them more accessible, personalized, and intelligent. The combination of the functionality of PWAs and Artificial Intelligence allows the system to support offline interactions, be compatible across platforms, and provide users with real-time impressions with the push messages. The suggested application uses the latest technologies like React.js, Node.js, Firebase, and the Gemini API to provide voice-based search and AI-based content suggestions. The system utilizes machine learning modalities to study user behavior and dynamically adjusts an educational content which adapts with the personal preferences and learning tendencies. Also, voice-enable search capability enhances usability since users can call the information with the highest possible efficiency without having to use any text-based query. An end user testing of the system shows that learning accessibility and engagement have improved significantly. The findings show that AI implementation combined with the method of PWA technology improves the user experience, stimulates continuous interaction, and increases the desire to read educational materials. Moreover, the system has a high accuracy level when it comes to voice query response, which confirms its efficiency in the role of an intelligent learning assistant.
AOI :10.100.234513.0220
ABSTRACT:
An efficient bug tracking solution gives all members of the software development team (testers, coders, managers, administrators, etc.) a means to locate, submit, distribute and solve software bugs quickly and accurately. The purpose of this bug tracking solution is to provide a quick, accurate and reliable means for these individuals to work together with each other. The design of the bug tracking solution is based upon an established process for reporting bugs, secure user logins, and the ability to view real-time status updates of both software bugs and bug reports. Please note that the two primary views of the software for the bug tracking solution are: The back-end (software logic), and the end-user's experience managing and tracking bugs in the solution. The back-end logic of the solution is designed using Object-Oriented Programming (OOP), allowing for the creation of reusable, modular components and customized to each user-type. When a user (any user type) submits the status of their bug report, the user's request will first check that enough information on the bug being reported exists, that it has been validated as correct, and that this information can be stored in a manner that can track that bug in the future. Additionally, the bug tracking solution allows the insertion of bug severity and bug urgency levels, which enables users to determine how they should proceed in resolving a bug. In the user face thing, a Java Swing based GUI that gives users a place thats pretty easy to use. Where, users can put in bugs, update how far along they are, look over what they've been given to do, and take care of system data. Future stuff to hook it up with MySQL will make sure the data is safe, and Maven project management will help it grow and handle what it needs to work. These parts work together, to make a user-friendly tool that is definitely going to improve software quality and help the teams work together better.
AOI :10.100.234513.0221
ABSTRACT:
A majority of experts agree that THz (terahertz) communication will be a primary enabler for 6G (sixth generation) wireless networks due to both its ample spectrum resources and enormous capacity potential. Nevertheless, reliable performance evaluations of THz systems require precise modelling of both channel fading and phase-based uncertainties. In general, traditional methods for deriving symbol error rates (SERs) assume either that there is perfect phase alignment or use simple approximations that ignore the statistical fluctuations of phase. Recent studies have found that including a probability density function of the phase jitter into the derivation of the SER results in substantially better analytical accuracy and more reliable predictions.
In this paper, a general theoretical analysis of phase-jitter-based SER modelling for THz and mixed THz/RF communication systems will be outlined. This effort adds to the analytical framework presented by [1], extends the previous modelling to include THz-related issues, and provides a general discussion on phase-aware detection, enhanced statistical reliability by making use of phase-based measures, and performance scaling with respect to the SNR (signal-to-noise ratio). Additionally, we will include comparison tables illustrating that the use of phase-jitter-based modelling provides a consistent SER improvement over traditional methods. Results indicate that incorporating statistical phase-aware modelling will yield improved analytical accuracy, enable a lower SNR to meet target reliability levels, and provide a basis for optimising future 6G networks.
AOI :10.100.234513.0222
ABSTRACT:
In this paper, the author has suggested a contactless toll collection system, which employs Artificial Intelligence (AI) and Number Plate Recognition (NPR) to enhance efficiency at the toll plazas. The system will be constructed under Raspberry Pi with a USB camera which will capture pictures of the vehicles in real time. A deep learning model based on the YOLO system detects the license plates and the vehicle numbers are recognized through the Optical Character Recognition (OCR). A sensor in the ultrasonic mode detects the coming vehicles and the automatic toll processing and the opening of the gates with the help of a DC motor and a driver. A database stores vehicle-related information and account balances, and the deduction of tolls can be done automatically. Status of transactions is displayed on an LCD screen, and a buzzer is used to notify the user about any errors or confirmations. Experimental findings indicate that the system is reliable to use, highly accurate and lessens congestion, delays and manual work, which is very economical as compared to traditional method of collecting tolls.
AOI :10.100.234513.0223
ABSTRACT:
The prediction of time-series is a pillar of intelligent systems used in areas like energy, finance, medicine, and industrial Internet of Things; accurate predictions are the foundations of automated decision-making, effective resource distribution, and sound mechanisms of identifying anomalies. Archaeologies, such as recurrent neural networks (RNNs), long short-term memory (LSTM) networks, temporal convolutional network (TCNs), and Transformers, have significantly surpassed classical statistical models like ARIMA and exponential smoothing models in terms of precision in prediction since the emergence of deep learning. However, the natural incomprehensibility of such deep models the so-called black-box problem limits their application in safety-critical fields where regulatory standards, human trust and explainability cannot be compromised. The presented review is an overview of recent developments in the intersection of deep time-series forecasting and explainable artificial intelligence (XAI) along with a specific focus on their application to intelligent systems. We provide a taxonomy of deep forecasting architectures which includes RNN based, CNN/TCN based, Transformer based, hybrid and emergent foundation models. We then summarize a systematic review of XAI methods specific to time-varying data comprising feature attribution, time-sensitive saliency maps, attention-based explanations and intrinsically explainable architectures. Healthcare, energy, and industrial IoT Our synthesis will represent a more general finance, healthcare, energy, and industrial IoT, demonstrating how explanations assist in model debugging, building trust in a model, and human-AI collaboration. Among the open challenges that we address and discourse include scaling XAI to multivariate high-frequency data streams, standardizing measures of quality of explanations, and incorporating domain knowledge into pipelines in forecasting
|
Download Paper template of Proceedings of Micro2026 from this link.
For query about publishing for printed or online version of conference Proceedings, write to: info@actsoft.org or talk to: +91-6291839750
About indexing of AOI(Applied Object Indexing): it is a new and advanced indexing method for digital and real life hard objects. Any book, paper, picture, medicine, Land, Building etc. can be indexed with unique number. Clicking that number, the object can be identified with its contents. --------------------------------------------------------------------------------------------- AOI :10.100.234513.0224
ABSTRACT:
Human communication is inherently multimodal, relying on a complex interplay between facial expressions and vocal prosody. However, traditional Affective Computing systems often rely on unimodal analysis—typically visual—which renders them susceptible to error when subjects mask their true emotions (e.g., a "social smile" concealing anxiety). To address this limitation, this paper proposes a real-time Multimodal Emotion Perception System that integrates visual and acoustic cues using a sequential asynchronous pipeline. The architecture leverages DeepFace (VGG-Face) for facial feature extraction and a fine-tuned Wav2Vec2 Transformer for speech sentiment analysis. The Multimodal Congruence Index (MCI) is a new decision level fusion algorithm that is used to measure the semantic agreement between modalities. The results of the experiment prove that unimodal accuracy varies in noisy conditions, whereas the proposed fusion model manages to detect emotional conflict with accuracy 88% correct, which is a strong solution to behavioral analysis and detecting lies in the interview context.
AOI :10.100.234513.0225
ABSTRACT:
Various 8T SRAM cell designs are comparatively analyzed to determine the most performance-efficient configuration. Monte Carlo simulations were performed at the 22-nm technology node to evaluate key design metrics. The Read Stack 8T (RS8T) and Low Power 8T (LP8T) cells demonstrate the shortest read access time (TRA), whereas the Zigzag 8T (ZZ8T) exhibits the longest. The Differential Data-Aware Power-Supplied 8T (D2AP8T) achieves the fastest write access time (TWA), while the Feedback Control 8T (FCS8T) shows the slowest. Among all designs, D2AP8T also has the lowest hold and write power consumption, making it the most efficient 8T SRAM cell overall, with superior write performance and energy efficiency at the cost of slightly higher read access time.
AOI :10.100.234513.0226
ABSTRACT:
The BERT model shows a steady decrease in both training and validation loss, as well as an improvement in accuracy. This shows that the model is converging and learning the context effectively, without any signs of overfitting. The digital news media and social media have significantly changed the manner in which people consume information. Although the digital news media has improved the dissemination of information, it has also resulted in an increase in the dissemination of false information. False information is disseminated rapidly due to its sensational headlines and the fact that it is disseminated without verification. Proposed in this paper is a comprehensive text-based fake news detection system using machine learning and deep learning techniques. Several traditional machine learning classifiers are experimented with TF-IDF feature extraction techniques, including Naive Bayes, Logistic Regression, Support Vector Machines, Random Forest, and SGD Classifier. In addition to this, several deep learning models like Convolutional Neural Networks (CNN), Bidirectional LSTM (BiLSTM), and a fine-tuned BERT transformer model are also employed for semantic feature extraction of the text. The performance of all models is compared using standard evaluation metrics such as accuracy, precision, recall, and F1-score. The experimental results show that the transformer models are performing better than the traditional models because of their understanding of the context, and the traditional models are performing well with less computational complexity. The importance of the combination of classical and modern approaches in NLP for effective misinformation detection is the focus of this paper.
AOI :10.100.234513.0227
ABSTRACT:
The world where we are currently living in is growing very fast, as everything has started to shift online. But this has its own cons, as the career preparation environments have its own extremely discontinuous methods where skills do not match with interests, project building, proper mentorship for chosen career path, and job listings are used by isolated and uncoordinated systems. This environment presentation is difficult to the learners’ because it is hard to keep track of their progress, time to time career guidance by mentors who are already in that position and can guide from their own mistakes so the new comers do not get stuck in a loop. The rat race that we are currently in, where everyone is going in the same direction and almost none are following their own interests as they do not even know what career would they pursue with their set of skills, they do not know if there is any career path with their match of interest. The important part is interest creates efficiency. The solution to this dilemma presents this paper Luminary X, which is a single web-based system that has it all. This platform has a section to choose the skills that match with the learners’ interests. After that they will get AI powered assistance to start a career that uses those skills. They will get guidance from alumni who’s has done great in their chosen career. Users can upload their projects and get help in building it to an advanced level. This platform enables an AI powered guide, who will help them to reach their desired goals. This makes the model from traditional linear learning environment to a structured in order platform. To plan project work that can turn into portfolio ready projects. Luminary
X is built on a scalable web architecture using Next.js, React.js, Tailwind CSS, Node.js, and Express.js. The platform integrated user profile, skill tracking, project management, alumni interaction, and AI-powered career guidance within a single responsive interface. It enables learners to monitor their progress, receive personalized mentorship, and achieve structured career development through a unified ecosystem.
AOI :10.100.234513.0228
ABSTRACT:
QFed-Anomaly During this work, we present QFed- Anomaly, which is a hybrid quantum-classical federated learning framework that offers a privacy-saving method of data detection of anomalies in a massive Internet-of-Things (IoT) system. The suggested architecture is also concerned with solving three essentially overlapping issues: the bulky and distributed (around the world) character of IoT information streams, high sensitivity to privacy, and limited computational ability of modern Noisy Intermediate Scale Quantum (NISQ) devices.
Quantum feature maps are trained on the node of the edge, with the model parameters being interpreted using the classical Federated Averaging (FedAvg) protocol. This system is a guarantee that raw data will be stored in the local devices, and difficult data processing will be provided by the global model operation, which will allow applying differential privacy. A quantum feature-mapping pipeline is again enhanced with a Gaussian mechanism that gives formal (ϵ, δ) -differential privacy without any: This is not at all a significant loss in detection performance.
Empirical analysis conducted on four most popular intrusion and IoT security datasets, which include NSL-KDD, CI- CIDS2017, IoT-23, and UNSW-NB15 containing more than four million samples and deployed in a realistic federated environment with heterogeneously partitioned data in a plurality of edge nodes, confirms that QFed-Anomaly achieves an average detection rate of about 95.0%. It performs better than conventional support -vector machines (SVMs), federated SVM baselines, and centrally trained quantum support -vector machines (QSVMs), and has an up to three-fold speed advantage over purely classical methodologies.
Also, the model has a high resilience to adversarial perturbation, with more than 90% accuracy to attack by FGSM, PGD, and Carlini Wagner (CW), and has practical privacy-utility trade-offs consistent with GDPR-compliant IoT applications. Therefore, the given research provides a workable framework of implementing the implementation of quantum enhanced, privacy preserving network anomaly devices in actual secure IoT settings. -confidential network detection devices in the context of real-world secure IoT systems.
AOI :10.100.234513.0229
ABSTRACT:
Smart homes and intelligent security systems have gradually emerged as a result of increasing risk of intrusion, damage, or unauthorized access to the residential as well as institutional environments. To overcome the challenges, this paper proposes a smart automation security system using EAP32 that combines motion sensors, environmental, real-time connectivity, as well as automatic notification for a holistic, economical, and effective surveillance system. The system applies PIR sensors, ultrasonic sensors, magnetic door sensors, as well as ESP32-CAM for smart movement detection, as well as image capture, that too only when required, significantly lowering energy as well as network costs. Since it operates with the microcontroller as its main part, decision-making, as well as sensor data integration, is locally performed, as well as it securely communicates with cloud services. Automated notification messages have been implemented for smartphones, allowing for immediate user action regardless of physical presence within the secure locations. Results have confirmed high accuracy, decreased false alarms, as well as significantly fast notification times. This research establishes the importance as well as importance of effective integration of IoT connectivity with affordable micro-controllers for developing a trustworthy, adaptable, as well as autonomous surveillance system for smart homes/offices for the coming years.
AOI :10.100.234513.0230
ABSTRACT:
This work provides a comprehensive evaluation and comparison of various active two-phase charge pump voltage multiplier topologies for low-current and high-voltage applications. Topologies studied include Dickson, series-parallel, cross-coupled, exponential, and Fibonacci charge pumps. All circuits are simulated with 180 nm CMOS technology to generate an output voltage of around 5 V from 1.8 V input. The performance is measured using essential metrics such as rise time, power dissipation, voltage conversion efficiency (VCE), power conversion efficiency (PCE), and area. The results indicate that the cross-coupled architecture achieves the best efficiency with VCE of 96.33% and PCE of 71% while consuming less power and having fewer stages. The series-parallel topology provides the quickest response, while the Fibonacci structure offers minimal area. A figure of merit (FoM) is introduced for the overall comparative analysis, showing that the cross-coupled architecture performs best. This study supports the selection of appropriate charge pump topologies for low-power integrated applications like IoT systems, biomedical devices, energy harvesting circuits etc.
AOI :10.100.234513.0231
ABSTRACT:
The problem that I am going to address in this paper is a real-life issue of how to create a face recognition system with only a couple of handfuls of images of individuals? Majority of studies presuppose that we possess hundreds of photos, yet that is not the case in the real-life deployments. We have developed a system that operates with 3 or less to 10 enrolment photos to identity. This is not us developing a new neural architecture but approximately to have the existing facial embeddings significantly more effective with thorough engineering. We present four techniques that use test-time augmentation to minimize noise, quality-based weighting to deal with bad photos, k-means clustering to make use of pose diversity, and a hybrid similarity measure that, as a combination of cosine and Euclidean distances. Comparison with standard ImageNet models (VGG-16, ResNet-50, etc.) demonstrates that our method achieves 98.7% accuracy, within 0.3 points of VGG-16 with an incredibly low number of samples. The system is real-time and uses standard hardware and an automatic threshold calibration technique which eliminates manual tuning. We've also addressed the very often ignored problem of cross-device enrolment. When a person signs in using the phone, but is not identified by a web camera. The full implementation is modular so you can replace more desirable embedding models as they are developed.
AOI :10.100.234513.0232
ABSTRACT:
The Industrial Internet of Things (IIoT) has emerged as a fundamental component of Industry 4.0, enabling intelligent automation, continuous monitoring, and data-driven decision-making in modern industrial environments. However, the growing connectivity among industrial devices and systems has introduced significant cybersecurity challenges, particularly insider threats originating from authorized users or internal entities with malicious intent. Unlike external cyberattacks, insider threats are difficult to detect due to their legitimate access privileges and the subtle nature of anomalous behavior, posing serious risks to operational integrity, data confidentiality, and industrial safety.
This study proposes an intelligent anomaly detection framework for identifying and mitigating insider threats in IIoT environments. The proposed approach employs a hybrid deep learning architecture that integrates Autoencoders and Long Short-Term Memory (LSTM) networks to effectively capture both structural and temporal patterns in industrial data. The Autoencoder component learns the characteristics of normal operational behavior and identifies deviations through reconstruction errors, while the LSTM network models sequential dependencies and temporal correlations among user activities and device interactions.
The effectiveness of the proposed framework is evaluated using the benchmark TON_IoT dataset and assessed through widely used performance metrics, including accuracy, precision, recall, F1-score, and detection latency. Experimental results demonstrate that the hybrid Autoencoder-LSTM model outperforms conventional machine learning techniques, particularly in detecting subtle and sophisticated insider anomalies. The proposed framework provides a scalable and efficient security solution for Industry 4.0 ecosystems and has potential applications in smart manufacturing environments, critical infrastructure protection, and cyber-physical systems.
AOI :10.100.234513.0233
ABSTRACT:
The swift growth of online used-car market place has raised the pressure on proper, transparent, and automated systems of valuations od car prices. Pricing of used car is a complicated fact, because it is related to several variable which are interconnected including age of the vehicle, miles covered, type of fuel used, transmission, ownership record, and market demand. Traditional valuation techniques founded on subjective appraisal or rudimentary statistical model may be subjective as well as weak in predictive ability particularly when non-linear depreciation patterns are dealt with. In this paper, Car-Valuate is proposed, a car valuation system that works based on machine learning and uses historical data to predict real used car prices. First, a Linear Regression model is used as a baseline because they are not complex and easy to interpret but in the actual experiment, it is shown that it has significant mistakes in prediction because it assumes linear relations between variable. In order to address these shortcomings a Random Forest Regress or is used which uses ensemble learning to learn the intricate non-linear interactions of features. The systematic data pre-processing steps to be included in the proposed system are the processing of missing values, encoding categorical features, feature selection, and the logarithmic transformation of the price values. The trained model will be implemented with the help of the Flask-based backend and combined with a web-based frontend written in HTML, CSS, and Bootstrap, which will allow interacting with the user in real time. The evaluation of the results of the performance shows that the predictions made by the Random Forest model are in close correlation with the real prices in the market, and they are far more effective and robust as compared to the Linear Regression. Evidence proves that Car-Valuate is a powerful system.
AOI :10.100.234513.0234
ABSTRACT:
Ambiguity in natural language queries continues to be a significant barrier to accurate question answering and conversational systems, resulting in many incorrect or unreliable answers. The prevailing methods, ranging from AmbigQA style LLM pipelines to rule-based systems such as TaskLint, generally rely on large models that require powerful GPU capabilities or rigid taxonomy-based structures; therefore, usability and real-world practicality are inhibited by these limitations. Detect-Then-Clarify is an agile and lightweight system for detecting and clarifying ambiguities via the detection, classification, and clarification processes, reducing the reliance on large language models or expensive computational resources. Experimental evaluation demonstrates that this system effectively distinguishes between ambiguous requests, resolves ambiguities through detailed exchanges, and builds user trust efficiently while minimizing unnecessary communication between parties.
AOI :10.100.234513.0235
ABSTRACT:
Blockchain technology, a decentralized and immutable distributed ledger, has emerged as a promising solution for addressing modern cybersecurity challenges beyond cryptocurrency applications. This paper provides a comprehensive and systematic overview of the advances in blockchain-based cybersecurity over the last few years, through an examination of current academic research, industry publications, and technical reports. The review explores how blockchain is being incorporated into a broad and growing range of cybersecurity applications, such as blockchain for identity and access management (IAM), data integrity, and Internet of Things (IoT) security, and discusses new types of threats including smart contract vulnerabilities, attacks on cross-chain bridges, privacy issues, and emerging cyberattacks. In addition, the study explores innovative defense strategies and techniques which integrate the concepts of blockchain, Artificial Intelligence (AI), Machine Learning (ML), Zero-Knowledge Proofs (ZKPs), and Post-Quantum Cryptography (PQC) to bolster security, privacy, scalability, and resilience against emerging quantum-based challenges. The paper also addresses representative case studies, adoption barriers, regulatory aspects and scalability considerations affecting mass adoption of blockchain. Lastly, research gaps and future research directions are outlined, highlighting the necessity for secure, scalable, interoperable and regulatory compliant blockchain-based cybersecurity frameworks. The results prove that blockchain technology combined with the new technologies offers a solid basis for the creation of new generation cyber security solutions.
AOI :10.100.234513.0236
ABSTRACT:
Brute-force attacks are the most common type of attacks to gain unauthorized access. The use of some of the most common brute-forcing tools includes "Hydra", "John the Ripper", "Hashcat" and "Aircrack-ng". These are some automated tools that are used nowadays to crack username and password through continuous attempt to guess the credentials using a file which is called a wordlist. A wordlist contains all the possible list of words and combination of letters, numbers or characters that could be the username or password. The traditional way of rate limiting or account lockout policy is mostly futile against new technological advancements of Brute-Force attacks that closely imitate legitimate user conduct. "Anomaly detection refers to the problem of finding patterns in data that do not conform to expected behaviour."[6] This research proposes a Brute-force Prevention Framework that uses Machine-Learning to prevent attacks. A dataset of 50,000 authentication logs was generated, that contained legitimate user logs as well as attackers. It uses the mechanism of differentiating through different factors that include number of failed attempts, time interval between attempts, IP address and session behaviour. It uses authentication logs and typical user conduct to differentiate between a legitimate user and an intruder who is trying to brute-force credentials to gain unauthorized access. Multiple classification algorithms were used, in which Random Forest model achieved highest accuracy with staggering 96.4%, 94% precision and only 4.3% false positives. After experimental analysis, we found out that machine learning models keep improving their detection accuracy to the point where the chances of trespassing become minimal. This project provides an intelligent and adaptive solution for securing systems against these types of attack.
AOI :10.100.234513.0237
ABSTRACT:
The problem of assessing reliability when there is uncertainty in the available information on network failures constitutes a serious limitation of classical probabilistic approaches to reliability estimation. Fuzzy set theory provides an attractive solution through the use of fuzzy numbers to represent the reliability of the components involved, thereby providing a mathematical foundation for dealing with epistemic uncertainty. Although several studies have investigated fuzzy reliability models for series-parallel and bridge networks, the issue of fuzzy K-terminal reliability, which involves the connection of only a selected set of nodes, has not yet received substantial attention in the context of fuzzy sets. Each edge in the network is modeled using trapezoidal fuzzy probabilities defined by four-parameter membership functions to capture uncertainty. The analysis incorporates minimal Steiner tree concepts for representing feasible connectivity among terminal nodes, while system reliability is evaluated using an inclusion-exclusion framework under fuzzy arithmetic operations. α-cut representation is employed to characterize and reconstruct the resulting fuzzy reliability structure, and Monte Carlo simulation is used as an external validation mechanism. The methodology is implemented computationally in Python and supported through graphical analysis for interpretability. The final results indicate that the fuzzy K-terminal reliability has a defined support and core, showing high but not near-certain connectivity among the designated terminal nodes. Overall, the proposed framework demonstrates the effectiveness of fuzzy arithmetic in assessing K-terminal network reliability under epistemic uncertainty.
AOI :10.100.234513.0238
ABSTRACT:
Suicide is a relevant worldwide social health issue and text analysis in social media can present an opportunity to intervene in time. This paper provides a comparative analysis of the conventional feature extraction strategies and transformer-related methods of automated suicide risk identification based on posts in online mental health groups. Five methods of extracting features were tested: TF-IDF, Bag of Words (BoW), N-gram, fine-tuned BERT and MentalBERT embeddings. Two BERT models were explored: general-purpose BERT and MentalBERT, a domain-targeted model that was pre-trained on mental health discussions on Reddit communities. 20,000 balanced samples were fine-tuned with Five machine learning classifiers: Random Forest, Support Vector Machine (SVM), Logistic Regression, XGBoost, and LightGBM. Findings have shown that the MentalBERT model with the Logistic Regression model surpassed all the other traditional feature extraction methods and baseline BERT Embedding by a significant margin with a high F1-score of 0.9808 and AUC of 0.98. Feature importance analysis showed that there are linguistic patterns with high linkage with suicidal ideation. Specialized linguistic patterns in mental health discourse: domain-specific pre-training of MentalBERT was very effective in capturing subtle linguistic patterns in mental health discourse, showing the vital role of specialized language models in sensitive text classification tasks in clinical Domains.
AOI :10.100.234513.0239
ABSTRACT:
The increasing demand for eco-friendly and high-efficiency photovoltaic technologies has intensified potential research on lead-free perovskite solar cells. However, in planar and tandem perovskite material based solar cells it is still a challenge to maintain proper band alignment between successive layers along with maintaining lower rate of charge carrier recombination yielding high photo conversion efficiency (PCE). This paper proposes an environment friendly lead free ITO/TiO2/CsSnI3/PEDOT:PSS/Al-based planar solar cell designed for efficient capture of the solar radiation. CsSnI₃ acts as the narrow-bandgap active absorber layer, enabling effective photon absorption and generation of electron–hole pairs. The layer below the front electrode serves as an electron transport layer (ETL) whereas the layer above the back contact works as hole transport layer (HTL), and the materials are selected based on favorable band alignment with the absorber layer. The layers above absorber are chosen to function as window layer for the incident solar irradiation. The energy band diagram is also presented to illustrate carrier transport and reduced recombination losses across the ETL and HTL. The simulation study is carried out using Ansys Lumerical implementing FDTD method along with using it’s CHARGE solver module which suggest that a perovskite layer thickness of 300 nm yields the optimal performance, achieving a photo conversion efficiency (PCE) of 32.1%, along with a short-circuit current density (Jsc) of 42.71 mA/cm², an open-circuit voltage (Voc) of 0.91 V, and a fill factor (FF) of 82.4%. The obtained PCE is among the highest values reported recently for perovskite lead-free planar solar cells.
AOI :10.100.234513.0240
ABSTRACT:
Given the capability of the “Software Defined OMR” engine to work with black and white photocopied forms, the proposed approach allows for eliminating the need for special-purpose scanning equipment as it allows answering sheets to be analysed using regular smartphone photos. The technology makes use of micro services architecture, a Convolutional Recurrent Neural Network (CRNN) for transcribing handwritten student IDs, circumventing optical constraints of conventional OMR technologies and Four Point Perspective Transforms to deal with issues of illumination and geometric distortions. An additional four-layered Adaptive Intelligence Layer improves the post-OCR identity matching precision from 71.3% (raw CRNN) to 99.1% in the case of enrolment numbers by employing an adaptive corrections log, session memory, and fuzzy database matching and human-in-the-loop approach. According to the performance indicators, ADEL offers an efficient and hardware-independent system for education-related institutions looking to combine their offline and online performance data with 99.7% bubble detection accuracy, 99.1% enrolment number recognition accuracy after adaptive identity resolution and 0.7 seconds per form processing speed.
AOI :10.100.234513.0241
ABSTRACT:
Topological band structure engineering enables the development of robust light propagation states known as topological interface states, across a variety of media. These light states emerge at the interface between two topologically distinct lattices and are characterized by their remarkable robustness against structural imperfections. To highlight the distinctive properties of topological interface states in direct comparison with conventional light-guiding mechanisms, we present a systematic investigation of light transport in two-dimensional photonic crystal (PhC) waveguides. The study traces a clear progression from conventional defect-engineered designs to topologically robust architectures, enabling a direct and meaningful comparison between the two approaches. Initially, a topologically trivial 2D square unit cell with lattice constant a is designed and its photonic band structure is computed using COMSOL Multiphysics®. Based on this unit cell, a line-defect waveguide of width d=2a is introduced. Further, a tapered 2D waveguide is realized using point defects, with the defect width gradually reduced from d=4a at the input to d=2a at the output. Although both of these conventional PhC waveguides show efficient light transmission as obtained from comparable input and output field intensities, they lack robustness against structural perturbations and fail to support propagation of light through a bent path. Consequently, a topological defect-guided waveguide is demonstrated, utilizing a topologically non-trivial 2D unit cell design, verified via parity analysis of eigenmodes obtained from eigenfrequency simulations in COMSOL Multiphysics®. The conjoint array of trivial and non-trivial unit-cells, enables robust light transport through topological interface states along a 15-degree bend.
Donate to book project: from India pay by scanning QR code:
Account details: Account Name: Applied Computer Technology Account Number: 35532552072 Bank name: STATE BANK OF INDIA IFSC Code: SBIN0001404 Account Type: Current Branch Name: Kamarhati Address: 1, B.T.Road, Kolkata-700058, West Bengal, India. PAN : AVPPA7870E Phone : +91-8420582707 -------------------------------------- |