Thesis (Bachelor of Science in Computer Science)
Permanent URI for this collectionhttps://hdl.handle.net/10361/28437
Browse
Recent Submissions
Open Accesslistelement.badge.dso-type Item , Neurodiagnostic advancements: Deep learning-driven MRI processing for ADHD detection(BRAC University, 2026-01) Shahat, Md. Shajjat Hossain; Tamim, Fardin Hassan; Sagor, Al-Amin; Mayabi, Simran Zaman; Biswas, Smita; Prome, Tasnim Ahsan; Tasnim, Anika; Department of Computer Science and EngineeringThe vast advancement of neurodiagnostic technology offers an opportunity to transform mental health assessment with more precise, objective evaluations. Which previously was reliant on subjective matters like individual perspectives (Doctors in this case), interpretations, and such. With the advancement of technology in the field of medical science, there are so many resources that can be used and worked on to find efficient and better solutions for any disease or health issue. Assessing the state of one’s mental health will have an impact on unleashing one’s potential fully. This research solely focuses on the application of deep learning-driven MRI processing in neurodiagnostic frameworks to improve mental health assessment. In this research, our main focus was on Attention Deficit Hyperactivity Disorder (ADHD), which is a very common and concerning mental health issue, especially amongst young children. To understand its mechanism, it’s crucial to investigate how functional connectivity patterns between various brain regions portray the neurobiological processes of the condition. With the help of the Graph Attention Network with Phenotype fusion (GAT-PhoNet) method, a deep learning model, we analyze the brain activity pattern through graphs, which is created from the ADHD-200 dataset resting-state fMRI data and phenotypic data that include age, gender, and IQ information. The model aimed to identify the brain activity patterns at the network level while also obtaining the crucial clinical covariates. Our GAT-PhoNet model achieved a classification accuracy of 72.09% along with an AUC-ROC score of 0.7517 and F1-scores of 0.74 for healthy controls and for ADHD subjects 0.68. This proved to be an understandable and powerful technique, which helps with ADHD detection through objective methods. The study demonstrates how multimodal graph-based deep learning systems can enhance neurodiagnostic procedures by providing more accurate quantitative analysis results. Open Accesslistelement.badge.dso-type Item , Development of blockchain-enabled radio communication protocol for secure information transmission(BRAC University, 2026-01) Nishu, Nusrat Jahan; Dipti, Faria Islam; Sarkar, Md. Arafat; Mukta, Jannatun Noor; Department of Computer Science and EngineeringIn modern communication systems, maintaining confidentiality, integrity, and availability plays a vital role in protecting information against people who aim to exploit vulnerabilities. Traditional radio communication protocols are reliable, but they lack robust security mechanisms, which causes data theft by interception and data tampering. Radio frequencies are crucial for emergencies, but are not reliable for transmitting confidential information without encryption, and Blockchain technologies are implemented in Radio Access Networks (B-RAN), but are still widely unexplored in radio systems. In this aspect, blockchain technologies in radio protocols can help to enhance messaging integrity and security. Our research aims to develop a blockchain-based decentralized communication protocol for amateur radio systems, which is designed to ensure secure and tamper-proof information exchange in real-time. Open Accesslistelement.badge.dso-type Item , Proof of physical activity (PoPA): A novel blockchain consensus algorithm based on physical work(BRAC University, 2026-01) Advan, Rodel; Nabi, Md. Mashfiqun; Sahib, Faiaj; Raihan, Aarshi Durriya; Ferdous, Md Sadek; Department of Computer Science and EngineeringBlockchain technology, which has persisted for a considerable period and has achieved remarkable advancements since the emergence of Bitcoin, still faces a challenge in normalizing cryptocurrency mining using smartphones for general users. This difficulty stems from the complexity of certain algorithms, such as Proof of Work (PoW) and Proof of Stake (PoS), which smartphones cannot effectively handle. Despite numerous proposals for smartphone-based consensus algorithms, integrating various Internet of Things (IoT) accessories, like smartwatches, into the smartphone consensus for enhanced processing capabilities remains uncommon. To address this research gap, we present a novel smartphone-based blockchain consensus protocol called Proof of Physical Activity (PoPA). PoPA leverages human physiological data collected from smartwatch sensors to participate in the mining process. This approach not only enhances the mining efficiency but also serves as a proof of physical fitness, supporting various applications. Open Accesslistelement.badge.dso-type Item , Meta-learning for zero-shot skin lesion classification across unseen smartphone devices(BRAC University, 2026-01) Azmine, Ishmam; Rafi, Fuad Ibne; Dewan, Shadab Uddin; Ishraq, Quazi Tousif; Rahman, Chowdhury Mofizur; Ahmed, Md. Sabbir; Department of Computer Science and EngineeringThe biggest challenge to automated classification of skin lesions is robust crossdomain generalization especially when models trained on dermoscopic images are applied to smartphone images. This paper assesses how meta-learning can achieve cross-domain robustness on a strict zero-shot protocol, where no target-domain images, labels, or hyperparameter feedback are used during training or model selection. Models are only trained on HAM10000, BCN20000 dermoscopy datasets and evaluated on PAD-UFES-20 smartphone dataset with a single six-class taxonomy. We compare the baselines of supervised transfer learning with several meta-learning paradigms, which include Prototypical Networks, Meta-Baseline, FEAT, MetaOpt- Net. Experiments are evaluated on convolutional and transformer-based backbones and evaluated in a similar episodic format. Macro-averaged F1-score is used as the main measure of performance because it accounts for class imbalance. Across backbones, the monitored models record a significant drop in source to target performance irrespective of good source domain validation outcomes. Meta-learning techniques mitigate this degradation, with relative macro-F1 gains of about 20–70% more than supervised baselines in a variety of backbone architectures in various configurations. These findings prove that episodic meta-learning can reduce the domain-induced failures in the task of classifying skin lesions in the absence of any target-domain exposure. The suggested evaluation environment captures the actual restrictions of teledermatology implementation and gives empirical data of metalearning as a potential approach towards zero-shot cross-domain medical imaging analysis. Open Accesslistelement.badge.dso-type Item , Predicting a t20 cricket match result while the match is in progress(BRAC University, 2015-08) Munir, Fahad; Hasan, Md. Kamrul; Ahmed. Sakib; Quraish, Sultan Md.Data Mining and Machine learning in Sports Analytics, is a brand new research eld in Computer Science with a lot of challenge. In this research the goal is to design a result prediction system for a T20 cricket match while the match is in progress. Di erent machine learning and statistical approach were taken to nd out the best possible outcome. A very popular data mining algorithm, decision tree were used in this research along with Multiple Linear Regression in order to make a comparison of the results found. These two model are very much popular in predictive modeling. Forecasting a T20 cricket match is a challenge as the momentum of the game can change drastically at any moment. As no such work has done regarding this for- mat of cricket, we have decided to take the challenge as T20 cricket matches are very much popular now a days. We are using decision tree algorithm to design our forecasting system by depending on the previous data of matches played between the teams. This system will help the teams to take major decision when the match is in progress such as when to send which batsman or which bowler to bowl in the middle overs. It significantly expands the exposure of research in sports analytics as it was previously bound between some other selected sports. Open Accesslistelement.badge.dso-type Item , Spatiotemporal analysis of air pollution using advanced machine learning techniques(BRAC University, 2026) Rahman, Shafin; Islam, Naeem; Rahman, Md. Shoaibur; Hridoy, Md. Moniruzzaman; Alam, Md. Ahasanul; Department of Computer Science and EngineeringThis thesis presents a unified framework for spatiotemporal analysis of air pollu- tion using advanced machine learning to enable short-horizon, citylevel forecasting and operational decision support. A leakage-safe, multi-source dataset is curated for 20 cities across Bangladesh and China, integrating pollutant observations with spatiotemporal covariates (e.g., meteorological and contextual signals) to model ur- ban pollution dynamics under heterogeneous conditions. The forecasting task is formulated as multi-output time-series regression over PM2.5, PM10, NO2, SO2, and CO. To capture short-term fluctuations and longer temporal dependencies while exploiting cross-pollutant structure, a multitask CNN–LSTM architecture is de- veloped with a shared feature backbone and pollutant-specific prediction heads. Performance is benchmarked against classical machine-learning baselines (including Random Forest and XGBoost) under cityaware evaluation to assess both accuracy and robustness. To address regional data imbalance, a cross-country transfer learn- ing strategy is evaluated by leveraging representations learned from data-rich source cities to improve forecasting in data-scarce target cities. Forecast reliability is en- hanced via Monte Carlo Dropout to estimate predictive uncertainty, while SHAP and Integrated Gradients provide complementary explanations of feature influence and temporal attribution. Finally, an early-warning episode detection layer converts forecasts into event-oriented alerts and diagnostics to support practical monitoring workflows. Overall, the proposed pipeline delivers more accurate, uncertainty-aware, and interpretable multi-pollutant forecasts suitable for risk-sensitive air quality man- agement in heterogeneous urban environments. Open Accesslistelement.badge.dso-type Item , Culturally adaptive neural network for detecting cybersecurity vulnerabilities in Bangladeshi web applications(BRAC University, 2026) Shaolin, Mohosina; Nawar, Fariha; Siddique, Arik Ahmed; Shams, Shaikh Mohammad Ali; Maliyat, Nafisa; Mostakim, Moin.; Department of Computer Science and EngineeringAs cyber threats become more complex and frequent, conventional methods for detecting website vulnerabilities, such as rule-based and heuristic approaches, faces significant difficulties, including limited adaptability, high rates of false positives, and a lack of contextual insight. This study presents a predictive model based on neural networks aimed to actively evaluating website security. By applying essential features like security headers, SSL/TLS settings, and SQL injection vulnerabilities, the model detects complex patterns and irregularities, enabling precise identification of emerging threats and vulnerabilities. This approach uses data-driven feature engineering and training with custom neural architectures, for comparison we used random forest and gradient boosting, For explainability we used SHAP followed by evaluation metrics such as precision, recall, and F1-score. Key results show improved accuracy, reduction of false positives, automated monitoring of configurations, and enhancement of resilience against adversarial attacks. Although neural networks show significant potential for transformation, challenges related to transparency, computational demands, and data imbalance are acknowledged. This highlights the necessity for ongoing learning, scalability, and integration with current frameworks, laying the groundwork for robust and adaptable web security strategies. Open Accesslistelement.badge.dso-type Item , An end-to-end framework for anomaly detection and categorization(BRAC University, 2025) Islam, MD. Farhan; Islam, Rehnuma; Reza, Syed Rahin; Tasnim, Saifa; Nipu, Anipa Akter; Rahman, Rafeed; Department of Computer Science and EngineeringIn this study, we proposed an end-to-end framework for anomaly detection, classification in Industry 4.0 using deep learning models YOLO V8 and ResNet on the MVTec Anomaly Detection(MVTec AD) dataset. The framework is based on defect detection, anomaly localization. The multitask queues in YOLO V8 guarantee both: fast and precise detection in real time, while ResNet primarily suited for classification, complete with top notch precision and recall metrics. The metrics used for evaluation (including AUC, accuracy, precision, recall, F1 score and AP) confirm the good performance of the models. We also provide decision surface visualizations through Grad-CAM and Integrated Gradients that will help you understand some of the decisions made by the model. The YOLO V8 performed optimal on real-time detection tasks and ResNet performed best on classification accuracy, as highlighted through the results. This framework allows for the automation of anomaly detection and the resolution through investigation, unlocking future opportunities for real time anomaly detection and management. Open Accesslistelement.badge.dso-type Item , A universal photography suggestion system utilizing composition detection, orientation detection, and subject position detection(BRAC University, 2025-06) Niloy, Iftikhar Shams; Proma, Syeda Mahjabin; Dofadar, Dibyo Fabian; Ahmed, Md. Sabbir; Department of Computer Science and EngineeringPhotography is one of the most popular hobby and images are one of the most important content types on social media, and the impact of a photo often hinges on its composition as much as its subject. In response to this, we proposed a system that classifies the compositional structure, detects orientation and subject of a given photo and suggests improvements based on established photography rules. For the classification of the composition, the photo will be categorized into one of five classes(CC, ROT, LL, FIF, PAT). Then, it will determine the orientation of an image. Lastly, this system uses YOLOv8 object detection model to find the objects of a photograph and through logics and conditions the subject is determined. The proposed system will provide the final suggestion based on the three results of the three proposed models. The main goal of the research is to develop a suggestion system that utilizes the detection models built using Deep Learning(DL) algorithms and find the optimal models that will accurately determine the composition, orientation and subject (if any) of a photograph. We have achieved up to 74.34% accuracy in our composition detection model and a minimum of 0.5870 mean square error (MSE) on our orientation detection model. The subject detection conditions capable of properly detecting the subject of an image most of the cases. Our approach aims to assist users in improving their photography skills and elevating the quality of visual content on any media platforms. Open Accesslistelement.badge.dso-type Item , Real-time aviation anomaly detection and multi-label classification using deep learning on multivariate sensor data(BRAC University, 2026) Yeasin, Sakib Rayhan; Rahat, Md. Atik Hasan; Nakib, Shafaat Jamil; Mitra, Debjoty; Alam, Md. Golam Rabiul; Reza, Md. Tanzim; Department of Computer Science and EngineeringGeneral aviation records a fatal accident rate of approximately one per 100,000 flight hours, with loss-of-control and stall events remaining leading preventable causes. Existing flight safety systems rely on fixed expert-defined thresholds and cannot detect complex multi-sensor anomaly patterns or identify specific event types in real time. This thesis proposes a lightweight two-stage deep learning framework for real-time aviation anomaly detection and multi-label event classification on raw flight sensor data from the NGAFID General Aviation Training Set. Stage 1 uses an ensemble of two novel Transformer architectures, the Cross-Sensor Patch Transformer (CSPT) and the Hierarchical Cross-Sensor Transformer (HiCST), each incorporating a cross-sensor multi-head attention module that explicitly models inter-sensor dependencies before temporal processing. Root Mean Square ensemble fusion achieves an anomaly-class F1-score of 0.8815, recall of 0.9076, and AUPRC of 0.9523 on a test set with a 339:1 class imbalance ratio. Stage 2 uses MHANet, which introduces per-sensor independent linear projections to classify each anomalous timestep into any combination of ten simultaneous event types, achieving a macro F1 of 0.9563 and subset accuracy of 0.9627. The complete pipeline runs in 6.08 milliseconds per sensor reading on a standard CPU, confirming real-time feasibility. Both stages outperform all established deep learning and classical machine learning baselines while using significantly fewer parameters, demonstrating that domain-aware architectural specialization consistently outperforms general-purpose approaches for aviation safety monitoring. Open Accesslistelement.badge.dso-type Item , WasteRefine: boundary-aware semantic segmentation of waste materials using a DINOv2 backbone with multi-scale feature fusion decoder(BRAC University, 2026-04) Khan, Talha Islam; Das, Trisha; Iqbal, Md. Ahnaf; Tawseef, Farhan; Alam, Md. Golam Rabiul; Datta, Nirjhor; Department of Computer Science and EngineeringThe rapid increase in world waste production needs smart, data-driven frameworks for efficient material identification and sustainable resource management. Intelligent recycling systems and waste materials spontaneous segmentation often lack behind due to scarcity of proper annotated datasets, visual ambiguities and severe class imbalancement of rare objects. The research aims to propose WasteRefine, utilizing DINOv2 Vision Transformer backbone with boundary aware semantic segmentation and multi scale feature fusion decoder for waste materials. To capture and accumulate the global context, an advanced dense predictive transformer is used consisting top-down fusion of features, Pyramid Pooling Module, Squeeze and Excitation channel attention and boundary composition component, for the proper identification of cluttered, deformed and visually ambiguous waste objects. The paper also introduces WasteRefine dataset consisting of 2,213 annotated images across four different categories: paper, soft plastic, rigid plastic and metal, marking it as the first waste semantic segmentation dataset from Bangladesh which contains visuals across various regions and annotated precisely. The proposed framework is rigorously evaluated on three different dataset WasteRefine, ZeroWaste-F and SpectralWaste (RGB) and assessed across notable published baselines. The ViT-B achieved 96.64 ± 0.16% mIoU on WasteRefine dataset, 61.94 ± 0.84% mIoU on extremely class imbalanced and deformed ZeroWaste-F dataset and 70.73 ± 0.10% FG mIoU on SpectralWaste beating all the published reports. Competitive results of the ViT-S variant with only 25.16M parameters demonstrated efficient parameter count without severe performance degradation. Open Accesslistelement.badge.dso-type Item , An efficient technique for real-time transformation of 2D to 3D images with GPU using CUDA programming(BRAC University, 2026-02) Mahmud, Sadat; Mustafa, Md. Rana; Mitra, Ananda; Shanto, Sajjad Hossain; Chowdhury, Mohammad Nazibul Bashar; Alam, Md. Ashraful; Department of Computer Science and EngineeringThis thesis presents a complete 2D to 3D reconstruction system designed to run reliably on a low computational powered PC, where GPU memory, host memory, and disk bandwidth impose strict constraints. The pipeline begins with large-scale synthetic data generation from ShapeNet models, producing aligned RGB and depth observations for supervised learning. A ResUNet18 based monocular depth network is trained in LibTorch using a mask-aware objective to promote numerical stability and reduce invalid-depth regions in the predicted maps. To ensure continuous training without data starvation under limited resources, the system is implemented a producer consumer scheduling system design: a producer renders and stages batches to fast local storage, consumers stream and pre-process shards into the training loop, and a destroyer reclaims storage deterministically once a batch is fully consumed. This design bounds disk usage, prevents host RAM accumulation, and decouples rendering from training so the GPU remains saturated even when CPU-side work fluctuates. After inference, predicted camera-centric depth is lifted into explicit 3D geometry using CUDA-accelerated reconstruction, enabling dense point cloud and grid-mesh generation at image resolution with minimal overhead. The system is evaluated using runtime traces (GPU utilization, GPU/host memory, and CPU load) alongside standard depth estimation metrics aggregated across training batches, demonstrating sustained execution, stable memory behavior, and reconstruction-ready depth quality on resource-constrained hardware. Open Accesslistelement.badge.dso-type Item , TRACER: task-aware risk-adaptive architecture for continual edge learning(BRAC University, 2026-02) Islam, Md.Hasibul; Fuad, Mir Muhammad; Tanzin, A.B.M. Fahim Hasan; Rahman, Md. Khalilur; Department of Computer Science and EngineeringCurrent computer-vision architectures are being deployed as long-lived services, especially on edge and on-device platforms, where input distributions change with changes in environment, users, sensors, and class frequencies. In these cases, to achieve sustainable performance, continual learning is required. Also, we need to keep in mind that the process needs to be feasible under strict constraints like latency and memory. Previous experience demonstrates that device-centric measures of deployment efficiency should be used instead of proxy metrics like FLOPs, and that tail latency (e.g., p95) is a more constrained measure of deployment efficiency than mean latency. At the same time, full neural architecture search (NAS) is generally too costly to integrate into a repeated learning loop, motivating restricted, hardware-aware search strategies. This thesis presents TRACER: Task-aware Risk-adaptive Architecture for Continual Edge leaRning, a deployability-oriented continual learning pipeline that keeps a fixed feature backbone and repeatedly selects and adapts a lightweight MLP classifier head. The system follows a restricted design space with a NAS-inspired controller and a Net2Net-optimized evolutionary population so that it can adapt efficiently. The stability between tasks is ensured through risk-aware exemplar rehearsal (high-risk samples are prioritized) and knowledge distillation. Experiments on Split CIFAR-100 (10 tasks x 10 classes) and CIFAR-10 (5 tasks x 2 classes) report class-incremental (CIL) and task-incremental (Task-IL) performance. Our proof-ofconcept implementation has a final CIL mean accuracy of 0.8145 and average forgetting of 0.0341, and TIL has a final mean accuracy of 0.970 on CIFAR-10. Also, in CIFAR-100, we got 0.6136 final CIL mean accuracy, and TIL has a final mean accuracy of 0.9055 with 0.0451 forgetting. These results, which are derived from a single deterministic run, demonstrate that Lagrangian-relaxation-based constraintaware head selection, combined with risk-sensitive stabilization, provides a practical accuracy-feasibility trade-off for continual learning under explicit latency targets. Open Accesslistelement.badge.dso-type Item , PAMM: pathway-aware masked representation learning for interpretable multi-cancer prediction(BRAC University, 2026-01) Chowdhury, Chandrima Roy; Rodoshi, Zarrin Tasnim; Surovi, Sumaiya Hossain; Hasan, Labib; Chakrabarty, Amitabha; Department of Computer Science and EngineeringIn this thesis, PAMM, a new paradigm of interpretable multi-cancer prediction based on Pathway-Aware Masked Representation Learning is introduced. To tackle the challenge of the ‘Small n, Large p’ of transcriptomics it is our holding that we apply the rigorous seven-stage pipeline of preprocessing (i.e. Log2 transform, ANOVA filter, Lasso regularization and Recursive Feature Elimination) to reduce the original high-noise 57,750 genes in Breast, Lung, GBM, and HC samples to a high-signal feature set. The basic architecture goes beyond the usual deep learning of black boxes by incorporating biologically relevant priors of KEGG 2021 Human library in a self-supervised masking scheme. In contrast to stochastic masking, the pretraining phase of PAMM uses a Pathway-Aware Masking logic where complete sets of functional genes are zeroed, requiring the model to recreate missing biological units and learn complicated inter-pathway relationships. The latent representations of the model are optimized with Optuna, and the statistical robustness is verified with twenty independent iterations, and the latent representation is further interpreted with Single-sample Gene Set Enrichment Analysis (ssGSEA). The resulting visualizations of mean pathway activity indicate that PAMM is able to capture different, clinically viable biological signatures of each cancer type. PAMM provides a clear and very precise diagnostics platform of precision oncology by filling the gap between high-dimensional self-supervised learning and functional biology. Along with closed-set multi-cancer, PAMM is also explicitly tailored to open-set recognition. Through a combined study of softmax confidence and latent space distances from class centroids, the framework can discard samples that do not adhere to any known cancer manifold. This allows the certainty of identifying unknown or non-cancerous gene expression patterns, which is very essential when it comes to a real-life clinical implementation in which unobservable conditions are the norm. This two-fold feature sets PAMM apart from the traditional classifiers and guarantees the accuracy of the diagnosis and its safety Open Accesslistelement.badge.dso-type Item , ProtReason: a reasoning-based framework for interpretable protein function prediction(BRAC University, 2025-06) Ayon, Sartiz Alam; Orin, Alvi Sakib; Biswas, Arpon; Fahad Al Shahid; Shahriyer, Shaikh Faiyaz; Sadeque, Farig Yousuf; Department of Computer Science and EngineeringUnderstanding how protein sequence determines function remains a central challenge in computational biology. While some protein language models have advanced function prediction but most of them produce outputs without any justification or explainability. Protein function can be justified by connecting biological evidence to functional conclusions. We present ProtReason: A reasoning-augmented framework that generates interpretable protein function predictions with structured reasoning traces. In this study, a curated dataset of 87K proteins is constructed which is enriched with protein domain motifs, localization predictions and structural features transformed into reasoning traces linked to functional labels. ProtReason employs a two-stage architecture that first aligns protein sequence embeddings with textual representations and then generates structured outputs including reasoning traces, functional descriptions, and confidence scores. Compared to a sequence-tofunction baseline without reasoning, ProtReason achieves significantly improved BERT F1 scores, demonstrating the benefit of incorporating reasoning prior to function prediction. A systematic ablation study with 16 model variants shows the best design principles: a single unified reasoning path is better than a multi-step chain of reasoning and generating reasoning before function prediction yields superior performance. ProtReason performs competitively on standard benchmarks while providing biologically interpretable explanations with calibrated confidence estimates. Open Accesslistelement.badge.dso-type Item , Efficacy of multiple curriculum-based large language models in Bangladesh’s education system(BRAC University, 2026) Rahman, Aumio; Mazumdar, Tanjila; Suzana, Anika Afsara; Sadeque, Farig Yousuf; Department of Computer Science and EngineeringIn Bangladesh, the education system follows multiple curricula, each o”ering di”erent perspectives on culture and society. This research aims to evaluate the e”ectiveness of large language models (LLM) to identify biases in textbooks across di”erent curricula in Bangladesh. As the various educational systems di”er, there is a high possibility that students from di”erent backgrounds may have diverse perspectives on culture and society. Henceforth, this study investigates the potential influences embedded in the curriculum that can a”ect young minds. Our goal is to achieve a comprehensive analysis of these biases, expecting insights that could inform curriculum development and promote balanced educational content in Bangladesh’s diverse educational streams. Open Accesslistelement.badge.dso-type Item , Automated hazard detection for AR/VR Mars terrain navigation using computer vision(BRAC University, 2026-01) Rhidy, Tasin Ahsan; Rahman, MD Touhidur; Shuvo, Istiak Zaman; Mahmood, Saiyed Mubasshir; Rafi, Abrar Mojahid; Alam, Md. Ashraful; Alam, Md. Golam Rabiul; Tasnim, Sanjida; Department of Computer Science and EngineeringThe exploration of Mars brings about a series of challenges that are occasioned by the risky topographical features, random weather patterns, and the fundamental need to have self-driving equipment. The study builds a combined computer vision and immersive technology system to improve the safety of the human astronauts and robot rovers in their navigation on the surfaces of the Martian environment. Our solution is a multi-modal deep-learning system consisting of object detection, semantic segmentation, and monocular depth estimation to generate complete hazard awareness in simulated Mars environments. We use datasets to train terrain classification models that are able to detect important surface features such as rocks, boulders, and potholes as well as other geological features. The system combines a number of deep-learning networks to detect hazards in real-time and locate bounding-boxes, semantic-segmentation, and pixel-level terrain-classification as well as a depth-estimation architecture to give the system spatial information of the Martian terrain. These models are synergistically used to produce an environmental cognition that drives into an AR/VR interface that provides users with visual cues in safe path planning. The AR/VR element converts raw computer-vision data into usable navigation data, and deciphers warnings of hazards and terrain complexity data to the Martian landscape. The initial studies have shown strong detection of varied terrain conditions, and the multi-modal strategy has a great benefit on improving the safety of navigation in comparison to the single-modality systems. The study has been applied to the development of autonomous planetary exploration technologies and created a scalable model of pre-mission astronaut training and rover operation plan. Open Accesslistelement.badge.dso-type Item , Temporal state-aware unsupervised anomaly detection for industrial control system(BRAC University, 2026-01) Anik, Khandoker Wahiduzzaman; Ontu, Md. Rakib Hossain; Sohag, Md. Mehedi Hasan; Badhon, Fardin Jahan; Chakrabarty, Amitabha; Department of Computer Science and EngineeringThe increasing interrelationship with Information Technology infrastructure between the Industrial Control Systems (ICS) and critical infrastructure has presented advanced cyber-attacks to critical infrastructure, which not only places data security at risk but also threatens the physical safety, to operational continuity. The thesis will visit the issue of creating an efficient, interpretable and computationally efficient intrusion detection system in an ICS environment through a proposed novel LSTM auto-encoder architecture, specifically trained with edge deployment in mind. The article takes a rigorous approach where physics-conscious feature engineering is embraced, deep-learning architecture creation, and thorough assessment of the WADI (Water Distribution) benchmark data. The proposed system achieves a score of 0.7018 in F1 (Precision=0.7196, Recall=0.7149) by performing the dimensionality reduction of 127 sensors to 30 (which is a reduction of 76 per cent), and by adding the zero-crossing-rate features to the frequency-domain analysis, which is drastically higher than more traditional statistical methods, including Isolation Forest (0.58 F1), or the current state-of-the-art methods, including STADN. To be practical, the system is edge-compatible with an inference latency of 1.84ms, a million parameters, and consumes 5.38W of power when run on simulated NVIDIA Jetson Nano hardware, which is a ten-fold faster inference time than graph-based algorithms. Unsupervised approach, which learns only based on normal operational data, helps to detect novel, zero-day attacks, therefore overcoming the limitation of labelled attack data in operational settings. The study establishes that advanced deep-learning systems can be deployed on the tight computational requirements of industrial edge devices, thus creating a reproducible model of secure and real-time secure critical infrastructure protection. Open Accesslistelement.badge.dso-type Item , Integrating single sign-on within the WebAuthn framework(BRAC University, 2026-01) Adnan, Asir; Anika, Nafisha Tabassum; Sobahan, Saima; Istiaque, A.J.M; Ferdous, Md Sadek; Department of Computer Science and EngineeringIn the world of digital identity, preserving user privacy while maintaining seamless access across platforms has become a challenge. WebAuthn, developed by World Wide Web Consortium (W3C) is mainly a web-based authentication standard. This system enhances security by enabling passwordless login through hardware-based and biometric authentication mechanisms. Another popular approach, to simplify authentication for users across the internet is Single-Sign-On (SSO) which allows a single credential to access multiple services or applications. This way users can get rid of the liability to manage multiple credentials, rather they can rely on only one credential to authenticate in a trusted manner and use that to authenticate in many other websites. Despite the potential of the SSO system, it has not been integrated with the WebAuthn framework till date. Through our research work, we have introduced a system that ensures passwordless authentication via WebAuthn and supports seamless access to service providers through SSO eliminating the requirements of repeated login. Moreover, this system empowers users with the full control over sharing their personal information by selective disclosure mechanism. Security Assertion Markup language (SAML) is used as the federated identity to exchange the authentication assertion securely between identity providers and service providers to enable seamless SSO. Thereby, introducing a new horizon of research on WebAuthn and SSO. Open Accesslistelement.badge.dso-type Item , Using machine learning to predict optimal erasure coding policies for object storage system in OpenStack Swift(BRAC University, 2026-01) Ankon, Amio Malakar; Chaki, Boloy; Sayan, Mashrur Shakhawat; Sarker, Debashish; Mukta, Jannatun Noor; Department of Computer Science and EngineeringErasure coding helps to reduce storage overhead and improve fault tolerance. But the procedure to select an appropriate erasure coding policy is complex, which often involves tradeoffs among various metrics such as access latency, recovery behavior, storage efficiency, etc. Generally, these are handled using static or heuristic-based configurations, which can not account for variations in workload. In the industry, service providers like Ceph do benchmarking based on throughput/latency without considering workload diversity. OpenStack Swift, one of the most widely used open-source object storage systems, supports erasure coding, but it allocates policy selection in a manual, static way - without it being workload-aware. To address this problem, this thesis showcases a data-driven performance modeling framework for erasure coding in object storage systems using machine learning. A structured dataset- ABDS-30k was constructed in a controlled execution of varying workload conditions. It was created on a Swift All-In-One (SAIO) testbed under different erasure coding policies with the addition of failure injections. The dataset collects several empirically observed performance metrics such as read and write latency, tail latency, success rate, and reconstruction time in case of disk failures. For the job of selecting the optimal erasure coding policy, we adopt two distinct Machine Learning paradigms - a regression-based approach of performance modeling and a classification-based approach for direct data-driven policy recommendation. For regression, CatBoost, XGBoost, and Random Forest Regression are used to predict target metrics in a given workload context and failure scenario, and the optimal policy is chosen based on a weighted score. In the classification-based approach, CatBoostClassifier, XGBoostClassifier, and Logistic Regression are used to label workload–policy pairs using an oracle cost function, and the most optimal erasure coding policy is directly recommended. Experimental results show that the regression models have moderate prediction errors for latency and recovery metrics because of the inherently noisy and heavy-tailed nature of the system, but they remain effective in optimal policy recommendation by exhibiting top-1 accuracy with 48.16% in XGBoost Regression and a top-3 accuracy 100% in Random Forest Regression model. Also, the regret mean (0.010 - 1.81) and regret median (0.00216 - 0.089) values showed a very low margin of error. Classification-based approach shows comparatively weaker metrics, indicating that it is not optimal for data-driven policy recommendation. Overall, this shows the feasibility of using ML-based performance modeling as opposed to static and heuristic-based policy selection, which can be later used as a foundation for future control-plane automations.