Search Articles & Publications

Showing 182 articles found for "Four"

FORENSIC ANALYSIS OF MITM ATTACK ON ‘AISYIYAH UNIVERSITY YOGYAKARTA NETWORK USING NIST METHOD

Ridwan, Virgiawan aqil, Firdonsyah, Arizona
Abstract: Abstract: Man-in-the-Middle (MITM) attacks are a threat that can occur on public wireless networks, including campus Wi-Fi environments. This study aims to analyze MITM attacks on the Wi-Fi network at Universitas ‘Aisyiyah… iyah Yogyakarta using the National Institute of Standards and Technology (NIST) digital forensics methodology. The study applied the four NIST phases: collection, examination, analysis, and reporting. The digital evidence analyzed included packet capture (PCAP) files, as well as digital traces such as browser history, cookies, and cache data obtained from the victim’s device. The analysis process utilized Wireshark, the SQLite Database Browser, and ChromeCacheView to identify suspicious activity and correlate the discovered digital traces. The results of the study show that the MITM attack was successfully reconstructed through the correlation of digital traces, leading to the identification of ARP spoofing and DNS spoofing originating from a device with the IP address 192.168.200.12 and the MAC address a0:47:d7:73:ef:fb. The correlation of digital traces in the victim’s network and system traffic revealed communication redirection and web access manipulation. This study concludes that the NIST method is capable of reconstructing MITM attacks and identifying digital evidence from activity traces on both the network and the system.             Keywords: ARP spoofing; digital forensics; DNS spoofing; MITM; NIST     Abstrak: Serangan Man-in-the-Middle (MITM) merupakan ancaman yang dapat terjadi pada jaringan nirkabel publik, termasuk lingkungan WiFi kampus. Penelitian ini bertujuan menganalisis serangan MITM pada jaringan WiFi Universitas ‘Aisyiyah Yogyakarta menggunakan metode forensik digital National Institute of Standards and Technology (NIST). Penelitian menerapkan empat tahapan NIST, yaitu collection, examination, analysis, dan reporting. Bukti digital yang dianalisis meliputi file packet capture (PCAP), jejak digital berupa history browser, cookies, dan cache yang diperoleh dari perangkat korban. Proses analisis menggunakan Wireshark, SQLite Database Browser, dan ChromeCacheView untuk mengidentifikasi aktivitas mencurigakan serta mengorelasikan jejak digital yang ditemukan. Hasil penelitian menunjukkan bahwa serangan MITM berhasil direkonstruksi melalui korelasi jejak digital yang mengarah pada identifikasi ARP spoofing dan DNS spoofing dari perangkat dengan alamat IP 192.168.200.12 dan MAC address a0:47:d7:73:ef:fb. Korelasi jejak digital pada lalu lintas jaringan dan sistem korban menunjukkan adanya pengalihan komunikasi serta manipulasi akses web. Penelitian ini menyimpulkan bahwa metode NIST mampu merekonstruksi serangan MITM dan mengidentifikasi bukti digital dari jejak aktivitas pada jaringan maupun sistem.   Kata kunci: ARP spoofing; DNS spoofing; forensik digital; MITM; NIST

TOPSIS-BASED SYSTEM FOR THE SELECTION OF TRAINING PARTICIPANT CANDIDATES AT THE ASAHAN MANPOWER OFFICE

Maha Putra, Guntur, Wan Mariatul Kifti, Putri Amanda Nurhayati
Abstract: Abstract: Job training is one of the government’s efforts to improve the quality of human resources so that they possess competencies that meet labor market demands. The process of selecting training participants at the… e Department of Manpower of Asahan Regency is still carried out manually, which can lead to subjectivity and inefficiency in determining the most eligible candidates. This study aims to develop a decision support system using the Technique for Order Preference by Similarity to Ideal Solution (TOPSIS) method to assist the selection process objectively and systematically. The study applies four evaluation criteria, namely education level, age, work experience, and interview, with a dataset consisting of 31 training candidates. The system is developed as a web-based application using PHP programming language and MySQL database. The TOPSIS method is applied through decision matrix normalization, weighting, determination of positive and negative ideal solutions, and preference value calculation to produce a ranking of candidates. The results show that the proposed system can provide objective recommendations for selecting training participants, improve the efficiency of the selection process, and support decision makers in producing more accurate and reliable decisions. Keywords: decision support system; selection; training; TOPSIS.   Abstrak: Pelatihan tenaga kerja merupakan salah satu upaya pemerintah dalam meningkatkan kualitas sumber daya manusia agar memiliki kompetensi yang sesuai dengan kebutuhan dunia kerja. Proses pemilihan calon peserta pelatihan di Dinas Tenaga Kerja Kabupaten Asahan selama ini masih dilakukan secara manual sehingga berpotensi menimbulkan subjektivitas dan kurang efektif dalam menentukan peserta yang paling layak. Penelitian ini bertujuan untuk membangun sistem pendukung keputusan menggunakan metode Technique for Order Preference by Similarity to Ideal Solution (TOPSIS) untuk membantu proses seleksi peserta pelatihan secara objektif dan sistematis. Penelitian ini menggunakan empat kriteria penilaian yaitu pendidikan, usia, pengalaman kerja, dan wawancara dengan jumlah data sebanyak 31 calon peserta pelatihan. Sistem dikembangkan berbasis web menggunakan bahasa pemrograman PHP dan database MySQL. Metode TOPSIS digunakan untuk melakukan normalisasi matriks keputusan, pembobotan, penentuan solusi ideal positif dan negatif, serta perhitungan nilai preferensi untuk menghasilkan perankingan peserta pelatihan. Hasil penelitian menunjukkan bahwa sistem yang dibangun mampu memberikan rekomendasi peserta pelatihan secara objektif, meningkatkan efisiensi proses seleksi, serta membantu pihak dinas dalam pengambilan keputusan yang lebih akurat. Kata kunci: pelatihan; seleksi; sistem pendukung keputusan; TOPSIS.

DIGITAL FORENSIC INVESTIGATION ON STORAGE MEDIA BASED ON NIST WITH FORENSIC PROCESS METHODS

Gunawan, Indra, Satria Tambunan, Heru, Ahmad, Abdullah
Abstract: Abstract: Storage media is an inseparable tool in everyday life. With storage media, users can store important data, both personal and workplace. In addition, in many cases, Indonesian law uses storage media as evidence.… The Electronic Information and Transactions Law (UU ITE) regulates how the provision of digital evidence can be strong evidence in court. This study examines the forensics of digital evidence on storage media with four test scenarios. Digital forensic processing uses forensic processes based on the National Institute of Standards and Technology (NIST) guidelines. This study produces an analysis in which evidence processed with scenarios 1 and 4 is valid digital evidence to be submitted to court, while evidence 2 and 3 is invalid evidence. The results of this digital evidence can be used for investigations under the ITE law.   Keywords: autopssy; digital forensics; storage media; FTK Imager.     Abstrak: Media Penyimpanan merupakan alat yang tak terpisahkan dari kehidupan sehari-hari. Dengan Media Penyimpanan, pengguna dapat menyimpan data penting, baik pribadi maupun tempat kerja. Selain itu, dalam banyak kasus, hukum Indonesia menggunakan Media Penyimpanan sebagai alat bukti. Undang-Undang Informasi dan Transaksi Elektronik (UU ITE) mengatur bagaimana penyediaan alat bukti digital menjadi alat bukti yang kuat di pengadilan. Penelitian ini mengkaji forensik terhadap alat bukti digital pada Media Penyimpanan dengan empat skenario pengujian. Pemrosesan forensik digital menggunakan proses forensik berdasarkan panduan National Institute of Standards and Technology (NIST). Penelitian ini menghasilkan analisis di mana alat bukti yang diproses dengan skenario 1 dan 4 merupakan alat bukti digital yang sah untuk diajukan ke pengadilan, sedangkan alat bukti 2 dan 3 merupakan alat bukti yang tidak sah. Hasil dari barang bukti digital ini, dapat digunakan untuk penyelidikan didalam undang-undang ITE.   Kata kunci: otopsi; forensik digital; media penyimpanan; FTK Imager

OPTIMIZING CYBER ATTACK SIMULATION AS A RESPONSE TO ESCALATING SECURITY THREATS USING A MACHINE LEARNING APPROACH

Lubis, Rivaldi, Halim, Apriyanto, Tanjaya, Felix Jansen, Tandri
Abstract: Abstract: The growing intensity of cyber attacks, marked by rapid, large-scale, automated, and adaptive execution, requires analytical methods that represent the diversity of network environments, including variations in… target platforms such as IoT, traditional networks, and hybrid infrastructures. This study compares machine learning models for cyber attack classification under heterogeneous environmental conditions and formulates a conceptual optimization framework based on model performance. Four publicly available benchmark datasets were used, namely UNB CIC IoT 2023, UNB CIC IDS-2018, UNSW-NB15, and a Kaggle cyber security attacks dataset, comprising approximately 40,000 to over 3.6 million records and 25 to 80 features across IoT, conventional, and mixed network environments. Random Forest, XGBoost, Multilayer Perceptron, and Transformer were implemented within a unified pipeline involving preprocessing, feature selection, and Bayesian Optimization-based hyperparameter tuning. All models achieved F1-score and Cohen's Kappa above 96%, with XGBoost performing best (97.80%, 97.26%), followed by Random Forest (97.78%, 96.96%) and Transformer (97.44%, 96.82%), while MLP scored lowest (96.74%, 96.00%), a gap below one percentage point. Confusion matrix analysis revealed persistent misclassification in minority and overlapping attack classes, informing a proposed adaptive cyber attack simulation optimization framework.             Keywords: cyber attacks; optimization; machine learning; environmental variability.     Abstrak: Meningkatnya intensitas serangan siber yang berlangsung cepat, masif, otomatis, dan adaptif menuntut pendekatan analitis yang merepresentasikan keragaman lingkungan jaringan, termasuk perbedaan karakteristik platform sasaran seperti Internet of Things (IoT), jaringan konvensional, dan infrastruktur hibrida. Penelitian ini membandingkan model machine learning untuk klasifikasi serangan siber pada kondisi lingkungan heterogen, sekaligus menyusun kerangka optimasi konseptual berdasarkan performa model. Empat dataset benchmark publik digunakan, yaitu UNB CIC IoT 2023, UNB CIC IDS-2018, UNSW-NB15, serta dataset Kaggle cyber security attacks, dengan jumlah data berkisar 40.000 hingga lebih dari 3,6 juta rekaman dan 25 sampai 80 fitur, mewakili lingkungan IoT, konvensional, dan campuran. Random Forest, XGBoost, Multilayer Perceptron, dan Transformer diimplementasikan melalui pipeline terpadu mencakup pra-pemrosesan, seleksi fitur, dan optimasi hyperparameter berbasis Bayesian Optimization. Seluruh model mencapai F1-score dan Cohen's Kappa di atas 96%, dengan XGBoost menunjukkan performa terbaik (97,80%, 97,26%), diikuti Random Forest (97,78%, 96,96%) dan Transformer (97,44%, 96,82%), sementara MLP mencatat skor terendah (96,74%, 96,00%), dengan selisih kurang dari satu poin persentase. Analisis confusion matrix mengungkap misklasifikasi yang konsisten pada kelas minoritas dan serangan dengan karakteristik serupa, yang menjadi dasar kerangka optimasi simulasi serangan siber adaptif yang diusulkan.   Kata kunci: serangan siber; optimasi; machine learning; variabilitas lingkungan

ANALYSING STUDENT MENTAL HEALTH THROUGH K-MEANS CLUSTERING AND MULTI-STAGE SAMPLING METHODS

Rahmat Hidayat, Dede Pratama
Abstract: Abstract: Mental health is an essential aspect of overall well-being, particularly for university students vulnerable to emotional strain. This study aims to identify clusters of student mental health trends using the K-Means… Means clustering technique. The research involved 60 students from four academic programs at the Faculty of Science and Technology, selected using stratified and cluster sampling techniques. Data were collected using a modified Mental Health Inventory (MHI). The results revealed distinct commonalities among majors: the Statistics program was predominantly defined by the depressed cluster at 53.3%, while Mathematics followed at 40% within the same cluster. In contrast, Biology students predominantly fell under the neu-tral/stable cluster (66.7%), whilst Information Systems students exhibited an even distribution (33.3% per cluster) without a dominant trend. The clustering quality was evaluated using the Silhouette Coefficient, yielding a range of 0.39 to 0.60. Biology (0.60) and Statistics (0.54) exhibited a reasonable structure, but Information Systems (0.39) and Mathematics (0.34) demonstrated a deficient structure. In conclusion, K-Means effectively discerns mental health patterns, providing a data-driven basis for targeted psychological interventions in educational settings. Keywords: biology; information systems; k-means; mathematics; mental health; silhouette coefficient; statistics   Abstrak: Kesehatan mental merupakan komponen vital dari kesejahteraan total, terutama bagi maha-siswa yang rentan terhadap stres emosional. Penelitian ini bertujuan untuk mengidentifikasi kelompok tren kesehatan mental mahasiswa melalui penerapan metode pengelompokan K-Means. Studi ini mencakup 60 mahasiswa dari empat program studi di Fakultas Sains dan Teknologi, yang dipilih melalui metode pengambilan sampel bertingkat dan kelompok. Data dikumpulkan dengan menggunakan Inventaris Kesehatan Mental (MHI) yang dimodifikasi. Temuan menunjukkan kesamaan yang jelas di antara jurusan: program studi Statistika terutama ditandai oleh kelompok depresi (53,3%), diikuti oleh Matematika dengan 40% dalam kelompok depresi. Sebaliknya, mahasiswa Biologi terutama termasuk dalam kelompok netral/stabil (66,7%), sedangkan mahasiswa Sistem Informasi memiliki distribusi yang merata (33,3% per kelompok) tanpa pola yang dominan. Kualitas pengelompokan dinilai dengan Koefisien Sil-houette, menghasilkan rentang 0,39 hingga 0,60. Biologi (0,60) dan Statistika (0,54) memiliki struktur sedang, sedangkan Sistem Informasi (0,39) dan Matematika (0,34) menunjukkan struktur yang buruk. Kesimpulannya, K-Means secara akurat mengidentifikasi tren kesehatan mental, menawarkan landasan berbasis data untuk terapi psikologis yang ditargetkan di ling-kungan pendidikan. Kata kunci: biologi; kesehatan mental; K-Means; matematika; silhouette coefficient; sistem in-formasi; statistika

OPTIMIZING RETRIEVAL-AUGMENTED GENERATION FOR DOMAIN-SPECIFIC KNOWLEDGE SYSTEMS THROUGH FINE-TUNING AND PROMPT ENGINEERING

Ahmad Fajri, Rila Mandala
Abstract: Abstract: This study discusses the optimization of RAG for a FAQ system in the field of information technology product security certification at BSSN. Although LLM generate reliable responses, they often lack up-to-date… and domain-specific knowledge, which can be addressed through the RAG approach. This research aims to optimize a domain-specific RAG system by improving embedding performance, enhancing prompt robustness, and increasing retrieval accuracy. The research methods consist of three stages. The first stage involves fine-tuning the bge-m3 embedding model and evaluating its performance using MRR, Recall, and AUC. The second stage applies prompt engineering techniques, namely the SRSM and Autodefense, to mitigate direct-injection and escape-character prompt injection attacks. The third stage evaluates the proposed RAG system using Precision, Recall, and F1-Score metrics against four baseline models. The results of research show that the fine-tuned embedding model achieves higher performance than the original model, with MRR@1 and Recall@1 values of 0.80 and an AUC@100 of 0.7023. In addition, the proposed prompt engineering techniques demonstrate robustness against prompt injection attacks, while the overall RAG system attains a perfect Precision, Recall, and F1-Score of 1.00. In conclusion, the proposed approach effectively enhances retrieval accuracy, embedding quality, and system security, resulting in a more reliable RAG-based FAQ system for information technology product security certification. Keywords: embedding fine-tuning; large language model; prompt engineering; prompt injection mitigation; retrieval-augmented generation   Abstrak: Studi ini membahas optimasi RAG untuk sistem FAQ di bidang sertifikasi keamanan produk teknologi informasi di BSSN. Meskipun LLM menghasilkan respons yang andal, mereka seringkali kurang memiliki pengetahuan terkini dan spesifik domain, yang dapat diatasi melalui pendekatan RAG. Penelitian ini bertujuan untuk mengoptimalkan sistem RAG spesifik domain dengan meningkatkan kinerja embedding, meningkatkan ketahanan prompt dan meningkatkan akurasi pengambilan. Metode penelitian terdiri dari tiga tahap. Tahap pertama melibatkan fine-tuning model embedding bge-m3 dan mengevaluasi kinerjanya menggunakan Mean Reciprocal Rank (MRR), Recall, dan AUC. Tahap kedua menerapkan teknik rekayasa prompt, yaitu Self- SRSM dan Autodefense, untuk mengurangi serangan direct-injection dan escape-character prompt injection. Tahap ketiga mengevaluasi sistem RAG yang diusulkan menggunakan metrik Presisi, Recall, dan F1-Score terhadap empat model dasar. Hasil penelitian menunjukkan bahwa model embedding yang disempurnakan mencapai kinerja yang lebih tinggi daripada model asli, dengan nilai MRR@1 dan Recall@1 sebesar 0,80 dan AUC@100 sebesar 0,7023. Selain itu, teknik rekayasa prompt yang diusulkan menunjukkan ketahanan terhadap serangan injeksi prompt, sementara sistem RAG secara keseluruhan mencapai Presisi, Recall, dan F1-Score sempurna sebesar 1,00. Kesimpulannya, pendekatan yang diusulkan secara efektif meningkatkan akurasi pengambilan, kualitas embedding dan keamanan sistem, menghasilkan sistem FAQ berbasis RAG yang lebih andal untuk sertifikasi keamanan produk teknologi informasi. Kata kunci: penyempurnaan embedding; model bahasa besar; rekayasa prompt; mitigasi injeksi prompt; retrieval-augmented generation

A FUZZY LOGIC BASED EVALUATION MODEL FOR THESIS TOPIC FEASIBILITY TO ENHANCE STUDENT RESEARCH RELEVANCE

Rizaldi, Dewi Anggraeni, Elly Rahayu
Abstract: Abstract: The determination of thesis topics is a fundamental stage in academic research, yet the evaluation process remains predominantly manual and subjective. This reliance on individual lecturer perception often leads… s to inconsistent feasibility assessments and fails to systematically measure the topic's alignment with strategic needs. This research aims to develop a Decision Support System (DSS) model based on fuzzy logic to assess the feasibility of thesis topics objectively and systematically, focusing on enhancing the relevance of student research. The research method employed the Fuzzy Inference System (FIS) with the Sugeno method. This model was designed through literature review and FGD to establish four criteria (Topic Relevance, Difficulty Level, Idea Novelty, Reference Availability) and 81 rule bases. The model validation results against expert judgment using 15 test data showed a high accuracy rate of 91.31%, with a Mean Absolute Percentage Error (MAPE) value of 8.69%. In conclusion, this DSS model is proven to be valid and consistent, and it can be relied upon as an objective tool to improve the quality and relevance of thesis topics. Keywords: academic evaluation; decision support system; fuzzy logic; fuzzy sugeno; thesis feasibility   Abstrak: Penentuan topik skripsi merupakan tahapan fundamental dalam penelitian akademik, namun proses evaluasinya hingga kini masih cenderung manual dan subjektif. Ketergantungan pada persepsi dosen secara individu sering kali menyebabkan penilaian kelayakan yang tidak konsisten serta kegagalan dalam mengukur keselarasan topik dengan kebutuhan strategis secara sistematis. Penelitian ini bertujuan mengembangkan model Sistem Pendukung Keputusan (SPK) berbasis logika fuzzy untuk menilai kelayakan topik skripsi secara objektif dan sistematis, dengan fokus pada peningkatan relevansi penelitian mahasiswa. Metode penelitian yang digunakan adalah Fuzzy Inference System (FIS) dengan metode Sugeno. Model ini dirancang melalui tinjauan pustaka dan Focus Group Discussion (FGD) untuk menetapkan empat kriteria (Relevansi Topik, Tingkat Kesulitan, Kebaruan Ide, Ketersediaan Referensi) serta 81 basis aturan. Hasil validasi model terhadap penilaian pakar menggunakan 15 data uji menunjukkan tingkat akurasi yang tinggi yaitu 91,31%, dengan nilai Mean Absolute Percentage Error (MAPE) sebesar 8,69%. Kesimpulannya, model SPK ini terbukti valid dan konsisten, serta dapat diandalkan sebagai alat objektif untuk meningkatkan kualitas dan relevansi topik skripsi. Kata kunci: evaluasi akademik; sistem pendukung keputusan; logika fuzzy; fuzzy sugeno; kelayakan skripsi

COMPARISON OF NAÏVE BAYES, SVM, K-NN, DECISION TREE, AND RANDOM FOREST IN SENTIMENT ANALYSIS BASED ON SEABANK APPLICATION ASPECTS

Fachrozi, Muhammad Al, Tania, Ken Ditha
Abstract: Abstract: The increasing use of digital banking applications has led to the need for a deeper understanding of user perceptions, especially through aspect-based sentiment analysis. This study aims to classify the sentiment… nt of SeaBank app users by focusing on four main aspects: learnability, efficiency, technical issues or errors, and satisfaction. Review data totaling 1,971 comments were collected from the Google Play Store and labeled with sentiments based on the scores (ratings) given by users. The CRISP-DM approach serves as the methodological framework for this study, which includes five classification algorithms: Naïve Bayes, Support Vector Machine (SVM), k-Nearest Neighbor (k-NN), Decision Tree, and Random Forest. The evaluation results show that the SVM algorithm provides the best performance with the highest average value of the four aspects achieving accuracy of 93.91%, Precision of 91.16%, recall of 97.96% and F1-Measure of 94.33%. According to the research findings, the Support Vector Machine (SVM) algorithm provides the best performance when performing aspect-based sentiment analysis on text data from digital banking application reviews. The findings are expected to serve as a reference for the development of automated evaluation systems that rely on user opinions as the basis for decision making.             Keywords: aspects; CRISP-DM; digital Banking; seabank; sentiment analysis     Abstrak: Peningkatan pemakaian aplikasi perbankan digital mendorong perlunya pemahaman yang lebih dalam mengenai persepsi pengguna, terutama melalui analisis sentimen berbasis aspek. Penelitian ini bertujuan untuk mengklasifikasikan sentimen pengguna aplikasi SeaBank dengan berfokus pada empat aspek utama: kemudahan dipelajari (learnability), efisiensi penggunaan (efficiency), kendala atau kesalahan teknis (error), serta tingkat kepuasan (satisfaction). Data ulasan berjumlah 1.971 komentar dikumpulkan dari Google Play Store dan diberi label sentimen berdasarkan skor (rating) yang diberikan oleh pengguna. Pendekatan CRISP-DM berfungsi sebagai kerangka metodologis untuk penelitian ini, yang mencakup lima algoritma klasifikasi: Naïve Bayes, Support Vector Machine (SVM), k-Nearest Neighbor (k-NN), Decision Tree, dan Random Forest. Hasil evaluasi menunjukkan bahwa algoritma SVM memberikan performa terbaik dengan nilai rata-rata dari ke empat aspek tertinggi yang mencapai accuracy sebesar 93.91%, Precision sebesar 91.16%, recall sebesar 97.96% dan F1-Measure sebesar 94.33%. Menurut temuan penelitian, algoritma Support Vector Machine (SVM) memberikan kinerja terbaik saat melakukan analisis sentimen berbasis aspek pada data teks dari ulasan aplikasi Seabank. Temuan ini diharapkan dapat menjadi referensi bagi pengembangan sistem evaluasi otomatis yang mengandalkan opini pengguna sebagai dasar pengambilan keputusan.   Kata kunci: Analisis Sentimen, Aspek, Bank Digital, SeaBank, CRISP-DM

COMPARATIVE ANALYSIS OF MACHINE LEARNING ALGORITHMS FOR COSMETIC SALES PREDICTION ON TOKOPEDIA

Sahira, Mutia, Tania, Ken Ditha, Afrina, Mira
Abstract: Abstract: The rapid growth of the cosmetics industry on e-commerce platforms has intensified competition, creating a critical need for effective, data-driven marketing strategies. This study aims to conduct a comparative… analysis of machine learning algorithms to predict the sales categories (High, Medium, Low) of cosmetic products on the Tokopedia marketplace. Four classification models; Random Forest, XGBoost, Logistic Regression, and Naive Bayes were trained and evaluated on data collected via web scraping. The methodology incorporates the Synthetic Minority Over-sampling Technique (SMOTE) to address significant class imbalance and GridSearchCV for hyperparameter optimization to ensure a fair and robust comparison. The experimental results conclusively show that the Random Forest model achieved the best performance, yielding the highest F1-Score Macro Average of 0.75 and an accuracy of 85.3%. The superior model was subsequently implemented in a simple recommendation system to simulate optimal discount strategies, demonstrating its practical utility in providing actionable insights for business decisions. Keywords: classification; comparative analysis; machine learning; sales prediction; SMOTE   Abstrak: Pertumbuhan pesat industri kosmetik pada platform e-commerce telah membuat persaingan ketat, sehingga menciptakan kebutuhan krusial akan strategi pemasaran yang efektif dan berbasis data. Penelitian ini bertujuan untuk melakukan analisis komparatif terhadap algoritma machine learning untuk memprediksi kategori penjualan (Tinggi, Sedang, Rendah) produk kosmetik di marketplace Tokopedia. Empat model klasifikasi, yaitu Random Forest, XGBoost, Regresi Logistik, dan Naive Bayes, dilatih dan dievaluasi menggunakan data yang dikumpulkan melalui web scraping. Metodologi penelitian ini menerapkan Synthetic Minority Over-sampling Technique (SMOTE) untuk mengatasi ketidakseimbangan kelas yang signifikan dan GridSearchCV untuk optimisasi hyperparameter guna memastikan perbandingan yang adil. Hasil eksperimen menunjukkan bahwa model Random Forest mencapai performa terbaik, dengan menghasilkan F1-Score Macro Average tertinggi sebesar 0,75 dan akurasi 85,3%. Model unggul ini kemudian diimplementasikan dalam sebuah sistem rekomendasi sederhana untuk menyimulasikan strategi diskon yang optimal, yang menunjukkan kegunaan praktisnya dalam memberikan wawasan yang dapat ditindaklanjuti untuk pengambilan keputusan bisnis. Kata kunci: analisis komparatif; klasifikasi; machine learning; prediksi penjualan; SMOTE

DEVELOPMENT RICE PLANT DISEASE CLASSIFICATION USING CNN WITH TRANSFER LEARNING

Fitrony, Fachri Ayudi, Utami, Ema
Abstract: Abstract: The rice plant, Oryza sativa, is a major food source in Indonesia. This plant is processed into rice, a staple food for the Indonesian people. Rice growth is crucial to ensure the rice produced is of good quality.… ty. One part of the rice plant that is susceptible to disease is the leaves, which can inhibit growth and reduce rice quality. Therefore, early detection and accurate classification of rice diseases are crucial to minimize these negative impacts. This has driven the development of a Deep Learning model capable of high-performance automatic classification. This study aims to create a rice leaf classification model using the CNN algorithm and several transfer learning architectures such as ResNet101, VGG16, and Xception. A dataset of 859 rice leaf images collected from the Kaggle website was then processed using augmentation techniques to a total of 2,439 images, plus 215 smartphone photos for external data validation. Thus, the total dataset increased to 2,656 images, covering four categories: leafblast, brownspot, healthy, and hispa. The model was processed in two stages: on the initial dataset (Non-Augmented Dataset) and the Augmented Dataset. The best experimental results were obtained using the ResNet architecture, with a training accuracy of 96.17% and a validation accuracy of 95.22%. Based on the research results, the rice plant disease classification model using deep learning demonstrated good performance.             Keywords: convolutional neural network; deep learning; fine-tuning; image classification; resnet; rice plant