Search Articles & Publications

Showing 348 articles found for "Most"

TOPSIS-BASED SYSTEM FOR THE SELECTION OF TRAINING PARTICIPANT CANDIDATES AT THE ASAHAN MANPOWER OFFICE

Maha Putra, Guntur, Wan Mariatul Kifti, Putri Amanda Nurhayati
Abstract: Abstract: Job training is one of the government’s efforts to improve the quality of human resources so that they possess competencies that meet labor market demands. The process of selecting training participants at the… e Department of Manpower of Asahan Regency is still carried out manually, which can lead to subjectivity and inefficiency in determining the most eligible candidates. This study aims to develop a decision support system using the Technique for Order Preference by Similarity to Ideal Solution (TOPSIS) method to assist the selection process objectively and systematically. The study applies four evaluation criteria, namely education level, age, work experience, and interview, with a dataset consisting of 31 training candidates. The system is developed as a web-based application using PHP programming language and MySQL database. The TOPSIS method is applied through decision matrix normalization, weighting, determination of positive and negative ideal solutions, and preference value calculation to produce a ranking of candidates. The results show that the proposed system can provide objective recommendations for selecting training participants, improve the efficiency of the selection process, and support decision makers in producing more accurate and reliable decisions. Keywords: decision support system; selection; training; TOPSIS.   Abstrak: Pelatihan tenaga kerja merupakan salah satu upaya pemerintah dalam meningkatkan kualitas sumber daya manusia agar memiliki kompetensi yang sesuai dengan kebutuhan dunia kerja. Proses pemilihan calon peserta pelatihan di Dinas Tenaga Kerja Kabupaten Asahan selama ini masih dilakukan secara manual sehingga berpotensi menimbulkan subjektivitas dan kurang efektif dalam menentukan peserta yang paling layak. Penelitian ini bertujuan untuk membangun sistem pendukung keputusan menggunakan metode Technique for Order Preference by Similarity to Ideal Solution (TOPSIS) untuk membantu proses seleksi peserta pelatihan secara objektif dan sistematis. Penelitian ini menggunakan empat kriteria penilaian yaitu pendidikan, usia, pengalaman kerja, dan wawancara dengan jumlah data sebanyak 31 calon peserta pelatihan. Sistem dikembangkan berbasis web menggunakan bahasa pemrograman PHP dan database MySQL. Metode TOPSIS digunakan untuk melakukan normalisasi matriks keputusan, pembobotan, penentuan solusi ideal positif dan negatif, serta perhitungan nilai preferensi untuk menghasilkan perankingan peserta pelatihan. Hasil penelitian menunjukkan bahwa sistem yang dibangun mampu memberikan rekomendasi peserta pelatihan secara objektif, meningkatkan efisiensi proses seleksi, serta membantu pihak dinas dalam pengambilan keputusan yang lebih akurat. Kata kunci: pelatihan; seleksi; sistem pendukung keputusan; TOPSIS.

STUDENT DEPRESSION SCREENING BASED ON THE OPTIMUM DATA BALANCING AND RANDOM FOREST

Adnan, M. Sayyidul, Budi Santoso, Irwan, Crysdian , Cahyo
Abstract: Abstract: Mental health issues, particularly depression among young adult university students, are often detected late due to stigma and reluctance to seek medical consultation. The objective of this study is to develop… an early screening model employing machine learning techniques, specifically the random forest algorithm, on a dataset of 268 students (aged 17-29 years; consisting of 98 males and 170 females) within a multicultural educational setting. The principal challenges associated with this dataset are class imbalance and the potential for data leakage from clinical scores. This study implements a rigorous feature selection approach that involves the elimination of depression score features and the utilization of the Synthetic Minority Over-sampling Technique (SMOTE) to balance the training data distribution. Furthermore, a Threshold Tuning strategy is employed to prioritize detection sensitivity (Recall). The findings indicate that reducing the decision threshold to an optimal value of 0.25 led to a substantial enhancement in the recall value, increasing it from 36% (baseline) to 77%. A feature importance analysis was conducted, the results of which indicated that Total Social Connectedness (ToSC) is the most dominant predictor. In summary, the present study corroborates the notion that optimizing sensitivity through threshold tuning is of paramount importance for medical screening. Furthermore, social isolation factors emerge as more significant indicators of depression risk than demographic attributes.             Keywords: data mining; depression; imbalanced data; random forest; smote; threshold tuning     Abstrak: Masalah kesehatan mental, khususnya depresi di kalangan mahasiswa dewasa muda, sering terdeteksi terlambat akibat stigma dan enggan mencari konsultasi medis. Tujuan studi ini adalah mengembangkan model skrining dini menggunakan teknik machine learning, khususnya algoritma random forest, pada dataset 268 mahasiswa (usia 17-29 tahun; terdiri dari 98 laki-laki dan 170 perempuan) dalam lingkungan pendidikan multikultural. Tantangan utama yang terkait dengan dataset ini adalah ketidakseimbangan kelas dan potensi kebocoran data dari skor klinis. Studi ini menerapkan pendekatan seleksi fitur yang ketat, yang melibatkan eliminasi fitur skor depresi dan penggunaan Teknik Over-sampling Minoritas Sintetis (SMOTE) untuk menyeimbangkan distribusi data pelatihan. Selain itu, strategi Penyesuaian Ambang Batas diterapkan untuk memprioritaskan sensitivitas deteksi (Recall). Hasil penelitian menunjukkan bahwa mengurangi ambang batas keputusan ke nilai optimal 0,25 menyebabkan peningkatan signifikan dalam nilai recall, dari 36% (dasar) menjadi 77%. Analisis pentingnya fitur dilakukan, hasilnya menunjukkan bahwa Total Social Connectedness (ToSC) adalah prediktor yang paling dominan. Secara ringkas, studi ini membenarkan bahwa mengoptimalkan sensitivitas melalui penyesuaian ambang batas sangat penting untuk skrining medis. Selain itu, faktor isolasi sosial muncul sebagai indikator risiko depresi yang lebih signifikan daripada atribut demografis.   Kata kunci: penambangan data; depresi; data tidak seimbang; hutan acak; smote; penyesuaian ambang batas

PREDICTING TEA HARVEST PRODUCTION AT BAH BUTONG USING RANDOM FOREST AND HISTORICAL DATA

Prayoga, Hafizd, Ramadhan Nasution, Yusuf
Abstract: Abstract: Accurate forecasts of tea harvest production are important for workforce planning, factory operations, and marketing decisions, yet conventional estimation in plantations often relies on field experience and can… n be biased and less adaptive to changing conditions. This study aims to develop a Random Forest Regression model to predict tea harvest production at the Bah Butong tea plantation using historical operational and climate-related data. The dataset consists of 60 monthly records (2020–2024) with six predictor variables: rainfall (mm), number of rainy days, pest level, weed level, number of harvested trees and land area. Data were split into 80% training (48 samples) and 20% testing (12 samples). Model hyperparameters were optimized using RandomizedSearchCV with RepeatedKFold cross-validation (5 folds, 3 repeats). The tuned model achieved MSE of 668,980,524.45, RMSE of 25,864.66 kg, MAE of 19,838.69 kg, and MAPE of 7.59% on the test set. The results indicate that the model can provide practical production estimates, with errors averaging about 7–8% of the actual production. Feature importance analysis shows that the number of harvested tea bushes and cultivated area contribute most to predictions. Future work should extend the historical period and incorporate time-based features (seasonality/lag) for improved forecasting.             Keywords: hyperparameter tuning; production prediction; random forest; regression; tea harvest   Abstrak: Perkiraan akurat produksi panen teh sangat penting untuk perencanaan tenaga kerja, operasional pabrik, dan keputusan pemasaran, namun estimasi konvensional di perkebunan seringkali bergantung pada pengalaman lapangan dan dapat bias serta kurang adaptif terhadap perubahan kondisi. Studi ini bertujuan untuk mengembangkan model Regresi Random Forest untuk memprediksi produksi panen teh di perkebunan teh Bah Butong menggunakan data operasional dan data terkait iklim historis. Dataset terdiri dari 60 catatan bulanan (2020–2024) dengan enam variabel prediktor: curah hujan (mm), jumlah hari hujan, tingkat hama, tingkat gulma, jumlah pokok panen, dan luas lahan. Data dibagi menjadi 80% data pelatihan (48 sampel) dan 20% data pengujian (12 sampel). Parameter model dioptimalkan menggunakan RandomizedSearchCV dengan validasi silang RepeatedKFold (5 lipatan, 3 pengulangan). Model yang telah disempurnakan mencapai MSE sebesar 668.980.524,45, RMSE sebesar 25.864,66 kg, MAE sebesar 19.838,69 kg, dan MAPE sebesar 7,59% pada set data uji. Hasil tersebut menunjukkan bahwa model dapat memberikan estimasi produksi yang praktis, dengan kesalahan rata-rata sekitar 7–8% dari produksi aktual. Analisis kepentingan fitur menunjukkan bahwa jumlah semak teh yang dipanen dan luas lahan budidaya paling berkontribusi pada prediksi. Pekerjaan selanjutnya harus memperpanjang periode historis dan menggabungkan fitur berbasis waktu (musiman/lag) untuk peramalan yang lebih baik.   Kata kunci: panen teh; prediksi produksi; random forest; regresi; tuning parameter

OPTIMIZATION OF FAST-MOVING DRUG INVENTORY USING THE WEIGHTED PRODUCT METHOD AT ANNISA DRUGSTORE

Dayanti, Rafika, Mulyani, Neni, Muhazir, Ahmad
Abstract: Abstract: Managing fast-moving drug inventory requires accurate supplier selection to ensure product availability and minimize the risk of overstock and out-of-stock conditions. At Annisa Pharmacy, the supplier selection… process has traditionally relied on experience and subjective judgment, which may lead to less optimal decisions. This study aims to design and implement a Decision Support System (DSS) for selecting fast-moving drug suppliers using the Weighted Product (WP) method. The WP method is applied because it is capable of processing multiple criteria simultaneously through structured weighting, including demand frequency, delivery lead time, remaining shelf life, purchase price, and profit margin. The system is developed as a web-based application using PHP and MySQL. The results show that the implementation of the Weighted Product method successfully produces preference values and accurate supplier rankings, enabling the system to correctly determine the most optimal fast-moving drug supplier based on the defined criteria. Therefore, the developed system can assist the owner of Annisa Pharmacy in making more precise, objective, and structured inventory procurement decisions. Keywords: decision support system; drug inventory; supplier selection; weighted product.   Abstrak: Pengelolaan stok obat fast moving di Toko Obat Annisa memerlukan ketepatan dalam menentukan supplier agar ketersediaan obat tetap terjaga dan risiko overstock maupun out of stock dapat diminimalkan. Selama ini, proses pemilihan supplier masih dilakukan secara konvensional berdasarkan pengalaman, sehingga berpotensi menghasilkan keputusan yang kurang optimal. Penelitian ini bertujuan untuk merancang dan mengimplementasikan Sistem Pendukung Keputusan (SPK) pemilihan supplier obat fast moving menggunakan metode Weighted Product (WP). Metode WP digunakan karena mampu mengolah beberapa kriteria secara simultan melalui pembobotan yang terstruktur, meliputi frekuensi permintaan, lead time, sisa masa kedaluwarsa, harga beli, dan margin keuntungan. Sistem dikembangkan berbasis web menggunakan bahasa pemrograman PHP dan basis data MySQL. Hasil penelitian menunjukkan bahwa penerapan metode Weighted Product mampu menghasilkan nilai preferensi dan perankingan supplier secara objektif, sehingga sistem berhasil menentukan supplier obat fast moving yang paling optimal sesuai dengan kriteria yang telah ditetapkan. Dengan demikian, sistem yang dibangun dapat membantu pemilik Toko Obat Annisa dalam mengambil keputusan pengadaan stok obat secara lebih tepat, objektif, dan terstruktur. Kata kunci: sistem pendukung keputusan; toko obat; pemilihan pemasok; weighted product

ANALYTIC NETWORK PROCESS IN DETERMINING RECIPIENTS OF EDUCATION GRANTS NORTH SUMATRA PROVINCE

Putri, Adelia Fariza, Fakhriza, M
Abstract: This study aims to apply the Analytic Network Process (ANP) method as a decision support tool in determining the eligibility of education grant recipients in North Sumatra Province. The background of this research arises&#8230; from the large number of grant applicants compared to the available budget, as well as the absence of clear and objective evaluation standards. The ANP method was chosen because it allows the interdependence between assessment criteria such as institutional feasibility, performance and achievement, social and educational impact, and accountability and transparency to be analyzed comprehensively. Data were obtained through interviews, documentation, and observation at the North Sumatra Provincial Education Office. The results of the ANP model show that the criterion with the highest weight is accountability and transparency (0.44), followed by social and educational impact (0.31). Among the three alternatives, community-based education foundations (A2) obtained the highest total weight (0.30), indicating that they are the most eligible recipients of education grants. The implementation of the ANP-based decision support system produces valid and consistent ranking results (CR < 0.1), enabling faster, fairer, and more transparent decision-making. Therefore, the ANP method contributes significantly to improving governance, objectivity, and accountability in the distribution of education grants in North Sumatra Province.

HEURISTIC GREEDY ALGORITHM FOR OPTIMAL TOURIST ROUTE RECOMMENDATION IN PATI REGENCY

Mohammad Ilham Kurnia, Alif Catur Murti, Rizkysari Mei Maharani
Abstract: Abstract: Tourism in Pati Regency currently lacks an integrated digital information system, resulting in suboptimal dissemination of information and trip planning. To address this issue, a tourism website for Pati Regency&#8230; y was developed, equipped with a recommended tourist route feature. This study aims to design and develop a web-based tourism information system that provides destination information based on categories, media galleries, and promotional YouTube videos, as well as a Patiways feature that allows users to select multiple tourist destinations. The system then calculates the most efficient visiting order using a greedy heuristic algorithm, based on the selected starting point. The system was developed using the Waterfall method, consisting of analysis, design, implementation, and testing phases. The system design is illustrated through UML diagrams such as Use Case, Activity, and Class Diagrams. With this system, the distribution of tourism information becomes more effective, and tourists can plan trips with optimized routes. Additionally, the website is expected to serve as a digital promotion medium that contributes to increasing tourist visits to Pati Regency. Keywords: heuristic greedy; recommendation route; tourism; waterfall   Abstrak: Pariwisata di Kabupaten Pati saat ini belum memiliki sistem informasi digital yang terintegrasi, sehingga penyebaran informasi dan perencanaan perjalanan wisata masih belum optimal. Untuk mengatasi permasalahan tersebut, penelitian ini mengembangkan sebuah website pariwisata Kabupaten Pati yang dilengkapi dengan fitur rekomendasi rute wisata terbaik. Penelitian ini bertujuan untuk merancang dan membangun sistem informasi pariwisata berbasis web yang mampu menyajikan informasi destinasi wisata berdasarkan kategori, galeri media, serta video promosi YouTube. Selain itu, sistem ini dilengkapi dengan fitur unggulan bernama Patiways yang memungkinkan pengguna memilih beberapa destinasi wisata dan secara otomatis memperoleh urutan kunjungan paling efisien menggunakan algoritma heuristik greedy berdasarkan titik awal perjalanan. Pengembangan sistem dilakukan menggunakan metode Waterfall yang meliputi tahapan analisis kebutuhan, perancangan sistem, implementasi, dan pengujian. Perancangan sistem direpresentasikan menggunakan diagram UML, meliputi Use Case Diagram, Activity Diagram, dan Class Diagram. Dengan adanya sistem ini, diharapkan penyebaran informasi pariwisata menjadi lebih efektif, wisatawan dapat merencanakan perjalanan dengan rute yang optimal, serta website dapat berfungsi sebagai media promosi digital yang berkontribusi terhadap peningkatan kunjungan wisatawan ke Kabupaten Pati. Kata kunci: heuristik greedy; pariwisata; rekomendasi rute; waterfall

COMPARISON OF CLUSTERING MODELS FOR GROUPING LIFESTYLE PATTERNS AND OBESITY FACTORS

Al Mas Ud, Khalid, Fathoni, Fathoni, Muhammad Kurniawan, Hafiz
Abstract: Abstract: Obesity is an escalating global health concern, with unhealthy lifestyle patterns contributing significantly to its development. This study aims to evaluate and compare three clustering techniques for categorizing&#8230; ing lifestyle patterns and obesity-related factors: K-Means, Agglomerative Clustering, and Gaussian Mixture Model (GMM). The data used in this study is sourced from the Food Nutrition dataset, which includes variables such as dietary habits, physical activity, and socio-economic status. The three clustering methods were assessed using evaluation metrics such as Silhouette Score, Davies-Bouldin Index (DBI), and Calinski-Harabasz Index (CHI). The findings revealed that K-Means exhibited the best performance in terms of cluster separation with a Silhouette Score of 0.5559, while GMM showed better flexibility in handling more complex data. Although Agglomerative Clustering produced acceptable results, it had a higher overlap between clusters compared to the other methods. This study offers valuable insights into selecting the most appropriate clustering technique based on the data characteristics.             Keywords: agglomerative; clustering; GMM; k-means; lifestyle patterns; obesity   Abstrak: Obesitas menjadi masalah kesehatan yang semakin meningkat di seluruh dunia, dengan pola hidup yang tidak sehat berperan besar dalam perkembangannya. Penelitian ini bertujuan untuk membandingkan tiga metode clustering dalam mengelompokkan pola gaya hidup dan faktor yang memengaruhi obesitas, yaitu K-Means, Agglomerative Clustering, dan Gaussian Mixture Model (GMM). Data yang digunakan diperoleh dari dataset Food Nutrition yang mencakup informasi terkait pola makan, aktivitas fisik, serta faktor sosial-ekonomi. Ketiga metode tersebut diuji dengan menggunakan beberapa metrik evaluasi, seperti Silhouette Score, Davies-Bouldin Index (DBI), dan Calinski-Harabasz Index (CHI). Hasil penelitian menunjukkan bahwa K-Means memiliki kinerja terbaik dalam hal pemisahan klaster, dengan nilai Silhouette Score sebesar 0.5559, sementara GMM lebih fleksibel dalam menangani data yang lebih kompleks. Meskipun Agglomerative Clustering memberikan hasil yang dapat diterima, tumpang tindih antar klaster lebih besar dibandingkan dengan kedua metode lainnya. Penelitian ini memberikan pemahaman yang lebih baik mengenai pemilihan metode clustering yang tepat berdasarkan karakteristik data yang digunakan.   Kata kunci: agglomerative; clustering; GMM; k-means; obesitas; pola gaya hidup

ANALYSIS OF INTEREST IN USING BLU DEPOSIT BASED ON TAM

Pangestu, Nathania Clarissa, Pratiwi , Heny, Yusnita, Amelia
Abstract: Abstract: Digital banking has brought various innovations in financial services, one of which is Blu Deposito by BCA Digital. However, the adoption rate of digital deposit services is still relatively low compared to digital&#8230; ital payment services. This study aims to identify and analyze the factors that influence customers' intentions and actual behavior in using Blu Deposito with reference to the Technology Acceptance Model (TAM). This study aims to analyze the factors that influence customers' intentions and actual behavior in adopting Blu Deposito using the Technology Acceptance Model (TAM) framework. Data was collected through a Google Form questionnaire from 54 customers at one BCA branch and analyzed using SPSS through validity and reliability tests, descriptive analysis, and multiple regression. The results show that Behavioral Intention (BI)is significantly influenced by Perceived Ease of Use (PEOU), Perceived Usefulness (PU), and Attitude Toward Using (ATU), with PEOU as the most dominant factor. In addition, BI has a significant effect on Actual System Use (AU), which confirms the relevance of applying the TAM model in the context of digital deposit products. These findings indicate that ease of use plays a greater role than financial benefits in encouraging users to adopt Blu Deposits. This study contributes to the understanding of digital deposit adoption and provides managerial insights to improve the usability and user engagement of digital banking services. Keywords: actual system use; attitude toward using; behavioral intention; perceived ease of use; perceived usefulness; technology acceptance model   Abstrak: Perbankan digital telah menghadirkan berbagai inovasi dalam layanan keuangan, salah satunya Blu Deposito oleh BCA Digital. Meskipun demikian, tingkat adopsi terhadap layanan deposito digital masih relatif rendah dibandingkan dengan layanan pembayaran digital. Penelitian ini bertujuan untuk mengidentifikasi dan menganalisis faktor-faktor yang memengaruhi niat serta perilaku aktual nasabah dalam menggunakan Blu Deposito dengan mengacu pada kerangka Technology Acceptance Model (TAM). Penelitian ini bertujuan untuk menganalisis faktor-faktor yang memengaruhi niat dan perilaku aktual nasabah dalam mengadopsi Blu Deposito dengan menggunakan kerangka Technology Acceptance Model (TAM). Data dikumpulkan melalui kuesioner Google Form dari 54 nasabah di satu cabang BCA dan dianalisis menggunakan SPSS melalui uji validitas, reliabilitas, analisis deskriptif, dan regresi berganda. Hasil penelitian menunjukkan bahwa Behavioral Intention (BI) dipengaruhi secara signifikan oleh Perceived Ease of Use (PEOU), Perceived Usefulness (PU), dan Attitude Toward Using (ATU), dengan PEOU sebagai faktor paling dominan. Selain itu, BI berpengaruh signifikan terhadap Actual System Use (AU), yang menegaskan relevansi penerapan model TAM pada konteks produk deposito digital. Temuan ini menunjukkan bahwa kemudahan penggunaan memiliki peran lebih besar dibandingkan manfaat finansial dalam mendorong pengguna untuk mengadopsi Blu Deposito. Penelitian ini berkontribusi terhadap pemahaman adopsi deposito digital serta memberikan wawasan manajerial untuk meningkatkan kegunaan dan keterlibatan pengguna pada layanan perbankan digital.   Kata kunci: actual system use; attitude toward using; behavioral intention; perceived ease of use; perceived usefulness; technology acceptance model  

COMPARISON SVM, RF, BERT PUBLIC SENTIMENT DATA MBG IN X

Gustri Efendi, Yandi, Rus, Aprilia, Rani, Amaroh Bit Taqwa, Irvan
Abstract: Abstract: MBG is a strategic program of the Prabowo-Gibran administration. This program has become a widely discussed issue in the public. To better understand public perception of this program, sentiment analysis is necessary.&#8230; essary. This study aims to compare the performance of algorithms machine learning SVM, RF, And BERT with preprocessing data analyzing public sentiment of the MBG program in media X. The total dataset for this study was 39,858 out of 42,465 successfully crawled tweets. The research methods included data collection, preprocessing data (cleaning, case folding, word normalization, stopword removal and stemming), feature extraction, model training (fine-tuning), handling class imbalance with SMOTE, and evaluation using accuracy, precision, recall, and f1-score. The research results show that without SMOTE, the best performing models are BERT with 89% accuracy, SVM 87%, and RF 78.4%. After SMOTE, the best algorithms were SVM with 92.94%, BERT with 88.3%, and RF with 86.59%. The results confirmed that SVM is the best algorithm if at leastclass imbalance. BERT is the best algorithm before and after SMOTE, because BERT is more effective in capturing the nuances of language on social media, so BERT is the most recommended in MBG sentiment analysis.             Keywords: sentiment analysis; machine learning; SVM, RF, and BERT   Abstrak: MBG merupakan program strategis pemerintahan Prabowo - Gibran. Program ini menjadi isu yang banyak diperbincangkan publik. Untuk mengetahui lebih dalam persepsi masyrakat tentang program ini, perlu dilakukan analisis sentiment. Penelitian ini bertujuan membandingkan kinerja algoritma machine learning SVM, RF, dan BERT dengan preprocessing data menganalisis sentiment public program MBG di media X. Total dataset penelitian ini adalah 39.858 dari 42.465 tweet yang berhasil di crawling. Metode penelitian mencakup pengumpulan data, preprocessing data (cleaning, case folding, normalisasi kata, stopword removal dan stemming), ekstraksi fitur, pelatihan model (fine-tuning), penanganan class imbalance dengan SMOTE, dan evaluasi menggunakan akurasi, presisi, recall, dan f1-score. Hasil peneltian menunjukkan, tanpa SMOTE model dengan kinerja terbaik adalah BERT dengan akurasi 89%, SVM 87%, dan RF 78,4%. Setelah SMOTE algoritma terbaik adalah SVM 92,94%, BERT 88,3% dan RF 86,59%. Hasil penelitian menegaskan bahwa SVM adalah algoritma terbaik jika minimal class imbalance. BERT adalah algoritma terbaik sebelum dan sesudah SMOTE, karena BERT lebih efektif dalam menangkap nuansa bahasa pada media sosial, sehingga BERT paling di rekomendasikan dalam analisis sentimen MBG.   Kata kunci: analisis sentimen; machine learning; SVM, RF, dan BERT

MACHINE LEARNING CONTENT-BASED FILTERING WOMEN EMPOWERING RECOMMENDATIONS ON YOUTUBE

Yuliana, Yuliana, Mira, Mira, Hari Kristianto, Aloysius
Abstract: Abstract: YouTube is one of the most popular video streaming platforms, but it has constraints that can cause problems when clients have difficulty finding content according to their wishes. The main objective of this study&#8230; udy is to increase user capacity in viewing content specifically in the field of women's empowerment. By using content-based filtering techniques, the system will analyze user preferences and interests through recommendations for women's empowerment content. The data source is via the YouTube API and is analyzed using PHP programming content-based filtering techniques. The system's recommendations provide a list of women's empowerment content with a user request display. The results of the research evaluation obtained a precision value of 62%, meaning that the recommendations match the topic being searched for, namely women's empowerment. The recall value of 84% indicates that the system has succeeded in finding relations from the database. The f1-score value of 72% indicates that there is a balance between precision and recall, meaning that a system is needed that is not only accurate but also complete. While the cosine value shows a score of 0.7071 approaching the maximum value (1.0). The recommendation of the content-based filtering method produces quite effective women's empowerment content. Keywords: content-based filtering, recommendations, women Empowerment, youtube