IMPLEMENTASI DAN EVALUASI WEB SCRAPING PADA DATA JURNAL SINTA

Authors

  • Indra Nasution Universitas Pembangunan Panca Budi
  • Muhammad Iqbal Universitas Pembangunan Panca Budi

DOI:

https://doi.org/10.54314/jssr.v9i3.6810

Keywords:

Web Scraping, SINTA, Scientific Journals, Python, Beautifulsoup, Data Extraction

Abstract

Abstract: The development of digital technology has increased the need for access to scientific information, especially related to academic journal publications. The Science and Technology Index (SINTA) website is one of the main platforms that provides information on accredited scientific journals in Indonesia. However, the process of searching and collecting journal data on the website is still done manually, so it is less efficient when the amount of data needed is large enough. This research aims to implement and evaluate web scraping techniques in collecting journal data on the SINTA website so that the data acquisition process can be carried out automatically, quickly, and in a structured manner. The research method uses an implementive approach through the stages of literature study, analysis of the structure of website pages, development of a system using the Python programming language with the BeautifulSoup library, data collection, database storage, and evaluation of extraction results. The results of the study show that the implementation of the web scraping technique  succeeded in obtaining as many as 15,456 journal data from the SINTA website in 2026 which includes information on journal names, publishing institutions, accreditation categories, ISSN, fields of science, and journal website links. The developed system is also capable of storing data in a structured manner into a database and supports regular information updates. The results of the evaluation show that the web scraping technique is able to improve the efficiency of data collection, minimize manual processes, and produce more systematic journal data. Thus, the implementation of web scraping techniques can be an effective solution in supporting the needs of academics and researchers to obtain journal information quickly and accurately.

Keywords: Web Scraping, SINTA, Scientific Journals, Python, Beautifulsoup, Data Extraction.

 

Abstrak: Perkembangan teknologi digital telah meningkatkan kebutuhan akses terhadap informasi ilmiah, terutama yang berkaitan dengan publikasi jurnal akademik. Situs web Indeks Sains dan Teknologi (SINTA) merupakan salah satu platform utama yang menyediakan informasi tentang jurnal ilmiah terakreditasi di Indonesia. Namun, proses pencarian dan pengumpulan data jurnal di situs web tersebut masih dilakukan secara manual, sehingga kurang efisien ketika jumlah data yang dibutuhkan cukup besar. Penelitian ini bertujuan untuk mengimplementasikan dan mengevaluasi teknik web scraping dalam pengumpulan data jurnal di situs web SINTA sehingga proses akuisisi data dapat dilakukan secara otomatis, cepat, dan terstruktur. Metode penelitian menggunakan pendekatan implementasi melalui tahapan studi pustaka, analisis struktur halaman web, pengembangan sistem menggunakan bahasa pemrograman Python dengan pustaka BeautifulSoup, pengumpulan data, penyimpanan basis data, dan evaluasi hasil ekstraksi. Hasil penelitian menunjukkan bahwa implementasi teknik web scraping berhasil memperoleh sebanyak 15.456 data jurnal dari situs web SINTA pada tahun 2026 yang mencakup informasi tentang nama jurnal, lembaga penerbitan, kategori akreditasi, ISSN, bidang ilmu, dan tautan situs web jurnal. Sistem yang dikembangkan juga mampu menyimpan data secara terstruktur ke dalam basis data dan mendukung pembaruan informasi secara berkala. Hasil evaluasi menunjukkan bahwa teknik web scraping mampu meningkatkan efisiensi pengumpulan data, meminimalkan proses manual, dan menghasilkan data jurnal yang lebih sistematis. Dengan demikian, implementasi teknik web scraping dapat menjadi solusi efektif dalam mendukung kebutuhan akademisi dan peneliti untuk memperoleh informasi jurnal dengan cepat dan akurat.

Kata kunci: Web Scraping, SINTA, Jurnal Ilmiah, Python, Beautifulsoup, Ekstraksi Data.

Downloads

Download data is not yet available.

References

N. E. Widiyastuti et al., Inovasi & pengembangan karya tulis ilmiah: Panduan lengkap untuk penelitian dan mahasiswa. PT. Sonpedia Publishing Indonesia, 2023.

A. Julia, S. Febriani, and S. Nurjanah, “PUBLIKASI ILMIAH DAN INOVASI PENDIDIKAN,” J. Ilmu Pendidik. Islam, vol. 23, no. 4, pp. 225–240, 2025.

N. Agustin and A. Fithriyah, “Pendampingan penulisan karya ilmiah bagi mahasiswa sebagai upaya peningkatan budaya akademik di perguruan tinggi,” J. Pengabdi. Masy., vol. 2, no. 2, pp. 235–246, 2025.

K. T. Ilmiah, “Peningkatan Pengetahuan Dan Keterampilan Dalam Penyusunan Karya Tulis Ilmiah Terakreditasi Sinta,” Community Dev. J., vol. 4, no. 2, 2023.

A. Saputra, “Pemanfaatan science and technology index (sinta) untuk publikasi karya ilmiah dan pencarian jurnal nasional terakreditasi,” Media Pustak., vol. 27, no. 1, pp. 56–68, 2020.

D. Darman, T. S. Maulana, and R. Mansur, “Pendampingan optimalisasi profil SINTA bagi dosen melalui perbaikan data dan penguatan dokumentasi publikasi,” Eastasouth J. Eff. Community Serv., vol. 4, no. 02, pp. 309–317, 2025.

D. Chrisinta and J. E. Simarmata, “Eksplorasi teknik web scraping pada data mining: Pendekatan pencarian data berbasis Python,” Fakt. Exacta, vol. 17, no. 1, 2024.

M. R. Fikri, R. T. Handayanto, and D. Irwan, “Web Scraping Situs Berita Menggunakan Bahasa Pemograman Python,” J. Students ‘Research Comput. Sci., vol. 3, no. 1, pp. 123–136, 2022.

N. ADILA, “IMPLEMENTASI WEB SCRAPING UNTUK PENGUMPULAN DATA JURNAL PADA WEBSITE SINTA,” Nusa Putra, 2022.

L. M. Purnomo and M. Ayub, “Analisis data hasil web scraping untuk menentukan kualitas jurnal ilmiah,” J. Strateg. Maranatha, vol. 3, no. 1, pp. 122–132, 2021.

E. R. Arumi and P. Sukmasetya, “Exploiting web scraping for education news analysis using depth-first search algorithm,” J. Online Inform., vol. 5, no. 1, pp. 19–26, 2020.

A. Rahmatulloh, R. Gunawan, and others, “Web scraping with HTML DOM method for data collection of scientific articles from Google Scholar,” Indones. J. Inf. Syst., vol. 2, no. 2, pp. 95–104, 2020.

U. Mufidah and M. Siahaan, “Perancangan aplikasi perbandingan harga produk (historical data) menggunakan teknik web scraping,” Pusdansi. org, vol. 1, no. 1, pp. 1–14, 2021.

Y. Sahria, “Implementasi Teknik Web Scraping pada jurnal SINTA untuk analisis topik penelitian kesehatan Indonesia,” Proceeding of The URECOL, pp. 297–306, 2020.

A. Priyanto, M. R. Ma’arif, and others, “Implementasi web scrapping dan text mining untuk akuisisi dan kategorisasi informasi dari internet (studi kasus: Tutorial hidroponik),” Indones. J. Inf. Syst., vol. 1, no. 1, pp. 25–33, 2018.

M. F. Adham, “Analisis implementasi sistem informasi: studi literatur,” J. Teknol. Sist. Inf., vol. 5, no. 1, pp. 264–275, 2024.

P. Deepublish, “Studi Literatur: Pengertian, Ciri, Teknik Pengumpulan Datanya,” Diakses pada tanggal, vol. 7, 2023.

F. Sembiring and A. Erfina, Bahasa Ular untuk Pemrograman Python. Insan Cendekia Mandiri, 2020

R. Mitchell, Web scraping with python. “ O’Reilly Media, Inc.,” 2024.

C. Chusna, S. Ilham, A. C. Fauzan, and others, “Implementasi Penjadwalan Round Robin pada Task Scheduler untuk Pembaruan Aplikasi Otomatis,” ILKOMNIKA, vol. 1, no. 1, pp. 11–14, 2019.

Downloads

Published

2026-06-30

Issue

Section

Artikel

How to Cite

IMPLEMENTASI DAN EVALUASI WEB SCRAPING PADA DATA JURNAL SINTA. (2026). JOURNAL OF SCIENCE AND SOCIAL RESEARCH, 9(3), 5220-5228. https://doi.org/10.54314/jssr.v9i3.6810

Most read articles by the same author(s)

1 2 > >>