Improvisation of Web Crawling on Web Maps Using Distributed Algorithm with Socket Programming-Based Coordination Model

Authors

  • Muhammad Ridho Rizqillah Universitas Negeri Jakarta, Indonesia
  • Muhammad Eka Suryana Universitas Negeri Jakarta, Indonesia
  • Med Irzal Universitas Negeri Jakarta, Indonesia

DOI:

https://doi.org/10.21009/j-koma.v2i1.47277

Abstract

Search engines require a large amount of data to operate optimally. Handling large data necessitates massive crawling processes. This research aims to enhance crawling efficiency by implementing a distributed crawler using sockets as the communication medium between devices. The distributed architecture of the crawler system includes a tracker, manager, and client. Its primary goal is to improve crawling efficiency and effectiveness through distributed means. The final outcome of this innovation shows a distributed increase of 30% in data acquisition with two crawlers compared to individual crawlers, with no duplicated data.

Published

2024-06-28

How to Cite

[1]
Muhammad Ridho Rizqillah, Muhammad Eka Suryana, and Med Irzal, “Improvisation of Web Crawling on Web Maps Using Distributed Algorithm with Socket Programming-Based Coordination Model”, J-KOMA, vol. 6, no. 2, pp. 49–62, Jun. 2024.