Sonali Gupta et.al / International Journal on Computer Science and Engineering (IJCSE) WebParF:A Web Partitioning Framework for Parallel Crawler

Sonali Gupta
Komal Bhatia
Pikakshi Manchanda

Publication date

September 2014

Abstract

Abstract—With the ever proliferating size and scale of the WWW [1], efficient ways of exploring content are of increasing importance. How can we efficiently retrieve information from it through crawling? And in this “era of tera ” and multi-core processors, we ought to think of multi-threaded processes as a serving solution. So, even better how can we improve the crawling performance by using parallel crawlers that work independently? The paper devotes to the fundamental development in the field of parallel crawlers [4], highlighting the advantages and challenges arising from its design. The paper also focuses on the aspect of URL distribution among the various parallel crawling processes or threads and ordering the URLs within each distrib...

Extracted data

We use cookies to provide a better user experience.

Data Protection

Sonali Gupta et.al / International Journal on Computer Science and Engineering (IJCSE) WebParF:A Web Partitioning Framework for Parallel Crawler

Abstract

Extracted data

Sonali Gupta et.al / International Journal on Computer Science and Engineering (IJCSE) WebParF:A Web Partitioning Framework for Parallel Crawler

Abstract

Extracted data

Related items

Related items