Repository logo
Communities & Collections
All of DSpace
  • English
  • العربية
  • বাংলা
  • Català
  • Čeština
  • Deutsch
  • Ελληνικά
  • Español
  • Suomi
  • Français
  • Gàidhlig
  • हिंदी
  • Magyar
  • Italiano
  • Қазақ
  • Latviešu
  • Nederlands
  • Polski
  • Português
  • Português do Brasil
  • Srpski (lat)
  • Српски
  • Svenska
  • Türkçe
  • Yкраї́нська
  • Tiếng Việt
Log In
New user? Click here to register.Have you forgotten your password?
  1. Home
  2. Browse by Author

Browsing by Author "Dr. Muhammad Hasanain Chaudary"

Filter results by typing the first few letters
Now showing 1 - 1 of 1
  • Results Per Page
  • Sort Options
  • No Thumbnail Available
    Item
    A Resource Efficient Method for Indexing Hidden Web Using Rank Based Web Crawling Techniques
    (Library Information Services, COMSATS University Islamabad, Lahore Campus, 2021) Abdul Mannan; FA18-RCS-026; LHR TP 7282; Dr. Muhammad Hasanain Chaudary
    Today we are seeing a shift in understanding and behavior of individuals toward anonymity and privacy. As a result, not only usage of virtual private network is increasing, but also more and more people are converging towards hidden web. Hidden web is a server less chain-based architecture, which can only be accessed using specific proxies and gateways and provides anonymity-using chain of interconnected nodes over public IP so that even if node is compromised anonymity of user is maintained. This form of security has its own drawbacks as overall network speed depends on node with lowest connectivity. Subsequently, crawling becomes costly since speed is proportional to the latency of generated Tor circuit. Another challenge in crawling Hidden web is the volatile nature of hosted services provided on it. Due to its anonymous nature, illegal services are prevalent. Consequently, this network is highly monitored becoming hot bed of banned websites some of which become live on new nodes while other stay down. This causes loss of time and resources crawler used to mine those dead URL. We want to propose a crawler that can mine data in low latency Tor network, auto tuning its configuration according to the state of the network and giving rank to websites in line with their content. This rank would be used to calculate crawling depth of a specific service at a given time, further improving as service stays alive

DSpace software copyright © 2002-2026 LYRASIS

  • Privacy policy
  • End User Agreement
  • Send Feedback
Repository logo COAR Notify