Skip to content

KrishivGubba/Concurrent-WebCrawler

View on GitHub ↗Related repositories →

Java Web Crawler Project: This repository implements a multi-threaded web crawler in Java, designed to fetch and parse web pages concurrently. It includes classes for managing URL queues, tracking visited URLs, parsing HTML content, and handling data storage, rate limiting, and robots.txt compliance.

active 2024-07-102024-07-15 (UTC)

Complete coverage26,581 / 26,581 hourly files (100%) · 2 absent upstream2023-08-152026-08-26 (UTC)
Events
21
Pushes
7
Pull requests
8
Issues
0
Stars
1
Forks
0

Activity over time

Daily event counts in the loaded window

Line chart, 6 days from 2024-07-10 to 2024-07-15. Pushes: 7 total, peak 3 in a day. Pull requests: 8 total, peak 4 in a day. Issues: 0 total, peak 0 in a day. Comments: 0 total, peak 0 in a day. Stars: 1 total, peak 1 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
KrishivGubba15780

Recent activity

Latest issues, pull requests and releases

  • Pull request#4KrishivGubba2024-07-13 13:43
  • Pull request#4KrishivGubba2024-07-13 13:43
  • Pull request#3KrishivGubba2024-07-13 10:22
  • Pull request#3KrishivGubba2024-07-13 10:22
  • Pull request#2KrishivGubba2024-07-12 18:07
  • Pull request#2KrishivGubba2024-07-12 18:07
  • Pull request#1KrishivGubba2024-07-11 18:09
  • Pull request#1KrishivGubba2024-07-11 18:09

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 1 stars here means stars gained during the window, not the repo's star count.