Skip to content

This repository houses a script that can download PDFs from a specified URL, convert them to text, and perform text analysis. This analysis includes identifying the language, eliminating stopwords, and counting word and phrase frequency. It's worth noting that the script is capable of analyzing texts in multiple languages.

active 2024-06-18 → 2025-02-06 (UTC)

Complete coverage27,280 / 27,283 hourly files (100%) · 2 absent upstream2023-08-15 → 2026-09-24 (UTC)
Events
47
Pushes
29
Pull requests
6
Issues
1
Stars
3
Forks
1

Activity over time

Daily event counts in the loaded window

Line chart, 234 days from 2024-06-18 to 2025-02-06. Pushes: 29 total, peak 14 in a day. Pull requests: 6 total, peak 3 in a day. Issues: 1 total, peak 1 in a day. Comments: 1 total, peak 1 in a day. Stars: 3 total, peak 1 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
cortega26372961

Recent activity

Latest issues, pull requests and releases

  • Pull request#5cortega262024-07-13 16:51
  • Pull request#5cortega262024-07-13 15:51
  • Pull request#4cortega262024-07-13 15:50
  • Pull request#4cortega262024-07-09 21:49
  • Issue#2cortega262024-07-06 18:45
  • Issue comment#2cortega262024-07-06 18:45
  • Pull request#3cortega262024-06-25 00:52
  • Pull request#3cortega262024-06-18 22:27

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 3 stars here means stars gained during the window, not the repo's star count.