Skip to content

pdf2html is a module which helps to convert PDF file to HTML pages using Apache Tika. This module also helps to generate thumbnail image for PDF file using Apache PDFBox.

active 2023-08-162026-02-18 (UTC)

Complete coverage26,415 / 26,415 hourly files (100%) · 2 absent upstream2023-08-152026-08-19 (UTC)
Events
251
Pushes
30
Pull requests
25
Issues
12
Stars
92
Forks
13

Activity over time

Daily event counts in the loaded window

Line chart, 918 days from 2023-08-16 to 2026-02-18. Pushes: 30 total, peak 10 in a day. Pull requests: 25 total, peak 3 in a day. Issues: 12 total, peak 3 in a day. Comments: 39 total, peak 10 in a day. Stars: 92 total, peak 3 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

Recent activity

Latest issues, pull requests and releases

  • Issue comment#77sonarqubecloud[bot]2026-01-14 21:13
    Bump undici from 7.11.0 to 7.18.2
  • Pull request#77dependabot[bot]2026-01-14 21:12
  • Pull request#77dependabot[bot]2026-01-14 21:12
  • Pull request#77dependabot[bot]2026-01-14 21:12
  • Pull request#76dependabot[bot]2025-11-19 02:27
  • Pull request#76dependabot[bot]2025-11-19 02:27
  • Issue comment#44maclonghorn2025-07-16 18:57
    pdf2html.pages & pdf2html.html does not return images
  • Releaseshebinleo2025-07-13 12:07
    v4.4.0
  • Issue comment#44shebinleo2025-07-13 12:06
    pdf2html.pages & pdf2html.html does not return images
  • Issue#44shebinleo2025-07-13 12:06
    pdf2html.pages & pdf2html.html does not return images
  • Issue comment#74github-actions[bot]2025-07-13 11:58
    added feature to extract all images from the pdf #44
  • Issue comment#74sonarqubecloud[bot]2025-07-13 11:57
    added feature to extract all images from the pdf #44
  • Issue comment#74sonarqubecloud[bot]2025-07-13 11:49
    added feature to extract all images from the pdf #44
  • Issue comment#74sonarqubecloud[bot]2025-07-13 11:37
    added feature to extract all images from the pdf #44
  • Issue comment#74sonarqubecloud[bot]2025-07-13 11:34
    added feature to extract all images from the pdf #44
  • Issue comment#42shebinleo2025-07-05 06:11
    Use pagesHtml but input is loss
  • Issue#42shebinleo2025-07-05 06:11
    Use pagesHtml but input is loss
  • Pull request#73shebinleo2025-07-05 06:06
  • Issue comment#73github-actions[bot]2025-07-05 06:05
    fixes potential xss, outdated dependencies, path traversal
  • Issue comment#73sonarqubecloud[bot]2025-07-05 06:04
    fixes potential xss, outdated dependencies, path traversal
  • Issue comment#73github-actions[bot]2025-07-05 05:54
    fixes potential xss, outdated dependencies, path traversal
  • Issue comment#73sonarqubecloud[bot]2025-07-05 05:50
    fixes potential xss, outdated dependencies, path traversal
  • Pull request#73shebinleo2025-07-05 05:50
  • Releaseshebinleo2025-06-27 16:25
    v4.3.0
  • Issue#71shebinleo2025-06-27 16:23
    Package is currently uninstallable due to dependency

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 92 stars here means stars gained during the window, not the repo's star count.