Skip to content

OpenPecha/ocr_error_filtering

View on GitHub ↗Related repositories →

active 2024-05-102024-05-13 (UTC)

Complete coverage26,491 / 26,491 hourly files (100%) · 2 absent upstream2023-08-152026-08-22 (UTC)
Events
31
Pushes
13
Pull requests
1
Issues
13
Stars
0
Forks
0

Activity over time

Daily event counts in the loaded window

Line chart, 4 days from 2024-05-10 to 2024-05-13. Pushes: 13 total, peak 10 in a day. Pull requests: 1 total, peak 1 in a day. Issues: 13 total, peak 12 in a day. Comments: 1 total, peak 1 in a day. Stars: 0 total, peak 0 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
gangagyatso4364251201
ta4tsering3110

Recent activity

Latest issues, pull requests and releases

  • Issue#1ta4tsering2024-05-13 08:37
    OCR0022: Filtering Norbuketaka data
  • Pull request#8ta4tsering2024-05-13 08:37
  • Issue comment#1gangagyatso43642024-05-13 07:03
    OCR0022: Filtering Norbuketaka data
  • Issue#6gangagyatso43642024-05-10 10:40
    download the transcript from aws.
  • Issue#5gangagyatso43642024-05-10 10:40
    delete the line images and transcript with issues in s3 bucket.
  • Issue#4gangagyatso43642024-05-10 10:40
    save the backup line images (with issues) and transcript to a s3 bucket.
  • Issue#3gangagyatso43642024-05-10 10:40
    parse transcript to get batch ids of pages.
  • Issue#2gangagyatso43642024-05-10 10:39
    filter the issue file and get page ids.
  • Issue#7gangagyatso43642024-05-10 09:50
    train a model to classify the error line images using the dataset of deleted line images and transcript.
  • Issue#6gangagyatso43642024-05-10 09:50
    download the transcript from aws.
  • Issue#5gangagyatso43642024-05-10 09:50
    delete the line images and transcript with issues in s3 bucket.
  • Issue#4gangagyatso43642024-05-10 09:50
    save the backup line images (with issues) and transcript to a s3 bucket.
  • Issue#3gangagyatso43642024-05-10 09:49
    parse transcript to get batch ids of pages.
  • Issue#2gangagyatso43642024-05-10 09:49
    filter the issue file and get page ids.
  • Issue#1gangagyatso43642024-05-10 09:49
    OCR0022: Filtering Norbuketaka data

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 0 stars here means stars gained during the window, not the repo's star count.