Skip to content

NVTabular is a feature engineering and preprocessing library for tabular data designed to quickly and easily manipulate terabyte scale datasets used to train deep learning based recommender systems.

active 2023-08-242026-05-22 (UTC)

Partial coverage20,537 / 26,293 hourly files (78%) · 2 absent upstream · 5,753 failed, retryable2023-08-152026-08-14 (UTC)— sampled evenly across the window, so rankings and trends hold; absolute counts scale up.
Events
192
Pushes
14
Pull requests
8
Issues
16
Stars
104
Forks
4

Activity over time

Daily event counts in the loaded window

Line chart, 1003 days from 2023-08-24 to 2026-05-22. Pushes: 14 total, peak 3 in a day. Pull requests: 8 total, peak 2 in a day. Issues: 16 total, peak 2 in a day. Comments: 37 total, peak 4 in a day. Stars: 104 total, peak 4 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

Recent activity

Latest issues, pull requests and releases

  • Pull request#1897jperez9992026-05-22 14:11
  • Issue comment#1895jorel6662026-05-10 05:38
    [QST] 139 SIGSEGV Error when creating NVTabular datasets from cudf dataframes
  • Issue#1895imran-marwat2025-11-28 13:12
    [QST] 139 SIGSEGV Error when creating NVTabular datasets from cudf dataframes
  • Issue#1895imran-marwat2025-11-28 13:12
    [QST] 139 SIGSEGV Error when creating NVTabular datasets from cudf dataframes
  • Issue comment#1894copy-pr-bot[bot]2025-10-23 16:35
    fix for load security
  • Issue#1891jordancaraballo2025-08-21 17:39
    [QST] PyPI Arm64 Support
  • Issue comment#1885Tottowich2025-06-16 19:41
    Slow performance of Categorify operation on Triton Inference Server
  • Issue#1889seneg0id2025-04-27 01:09
    [BUG] CUDA assert error
  • Issue comment#1886maciekrtb2024-12-11 07:29
    [BUG] ops.GroupBy after ops.Filter fails to group correctly, and produces unexpected NaNs
  • Issue#1885rahuljantwal-84512024-10-03 21:29
    Slow performance of Categorify operation on Triton Inference Server
  • Issue comment#1880rnyak2024-09-24 22:33
    [QST] NVTabular function is not supported for this dtype: size
  • Issue comment#1880anuragreddygv3232024-09-24 20:33
    [QST] NVTabular function is not supported for this dtype: size
  • Issue comment#1880rnyak2024-09-24 19:06
    [QST] NVTabular function is not supported for this dtype: size
  • Issue comment#1883rnyak2024-09-24 19:03
    [BUG] NVtabular.dataset.to_parquet(...) Improperly matched output dtypes detected in time, object and datetime64[ns]
  • Issue comment#1880anuragreddygv3232024-09-23 22:11
    [QST] NVTabular function is not supported for this dtype: size
  • Issue comment#1761Chevolier2024-08-18 13:38
    [QST] How can I fit a Workflow to a large dataset ?
  • Issue comment#1880rnyak2024-08-15 12:43
    [QST] NVTabular function is not supported for this dtype: size
  • Issue#1883Zachacy2024-08-15 08:00
    [BUG] NVtabular.dataset.to_parquet(...) Improperly matched output dtypes detected in time, object and datetime64[ns]
  • Issue comment#1880Chevolier2024-08-15 05:45
    [QST] NVTabular function is not supported for this dtype: size
  • Pull request#1882jperez9992024-08-09 21:01
  • Issue comment#1882copy-pr-bot[bot]2024-08-09 21:01
    fix blossom ci
  • Pull request#1882jperez9992024-08-09 21:01
  • Issue#1880LoMarrujo2024-07-08 23:09
    [QST] NVTabular function is not supported for this dtype: size
  • Issue comment#1876SunnyGhj2024-04-18 05:37
    [BUG] Distributed Training With (NVTabular + Pytorch DDP), I got this error: `RuntimeError: parallel_for: failed to synchronize: cudaErrorIllegalAddress: an illegal memory access was encountered`
  • Issue#1876SunnyGhj2024-04-18 05:37
    [BUG] Distributed Training With (NVTabular + Pytorch DDP), I got this error: `RuntimeError: parallel_for: failed to synchronize: cudaErrorIllegalAddress: an illegal memory access was encountered`

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 104 stars here means stars gained during the window, not the repo's star count.