[CVPR'23] MM-Diffusion: Learning Multi-Modal Diffusion Models for Joint Audio and Video Generation
active 2023-08-16 → 2026-04-10 (UTC)
Activity over time
Daily event counts in the loaded window
Line chart, 969 days from 2023-08-16 to 2026-04-10. Pushes: 3 total, peak 1 in a day. Pull requests: 0 total, peak 0 in a day. Issues: 27 total, peak 2 in a day. Comments: 27 total, peak 3 in a day. Stars: 220 total, peak 5 in a day.
- Pushes
- Pull requests
- Issues
- Comments
- Stars
Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.
Top contributors
Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity
| Contributor | Contributions | Pushes | PRs | Comments |
|---|---|---|---|---|
| TousenKaname | 6 | 0 | 0 | 2 |
| ludanruan | 5 | 3 | 0 | 1 |
| mayank-git-hub | 5 | 0 | 0 | 4 |
| kaiw7 | 4 | 0 | 0 | 3 |
| SoloMannn | 4 | 0 | 0 | 4 |
| TWTWTWTWTWTWTWTW | 4 | 0 | 0 | 2 |
| ltzheng | 3 | 0 | 0 | 0 |
| ChdDongyang | 3 | 0 | 0 | 1 |
| xuihan | 3 | 0 | 0 | 0 |
| henu77 | 2 | 0 | 0 | 1 |
| Fant4sti | 2 | 0 | 0 | 1 |
| kalelpark | 2 | 0 | 0 | 0 |
| Ashigarg123 | 2 | 0 | 0 | 2 |
| ShruthiVijayakumar | 1 | 0 | 0 | 0 |
| yuchenli-sony | 1 | 0 | 0 | 1 |
| TreeberryTomato | 1 | 0 | 0 | 1 |
| YinzhenWang | 1 | 0 | 0 | 0 |
| samarthtehri | 1 | 0 | 0 | 0 |
| andyl-flwls | 1 | 0 | 0 | 0 |
| MrGuanxi | 1 | 0 | 0 | 0 |
Recent activity
Latest issues, pull requests and releases
- Issue comment#23julkaztwittera2025-06-14 23:40training from scratch problem
- Issue#25henu772025-03-13 13:58Which encoders and decoders are used for each modality of data
- Issue comment#9henu772025-03-13 13:57question about paper
- Issue comment#15yuchenli-sony2025-01-25 13:56OOM in multimodal_sample_sr.sh
- Issue comment#18URRealHero2024-12-11 04:44Dependency problem
- Issue#23andyl-flwls2024-09-12 02:54training from scratch problem
- Issue#22Fant4sti2024-08-09 14:28"slow_conv2d_cpu" not implemented for 'Half'
- Issue comment#7Fant4sti2024-08-09 14:21RuntimeError: "slow_conv2d_cpu" not implemented for 'Half'
- Issue#21kalelpark2024-06-24 14:26The dataset doesn't include sound.
- Issue#21kalelpark2024-06-24 13:02The dataset doesn't include sound.
- Issue#17ludanruan2024-06-05 12:20Reproducing results in the paper
- Issue comment#17ludanruan2024-06-05 02:48Reproducing results in the paper
- Issue comment#17kaiw72024-05-08 08:29Reproducing results in the paper
- Issue comment#17mayank-git-hub2024-05-07 09:04Reproducing results in the paper
- Issue comment#17kaiw72024-05-05 21:12Reproducing results in the paper
- Issue comment#20mayank-git-hub2024-05-05 09:10How many iterations have the pretrained models been trained for?
- Issue comment#20kaiw72024-05-03 16:56How many iterations have the pretrained models been trained for?
- Issue comment#19TreeberryTomato2024-04-21 03:11computational complexity in paper
- Issue#20mayank-git-hub2024-04-17 08:16How many iterations have the pretrained models been trained for?
- Issue comment#17mayank-git-hub2024-04-17 05:44Reproducing results in the paper
- Issue comment#17mayank-git-hub2024-04-16 12:32Reproducing results in the paper
- Issue comment#13brandocaeser2024-04-11 14:50RuntimeError: "avg_pool3d_out_frame" not implemented for 'Half
- Issue#17ltzheng2024-04-07 14:04Reproducing results in the paper
- Issue comment#12TWTWTWTWTWTWTWTW2024-04-06 08:29About generating 10 seconds of audio
- Issue#18TWTWTWTWTWTWTWTW2024-04-06 08:27Dependency problem
Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 220 stars here means stars gained during the window, not the repo's star count.