Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-08-24 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-23 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-22 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-21 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-20 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-19 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-18 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-17 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-16 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-15 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-14 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-13 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-12 | 0 | 0 | 0 | 0 | 0 |
| 2026-08-11 | 0 | 0 | 0 | 0 | 0 |
Opened: 0
Closed: 0
Comments: 0
Events: 0
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#184 Energy Efficiency: 10 Mathematical Techniques for 60-70% AI Energy Reduction (Phi6Simple, FFT-Mix, Phi MoE) Opened 5 months ago | dancinlife | open | No labels | 0 | 0 | 5 months ago |
#83 ParallelDroplessMLP initialises self.mlp twice Opened 3 years ago | 152334H | open | enhancement help wanted | 7 | 1 | 9 months ago |
#180 Compatibility with CUDA 13.0 Opened 10 months ago | kiddyboots216 | open | No labels | 1 | 0 | 10 months ago |
#179 difference between cuda kernel code and using torch.ops Opened 11 months ago | abliao | open | No labels | 0 | 0 | 11 months ago |
#178 Issue with Sparse MLP Opened 1 year ago | Venkatesh3132003 | open | No labels | 1 | 0 | 1 year ago |
#170 Standard dmoe script crushes Opened 1 year ago | comane | closed - completed | No labels | 2 | 0 | 1 year ago |
#94 Unsharding scripts for megablocks models Opened 3 years ago | mayank31398 | closed - completed | No labels | 1 | 2 | 1 year ago |
#155 what devices are supported? Opened 2 years ago | Guodanding | open - reopened | No labels | 15 | 0 | 1 year ago |
#159 Pytorch 2.5.0 + Megablocks undefined symbol error Opened 2 years ago | jramapuram | closed - completed | No labels | 5 | 0 | 1 year ago |
#154 Memory cost increase gradually until oom Opened 2 years ago | maobenz | closed - completed | No labels | 3 | 0 | 1 year ago |
#166 MOE uses more memory than dense model and is slower Opened 1 year ago | samuelwheeler | open | No labels | 1 | 0 | 1 year ago |
#95 AMP + BF16 failing Opened 3 years ago | jramapuram | open | No labels | 7 | 0 | 1 year ago |
#165 bump version for PyTorch 2.6 Opened 1 year ago | ad8e | open | No labels | 0 | 1 | 1 year ago |
#134 Running into ValueError when running moe/dmoe scripts Opened 2 years ago | rtmadduri | closed - completed | No labels | 11 | 0 | 2 years ago |
#157 amp_C undefined symbol after installing Megablocks Opened 2 years ago | RachitBansal | open | No labels | 4 | 0 | 2 years ago |
#164 [rank5]: TypeError: SortOp.forward() takes from 2 to 3 positional arguments but 5 were given When running moe script Opened 2 years ago | rtmadduri | open | No labels | 0 | 0 | 2 years ago |
#163 Grouped GEMM execution not possible with HW Opened 2 years ago | cassanof | open | No labels | 2 | 0 | 2 years ago |
#161 CUDA Error When Running Single GPU Experiment Opened 2 years ago | kevin3567 | open | No labels | 0 | 0 | 2 years ago |
#160 Is current megablocks compatible with distributed optimizer in Megatron-LM? Opened 2 years ago | Spico197 | open | No labels | 1 | 0 | 2 years ago |
#156 Do you support the bias of mlp? Opened 2 years ago | maobenz | open | No labels | 1 | 0 | 2 years ago |
#153 Does it work with torch.compile? Opened 2 years ago | Muennighoff | open | No labels | 2 | 0 | 2 years ago |
#107 1-expert worse than dense model Opened 2 years ago | Muennighoff | open | No labels | 1 | 0 | 2 years ago |
#132 [Do not Merge] Add CI/CD Milestones Opened 2 years ago | eitanturok | open | No labels | 0 | 0 | 2 years ago |
#130 Torch ModuleNotFoundError when running pip install megablocks Opened 2 years ago | PaulMullerH | closed - completed | No labels | 0 | 0 | 2 years ago |
#61 Question on offsets in figures 5 Opened 3 years ago | DaehanKim | closed - completed | No labels | 2 | 0 | 2 years ago |