Current issue state, recent activity, and per-issue timelines from the indexed issue data.
| Date | Opened | Closed | Comments | Events | Open Backlog |
|---|---|---|---|---|---|
| 2026-07-24 | 1 | 0 | 5 | 20 | 4 |
| 2026-07-23 | 3 | 4 | 2 | 14 | 2 |
| 2026-07-22 | 0 | 0 | 2 | 7 | 4 |
| 2026-07-21 | 3 | 1 | 3 | 9 | 5 |
| 2026-07-20 | 0 | 2 | 2 | 4 | 4 |
| 2026-07-19 | 1 | 1 | 0 | 0 | 1 |
| 2026-07-18 | 0 | 0 | 0 | 0 | 3 |
| 2026-07-17 | 1 | 0 | 0 | 0 | 0 |
| 2026-07-16 | 5 | 1 | 4 | 16 | 5 |
| 2026-07-15 | 4 | 3 | 2 | 15 | 7 |
| 2026-07-14 | 2 | 3 | 4 | 22 | 9 |
| 2026-07-13 | 1 | 2 | 5 | 21 | 9 |
| 2026-07-12 | 4 | 0 | 9 | 31 | 28 |
| 2026-07-11 | 0 | 0 | 0 | 0 | 0 |
Opened: 8
Closed: 8
Comments: 14
Events: 54
| Issue | Author | State | Labels | Comments | Reactions | Updated |
|---|---|---|---|---|---|---|
#1966 preempt: JobOrderFn tiebreak has no resource-awareness, falls back to creation-time order Opened 1 day ago | CoolingCube | open | enhancement | 2 | 0 | 8 hours ago |
#1944 Knative scale-to-zero causes PodGroup creation failure when Knative gang scheduling is enabled Opened 3 days ago | NaelAsbi123 | open | No labels | 3 | 1 | 14 hours ago |
#1913 feat(scheduler): allow jobs to override stale gang eviction grace period Opened 10 days ago | enoodle | closed - completed | enhancement good first issue | 5 | 0 | 14 hours ago |
#1968 stalegangeviction evicts pods of successfully completing jobs Opened 19 hours ago | lixiang233 | open | bug | 1 | 0 | 14 hours ago |
#1967 [Bug] CUDA_DEVICE_MEMORY_LIMIT is not set by HAMI Core plugin on UMA nodes where nvidia.com/gpu.memory label is missing Opened 1 day ago | akshnevrekar | closed - not_planned | bug | 1 | 0 | 1 day ago |
#1827 perf(scheduler): bound per-task per-node fit-error retention during allocation Opened 20 days ago | enoodle | closed - completed | No labels | 0 | 0 | 1 day ago |
#1964 Segmented PyTorch grouper ignores elastic minReplicas and requires all worker segments Opened 1 day ago | rotembubrunai | open | No labels | 1 | 1 | 1 day ago |
#1927 podgrouper: segmented PyTorch and LWS emit invalid minMember on parent SubGroups Opened 9 days ago | rotembubrunai | closed - completed | No labels | 2 | 0 | 2 days ago |
#1756 [Numa awareness] Support best effort topology manager Opened 1 month ago | davidLif | open | enhancement | 1 | 0 | 3 days ago |
#1945 status-updater floods logs and API server retrying status/patch updates for deleted PodGroups Opened 3 days ago | david-gang | open | No labels | 2 | 1 | 3 days ago |
#1948 Pipeline-only allocation leaks speculative fit errors Opened 3 days ago | enoodle | open | No labels | 1 | 1 | 3 days ago |
#1888 PodGroup schedulingConditions not cleared after the workload schedules Opened 12 days ago | gshaibi | closed - completed | No labels | 1 | 1 | 3 days ago |
#1929 docs/metrics: inconsistent naming between queue_deserved_gpus and queue_quota_* metrics Opened 9 days ago | david-gang | open | No labels | 1 | 0 | 4 days ago |
#1933 How to schedule a pod onto a cordoned node? Opened 8 days ago | lasse-ii | open | No labels | 3 | 0 | 4 days ago |
#1936 Set annotation to ignore quota occupied by workload Opened 7 days ago | amy | open | enhancement needs-design | 7 | 2 | 4 days ago |
#1939 reclaim: FeasibleNodesForJob reuses the reclaimer's node set to re-home victims, needlessly killing relocatable CPU-only victims Opened 5 days ago | david-gang | open | No labels | 4 | 0 | 4 days ago |
#1873 DRA GPU count overflow can understate queue demand and prevent eligible GPU reclaim Opened 15 days ago | thc1006 | closed - completed | No labels | 5 | 0 | 5 days ago |
#1584 Design kai behavior for priority class preemptionPolicy Opened 2 months ago | davidLif | closed - not_planned | enhancement | 1 | 0 | 5 days ago |
#848 Scheduler Assigns Multiple Workloads to Fully-Allocated GPU Node Causing Infinite Retry Loop Opened 7 months ago | dttung-starling | closed - completed | bug | 19 | 3 | 5 days ago |
#1821 feat: Container level VRAM metrics - integration with HAMi Opened 21 days ago | dttung2905 | open | enhancement | 1 | 0 | 6 days ago |
#1934 Can reclaim enforce fair-share (not just deserved quota) between sibling queues? Opened 8 days ago | lasse-ii | open | No labels | 1 | 0 | 8 days ago |
#1930 shared DRA ResourceClaim double-counted per pod makes node unschedulable Opened 9 days ago | TensorRaya | open | No labels | 1 | 1 | 8 days ago |
#1856 NUMA scale tests Opened 17 days ago | itsomri | closed - completed | enhancement | 0 | 0 | 8 days ago |
#1912 Cannot find parameter scheduler.gpuSharing.hamicoreEnabled in values.yaml file Opened 10 days ago | shakir91 | closed - completed | No labels | 2 | 1 | 9 days ago |
#1923 queue validator: validateParentChildQuota only sums CPU across siblings Opened 9 days ago | kshitizlohia1994 | open | No labels | 0 | 0 | 9 days ago |