Repository Issue Activity (beta)

deepspeedai/DeepSpeedExamples

Current issue state, recent activity, and per-issue timelines from the indexed issue data.

Open Issues
27
New in 7 Days
0
Closed in 7 Days
0
Average Open Age
864 days
Stale 30+ Days
27
Stale 90+ Days
27
Last 2 Weeks
DateOpenedClosedCommentsEventsOpen Backlog
2026-09-1300000
2026-09-1200000
2026-09-1100000
2026-09-1000000
2026-09-09000027
2026-09-0800000
2026-09-0700000
2026-09-0600000
2026-09-0500000
2026-09-0400000
2026-09-0300000
2026-09-0200000
2026-09-0100000
2026-08-3100000
This Week

Opened: 0

Closed: 0

Comments: 0

Events: 0

Top Labels
deespeed chat (5)
enhancement (1)
hybrid engine (1)
question (1)
system (1)
Issue Explorer
IssueAuthorStateLabelsCommentsReactionsUpdated

#996 Compute_rewards in PPO:rewards[j, start:ends[j]][-1] += reward_clip[j] is wrong

Opened 9 months ago
CZH-THU
open
No labels
009 months ago

#995 Step1 failed when I run "bash training_scripts/opt/single_gpu/run_1.3b.sh"

Opened 9 months ago
niebowen666
open
No labels
009 months ago

#524 RuntimeError: Error building extension 'transformer_inference'

Opened 3 years ago
li995495592
closed - completed
deespeed chat
501 year ago

#989 One example, multiple config files

Opened 1 year ago
delock
open
No labels
001 year ago

#988 Gradient overflow issue while using deepspeed

Opened 1 year ago
jaydeepborkar
closed - completed
No labels
101 year ago

#703 how to understand the code for calculating rewards

Opened 3 years ago
lyzKF
open
No labels
111 year ago

#986 Is there a micro-benchmark that can test the last two optimizations mentioned in the FastPersist paper, "Parallelizing checkpoint writes across DP ranks and Pipelining checkpoint writes"?

Opened 1 year ago
Buddingpopp
closed - completed
No labels
001 year ago

#984 moe example 404

Opened 1 year ago
xumj1
open
No labels
001 year ago

#186 OOM despite ZeRO stage 3

Opened 4 years ago
kchu02
open
No labels
101 year ago

#977 FastPersist micro-benchmarks test results are inconsistent with expectations

Opened 1 year ago
Buddingpopp
closed - completed
No labels
101 year ago

#979 Possible to include an example of DeepNVMe + state dict

Opened 1 year ago
sayakpaul
open
No labels
601 year ago

#943 KV_cache offload

Opened 2 years ago
yuzhenmao
open
No labels
321 year ago

#941 A bug in argument parser.

Opened 2 years ago
ChenDaiwei-99
closed - completed
No labels
101 year ago

#172 My deepspeed code is very slow

Opened 4 years ago
zhaowei-wang-nlp
open
No labels
3001 year ago

#956 Why Does vf_loss Take the Maximum Value, Rendering Clamp Meaningless?

Opened 2 years ago
Morizhaoyang
open
No labels
201 year ago

#969 Apply Zero-3 and LoRA appears empty lora weight [0]

Opened 1 year ago
jiangxinke
open
No labels
101 year ago

#845 torch.distributed.DistBackendError: NCCL error in: ../torch/csrc/distributed/c10d/ProcessGroupNCCL.cpp:1333, remote process exited or there was a network error, NCCL version 2.18.6

Opened 3 years ago
Rainbowman0
open
No labels
401 year ago

#972 DeepSpeed-Chat step-1 take a long time and no response

Opened 1 year ago
StoKou
closed - completed
No labels
001 year ago

#960 DeepSpeed-FastGen support ascend npu?

Opened 2 years ago
RyanOvO
open
No labels
301 year ago

#892 Does Zero-Inference support TP?

Opened 2 years ago
preminstrel
open
No labels
1101 year ago

#831 [Discussion] Can anyone show the performance on every step with any dataset

Opened 3 years ago
EeyoreLee
closed - completed
No labels
002 years ago

#946 Assertion `srcIndex < srcSelectDimSize` failed

Opened 2 years ago
boqiny
open
No labels
102 years ago

#456 enable_hybrid_engine issue

Opened 3 years ago
llllooong
open
deespeed chat
hybrid engine
1032 years ago

#951 Is there any example about DeepSpeed Zero with Ulysses/Ulysses-offload

Opened 2 years ago
LSC527
open
No labels
002 years ago

#950 Domino + PP

Opened 2 years ago
XZQshiyu
open
No labels
002 years ago

Rows per page:

1–25 of 46