Suggested GPU to run the demo code of step2_reward_model_finetuning (DeepSpeed-Chat)
还没有人认领这个 Issue。
评估
- 难度
- 4/5
- 预计耗时
- 3-5 天
- 新手友好度
- 20/100
- Issue 类型
- 缺陷
- 描述清晰度
- 需要澄清
- 活跃度
- 停滞
- 技术栈
- python
调研方向
先从 step2_reward_model_finetuning 说明和 training_scripts/opt/single_gpu/run_350m.sh 开始,然后在报告中的 V100 配置上重现内存不足失败。完成的标准是确定一个能够成功运行的文档化配置,或清楚地记录 GPU 要求和所需的更改。
由索引模型根据 Issue 内容生成。
描述
I follow the instructions of this page to do step2_reward_model_finetuning with demo code.
On the Google Cloud platform, I create one instance with a single V100(16GB) and another instance with double V100(16GB). I directly use this command bash training_scripts/opt/single_gpu/run_350m.sh but always meet the out-of-memory issues.
Are there any modifications I could do to run this demo code on double V100(16BG)? Or Are there recommendations about which type of GPU I should use to run this demo code successfully?
Appreciate the help!
- 主要语言
- Python
- 星标
- 6.8k
- 派生
- 1.1k
- 平均合并
- 2 天 16 小时
- 30 天内合并 PR
- 1
贡献指南
这个仓库没有索引到贡献指南
从这里开始
- 先读完整个 Issue,再读项目的贡献指南。
- 在 Issue 下留言说明你要接手 —— 这能避免两个人做同样的事。
- Fork 仓库,在一个分支上完成修改。
- 提交 Pull Request,并在描述里引用这个 Issue 编号。
deepspeedai/DeepSpeedExamples 的其他 Issue
-
难度 2/5 1-3 小时 新手友好度 52/100
deepspeedai/DeepSpeedExamples#996 ·
-
难度 4/5 3-5 天 新手友好度 25/100
deepspeedai/DeepSpeedExamples#995 ·
-
难度 2/5 1-3 小时 新手友好度 52/100
deepspeedai/DeepSpeedExamples#989 ·
-
moe example 404 未关闭
难度 4/5 3-5 天 新手友好度 25/100
deepspeedai/DeepSpeedExamples#984 ·
-
难度 4/5 3-5 天 新手友好度 42/100
deepspeedai/DeepSpeedExamples#979 · 6 条评论 ·
查看 deepspeedai/DeepSpeedExamples 的全部 Issue
相似的 Issue
-
area: harness bug status: needs-triage
难度 2/5 1-3 小时 新手友好度 75/100
Human-Agent-Society/reef#625 ·
-
难度 2/5 1-3 小时 新手友好度 70/100
-
难度 1/5 1 小时以内 新手友好度 80/100
learningequality/kolibri#15351 · 2 条评论 ·
-
难度 2/5 1-3 小时 新手友好度 75/100
-
Name consistency 未关闭
难度 2/5 1-3 小时 新手友好度 75/100
eellak/triplestore#65 · 1 条评论 ·