Hacktoberfest 2026:维护者为十月标记出来的 issue,仍然开放、适合新手。 浏览 Hacktoberfest issue

About multiple-thread attention computation on CPU using zero-inference example.

未关闭
#886 0 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看

还没有人认领这个 Issue。

评估

难度
4/5
预计耗时
3-5 天
新手友好度
20/100
Issue 类型
缺陷
描述清晰度
需要澄清
活跃度
停滞
技术栈
python

调研方向

从 run_model.py 和提供的 DeepSpeed 命令开始,重点关注 CPU offload、pin-memory、KV offload 和 async KV offload 选项。重现 zero-interference run 并检查进程和核心的使用情况;完成的标准是确定缺少的配置,或确认该示例无法使用多个 CPU 核心。

由索引模型根据 Issue 内容生成。

描述

Hi,

I am trying to test the attention computation on the CPU with zero-interference.

I use the following command to run the script.

BSZ=96
LOG_DIR=$BASE_LOG_DIR/${MODEL_NAME}_bs${BSZ}
mkdir -p  $LOG_DIR
deepspeed --num_gpus 1 run_model.py --dummy --model ${FULL_MODEL_NAME} --batch-size ${BSZ} --cpu-offload --pin-memory 1 --offload-dir /tmp/data/dxu --gen-len 32 --pin-memory 1 --kv-offload --async_kv_offload

Duing the exection, I only saw one core is used.
Screenshot 2024-04-04 at 3 11 39 PM

However, there are many processes are created for the run.

Are there any other parameters or env I should configure to enable multiple-cores?

主要语言
Python
星标
6.8k
派生
1.1k
平均合并
2 天 16 小时
30 天内合并 PR
1

贡献指南

这个仓库没有索引到贡献指南

从这里开始

  1. 先读完整个 Issue,再读项目的贡献指南。
  2. 在 Issue 下留言说明你要接手 —— 这能避免两个人做同样的事。
  3. Fork 仓库,在一个分支上完成修改。
  4. 提交 Pull Request,并在描述里引用这个 Issue 编号。

deepspeedai/DeepSpeedExamples 的其他 Issue

查看 deepspeedai/DeepSpeedExamples 的全部 Issue

相似的 Issue

更多 Python Issue

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。