[Bug]: multi-arch `vgpu-device-manager` image digest not pinned in OLM bundle
维护者通常 1 天内回复
评估
- 难度
- 2/5
- 预计耗时
- 1-3 小时
- 新手友好度
- 70/100
- Issue 类型
- 缺陷
- 描述清晰度
- 描述清楚
- 活跃度
- 活跃
- 技术栈
- kubernetes, yaml
- 领域
- devops
调研方向
从bundle/manifests/gpu-operator-certified.clusterserviceversion.yaml第282行的镜像引用vgpu-device-manager开始,然后检查该 bundle 的镜像摘要是如何发布的。确认此引用指向多架构镜像,并验证基于 ARM 的 GPU 节点拉取的镜像能够运行,且不会出现报告中的 exec format error。
由索引模型根据 Issue 内容生成。
描述
Describe the bug
The vgpu-device-manager image digest in the OLM bundle refers to the amd64 digest instead of the multi-arch image digest. This leads to the following error message when running the image on an ARM system:
exec container process `/usr/bin/nvidia-k8s-vgpu-dm`: Exec format error
To Reproduce
Follow the procedure for deploying the GPU Operator with OpenShift Virtualization. Configure GPU nodes with the vm-vgpu workload type.
Expected behavior
The correct vgpu-device-manager image for ARM gets pulled on an ARM-based GPU node.
Environment (please provide the following information):
- GPU Operator Version: v26.7.1
- OS: RHEL
- Kernel Version: 6.12
- Container Runtime Version: cri-o
- Kubernetes Distro and Version: OpenShift
- 主要语言
- Go
- 星标
- 2.9k
- 派生
- 569
- 平均合并
- 1 天 10 小时
- 30 天内合并 PR
- 78
环境准备
- 没有 Dockerfile 或 Docker Compose 文件
- 有 Pull Request 模板
- 阅读贡献指南
从这里开始
- 先读完整个 Issue,再读项目的贡献指南。
- 在 Issue 下留言说明你要接手 —— 这能避免两个人做同样的事。
- Fork 仓库,在一个分支上完成修改。
- 提交 Pull Request,并在描述里引用这个 Issue 编号。
NVIDIA/gpu-operator 的其他 Issue
-
[Bug]: Driver upgrade does not evict pods that use nvidia.com/gpu only in a native sidecar可能已有人在做 关联的 PR 仍在进行中或已合并。 未关闭bug needs-triage
难度 2/5 1-3 小时 新手友好度 78/100
NVIDIA/gpu-operator#3026 ·
维护者通常 1 天内回复
-
Make NVIDIADriver node-pool rendering deterministic可能已有人在做 @efegokdemir 于 8 天前认领。 未关闭dsx-ws-0930 good-first-issue
难度 2/5 1-3 小时 新手友好度 88/100
NVIDIA/gpu-operator#2981 · 1 条评论 ·
维护者通常 1 天内回复
-
[Bug]: GPUCluster common name label breaks DRA validator selector可能已有人在做 @ajavanma 于 16 天前认领。 未关闭
难度 2/5 1-3 小时 新手友好度 88/100
NVIDIA/gpu-operator#2955 · 1 条评论 ·
维护者通常 1 天内回复
-
bug needs-triage
难度 4/5 3-5 天 新手友好度 48/100
NVIDIA/gpu-operator#3027 ·
维护者通常 1 天内回复
-
Ensure automated backport commits have verified signatures可能已有人在做 @asivanadi0 于 7 天前认领。 未关闭good-first-issue
难度 4/5 3-5 天 新手友好度 55/100
NVIDIA/gpu-operator#2997 · 1 条评论 ·
维护者通常 1 天内回复
查看 NVIDIA/gpu-operator 的全部 Issue
相似的 Issue
-
Idle compaction monitors LIST the replica every tick when the newest destination file spans more than one TXID可能已有人在做 @pishuv 今天认领。 未关闭
难度 2/5 1-3 小时 新手友好度 72/100
benbjohnson/litestream#1563 ·
维护者通常 2 天内回复
-
难度 1/5 1 小时以内 新手友好度 88/100
维护者通常 1 天内回复
-
agent-research agent-review-finding chore
难度 2/5 1-3 小时 新手友好度 66/100
jordansmall/spindrift#4922 ·
维护者通常 1 天内回复
-
gcsartifact: deleting a missing version returns an error可能已有人在做 @ktsoator 今天认领。 未关闭bug
难度 2/5 1-3 小时 新手友好度 78/100
维护者通常 2 天内回复
-
govulncheck
难度 2/5 1-3 小时 新手友好度 62/100
维护者通常 1 天内回复