Hacktoberfest 2026: những issue maintainer đã đánh dấu cho tháng Mười, đang mở và phù hợp người mới. Xem issue Hacktoberfest

Evaluate native HRX and Loom on gfx1100 with matched measurements

Đang mở
#3,080 0 bình luận 0 reaction 0 người được giao Xem trên GitHub

Maintainer thường phản hồi trong vòng 1 ngày

Chưa có ai nhận issue này.

Đánh giá

Độ khó
5/5
Thời gian dự kiến
Hơn một tuần
Mức phù hợp với người mới
25/100
Loại issue
Tính năng
Độ rõ ràng
Khá rõ ràng
Mức độ hoạt động
Sôi nổi
Công nghệ
cpp

Hướng nghiên cứu

Bắt đầu với .agents/specs/rocm-hrx-evaluation.md và các revision được ghim của HRX và llama.cpp, sau đó thiết lập các bản build theo cặp chính xác và việc thực thi native trên gfx1100. Ghi lại các hash có thể tái lập, lệnh, tính đúng đắn, thời gian khởi động, hiệu năng khi đã làm nóng, độ trễ, bộ nhớ và tranh chấp trên các workload được chỉ định. Được xem là hoàn tất khi có một khuyến nghị được đánh giá độc lập về việc áp dụng các phương pháp, tiếp tục theo đuổi một khoảng cách đã đo lường hoặc từ chối chúng.

Do mô hình lập chỉ mục viết ra từ nội dung của issue.

Mô tả

Row: BACKEND-ROCM

Evaluate native AMD HRX and Loom for the project on gfx1100. The developer requested hardware measurements before selecting an architecture. Existing vllm.cpp code is a comparator, not a preferred design or correctness oracle.

Build pinned HRX and the author-linked llama.cpp HRX integration. Establish native gfx1100 execution and correctness, compare HRX and conventional HIP on identical supported workloads, separate runtime-submission and kernel/compiler contributions, and assess vllm.cpp integration including Qwen3.5 GDN and quantized arms. Native HRX, its HIP compatibility layer, and selective Loom reuse are all candidates.

Starting anchors: ROCm/hrx-system 6bcd5a4ff111fa5bf160ab9f4592ca8e7cc810b1; AMD-Ecosystem/llama.cpp hrx-graph-develop-v2 at 6319038132ed12f968ea68f37753f705da830ea8; https://github.com/ggml-org/llama.cpp/discussions/27219. Establish exact-pair build compatibility. Spec: .agents/specs/rocm-hrx-evaluation.md, to be committed before implementation.

Record revisions, commands, binary/model hashes, device identity and contention, output correctness, cold-start/JIT and warm prefill/decode/latency/memory. Product correctness uses the applicable pinned oracle and unchanged thresholds. Upstream benchmark or compiler-smoke results are diagnostic evidence, not end-to-end vllm.cpp acceptance.

Finish with an independently reviewed empirical recommendation: adopt, pursue a specific measurable gap, or reject for measured reasons. Missing coverage is engineering scope to investigate. Shipped defaults require the normal acceptance gates.

The BACKEND-ROCM operator owns this evaluation. Task staging and builds stay in ignored build directories, with no /tmp or /dev/shm workspaces. Preserve compact reproducible evidence and remove disposable artifacts when no longer needed.

Ngôn ngữ chính
C++
Star
440
Fork
55
Merge trung bình
1 ngày 4 giờ
Pull request đã merge (30 ngày)
366

Chuẩn bị môi trường

Bắt đầu từ đâu

  1. Đọc hết issue, rồi đọc hướng dẫn đóng góp của dự án.
  2. Bình luận trên issue rằng bạn sẽ nhận — tránh hai người làm cùng một việc.
  3. Fork repository và làm thay đổi trên một nhánh.
  4. Mở pull request có tham chiếu số hiệu của issue.

Issue khác của mudler/vllm.cpp

Tất cả issue của mudler/vllm.cpp

Issue tương tự

Thêm issue về C++

Nhận issue mới trong hộp thư của bạn

Bản tóm tắt ngắn những issue GitHub phù hợp với người mới.