llvm/llvm-project

AMDGPU should handle SimplifyDemandedVectorElts for more trivial intrinsics

Ouverte

#131 734 ouverte le 18 mars 2025

 (9 commentaires) (0 réaction) (1 personne assignée)C++ (10 782 forks)batch import
backend:AMDGPUgood first issuemissed-optimization

Métriques du dépôt

Stars
 (26 378 étoiles)
Métriques de merge PR
 (Métriques PR en attente)

Description

As a follow up to https://github.com/llvm/llvm-project/pull/128647, more intrinsics should be handled in SimplifyDemandedVectorElts.

This includes: Intrinsic::amdgcn_readlane, Intrinsic::amdgcn_update_dpp, Intrinsic::amdgcn_permlane16, Intrinsic::amdgcn_permlanex16, Intrinsic::amdgcn_permlane64 and Intrinsic::amdgcn_mov_dpp8

This is mostly a matter of adding the intrinsics to the switch, some boilerplate to keep the other immediate operands as they are, and adding tests similar to the ones added in #128647

Guide contributeur