global_atomic_fmax GPU Native ISA AMD Vector
Vector/Global Memory
Select the maximum of two single-precision float inputs, given two values stored in the data register and a location in the global aperture.
Also written as
global_atomic_max_f32, global_atomic_max_num_f32.
AMD's machine-readable ISA specification lists these names for the same instruction.
Encoding
Operands
-
VDST
Written. Operand must be a vector GPR. Uses an 8-bit operand field. Data format: Value can be anything. (OPR_VGPR, FMT_ANY) -
ADDR
Read. Operand must be a vector GPR. Uses an 8-bit operand field. Data format: Value can be anything. (OPR_VGPR, FMT_ANY) -
DATA
Read. Operand must be a vector GPR. Uses an 8-bit operand field. Data format: Value can be anything. (OPR_VGPR, FMT_ANY) -
SADDR
Read. N/A Data format: Value can be anything. (OPR_SREG, FMT_ANY)
In AMD's order, as its machine-readable ISA specification lists them for the ENC_FLAT_GLBL encoding (AMD RDNA 2).
GFX Target Compatibility
Per-target GFX compatibility has not yet been verified for this instruction.
Related
More in Vector/Global Memory
Reference
Description
Example
global_atomic_fmax v[1:2], v2, off dlcA real instruction accepted by the LLVM assembler, taken verbatim from LLVM's own AMDGPU MC test suite (gfx10_asm_flat.s). Not from an AMD document, and not authored here.
Sources
- AMD Machine-Readable GPU ISA Specification ↗ - Advanced Micro Devices, Inc.
- LLVM MC assembler tests for AMDGPU ↗ - LLVM Project