global_atomic_add_f32 GPU Native ISA AMD Vector
Vector/Global Memory
Add two single-precision float values stored in the data register and a location in the global aperture.
Encoding
The ENC_VGLOBAL layout from AMD's machine-readable ISA specification (AMD CDNA 5). Opcode 86 in OP rebuilds 0x0000000000000000EE158000, an identifier AMD lists for this encoding. Bits marked unassigned have no field in the specification. Not the same in every generation: AMD CDNA 4 (layout not verifiable), AMD CDNA 3 (layout not verifiable), AMD CDNA 2 (layout not verifiable), AMD CDNA 1 (layout not verifiable), AMD RDNA 4 (field layout), AMD RDNA 3.5 (layout not verifiable), AMD RDNA 3 (layout not verifiable).
Operands
-
VDST
Written. Operand must be a vector GPR. Uses an 8-bit operand field. Data format: Value can be anything. (OPR_VGPR, FMT_ANY) -
VADDR
Read. Operand must be a vector GPR. Uses an 8-bit operand field. Data format: Value can be anything. (OPR_VGPR, FMT_ANY) -
VSRC
Read. Operand must be a vector GPR. Uses an 8-bit operand field. Data format: Value can be anything. (OPR_VGPR, FMT_ANY) -
SADDR
Read. Any scalar GPR operand including VCC and NULL. Data format: Value can be anything. (OPR_SREG, FMT_ANY)
In AMD's order, as its machine-readable ISA specification lists them for the ENC_VGLOBAL encoding (AMD CDNA 5).
GFX Target Compatibility
Per-target GFX compatibility has not yet been verified for this instruction.
Related
More in Vector/Global Memory
Reference
Description
Example
global_atomic_add_f32 v1, v2, vccA real instruction accepted by the LLVM assembler, taken verbatim from LLVM's own AMDGPU MC test suite (gfx11_asm_flat.s). Not from an AMD document, and not authored here.
Sources
- AMD Machine-Readable GPU ISA Specification ↗ - Advanced Micro Devices, Inc.
- LLVM MC assembler tests for AMDGPU ↗ - LLVM Project