s_fmac_f32 GPU Native ISA AMD Scalar
Scalar Arithmetic
Multiply two single-precision float inputs and accumulate the result into the destination register using fused multiply add.
Encoding
The ENC_SOP2 layout from AMD's machine-readable ISA specification (AMD CDNA 5). Opcode 71 in OP rebuilds 0xA3800000, an identifier AMD lists for this encoding. The same in the other 2 generations that define it.
Operands
-
SDST
Written. All scalar destination operands. Data format: 32-bit single-precision floating point value. (OPR_SDST, FMT_NUM_F32) -
SSRC0
Read. All scalar operands. Covers all operands that are allowed as scalar sources. Data format: 32-bit single-precision floating point value. (OPR_SSRC, FMT_NUM_F32) -
SSRC1
Read. All scalar operands. Covers all operands that are allowed as scalar sources. Data format: 32-bit single-precision floating point value. (OPR_SSRC, FMT_NUM_F32)
In AMD's order, as its machine-readable ISA specification lists them for the ENC_SOP2 encoding (AMD CDNA 5).
GFX Target Compatibility
Per-target GFX compatibility has not yet been verified for this instruction.
Related
More in Scalar Arithmetic
Reference
Example
s_fmac_f32 s5, 0, s2A real instruction accepted by the LLVM assembler, taken verbatim from LLVM's own AMDGPU MC test suite (gfx1150_asm_salu_float.s). Not from an AMD document, and not authored here.
Sources
- AMD Machine-Readable GPU ISA Specification ↗ - Advanced Micro Devices, Inc.
- LLVM MC assembler tests for AMDGPU ↗ - LLVM Project