v_fma_f16 GPU Native ISA AMD Vector
V FMA F16 Vector Arithmetic
Multiply two half-precision float inputs and add a third input using fused multiply add, and store the result into a vector register.
Encoding
Operands
Operand details have not yet been curated for this instruction.
GFX Target Compatibility
| Target | Support |
|---|---|
| gfx1100 | ✅ Supported |
0.5ULP accuracy, denormals are supported. If OPSEL[3] is 0 Result is written to 16 LSBs of destination VGPR and hi 16 bits are preserved. If OPSEL[3] is 1 Result is written to 16 MSBs of destination VGPR and lo 16 bits are preserved.
Related
More in Vector Arithmetic
Reference
Semantics
Example
v_fma_f16 v5, v1, v2, s3A real instruction accepted by the LLVM assembler, taken verbatim from LLVM's own AMDGPU MC test suite (gfx11). Not from an AMD document, and not authored here.
Sources
- User Guide for AMDGPU Backend ↗ - LLVM Project
-
"AMD Instinct MI300" Instruction Set Architecture: Reference Guide ↗
- Advanced Micro Devices, Inc.
Reference Guide, page 358. - LLVM MC assembler tests for AMDGPU (gfx11) ↗ - LLVM Project