OCP FP8 E4M3FN #
OCP E4M3FN has a nominal format identity and a direct UInt8 runtime carrier. For addition,
subtraction, multiplication, division, and square root, balanced and throughput planning select
exhaustive byte tables; latency planning selects the arithmetic kernels. Fused multiply-add uses
the proved generic single-rounding kernel because an eight-bit ternary table would contain
16,777,216 entries.
The binary-interchange descriptor specifies the format and its kernels. It is not stored in
FloatLib.Floats.ExecFloat E4M3FN.
Reference #
- Open Compute Project, 8-bit Floating Point Specification, revision 1.0, E4M3, https://www.opencompute.org/documents/ocp-8-bit-floating-point-specification-ofp8-revision-1-0-2023-12-01-pdf-1.
OCP E4M3 finite-number FP8 with direct byte storage.
Instances For
The E4M3FN encoding fits in its direct byte carrier.
Certified direct-byte E4M3FN addition table.
Instances For
Certified direct-byte E4M3FN subtraction table.
Instances For
Certified direct-byte E4M3FN multiplication table.
Instances For
Certified direct-byte E4M3FN division table.
Instances For
Certified direct-byte E4M3FN square-root table.
Instances For
Package the five dense byte tables with the proved arithmetic FMA baseline. The tables total roughly 257 KiB; retaining the single-rounding model kernel avoids a separate 16 MiB FMA table.
Construct OCP E4M3FN from a natural bit pattern, reduced to eight bits.
Instances For
Read the complete OCP E4M3FN encoding byte.