_mm256_bcstnebf16_psConvert scalar BF16 (16-bit) floating point element stored at memory locations starting at location
a to single precision (32-bit) floating-point, broadcast it to packed single precision (32-bit) floating-point
elements, and store the results in dst.
_mm256_bcstnesh_psConvert scalar half-precision (16-bit) floating-point element stored at memory locations starting
at location a to a single-precision (32-bit) floating-point, broadcast it to packed single-precision
(32-bit) floating-point elements, and store the results in dst.
_mm256_cvtneebf16_psConvert packed BF16 (16-bit) floating-point even-indexed elements stored at memory locations starting at
location a to single precision (32-bit) floating-point elements, and store the results in dst.
_mm256_cvtneeph_psConvert packed half-precision (16-bit) floating-point even-indexed elements stored at memory locations starting at
location a to single precision (32-bit) floating-point elements, and store the results in dst.
_mm256_cvtneobf16_psConvert packed BF16 (16-bit) floating-point odd-indexed elements stored at memory locations starting at
location a to single precision (32-bit) floating-point elements, and store the results in dst.
_mm256_cvtneoph_psConvert packed half-precision (16-bit) floating-point odd-indexed elements stored at memory locations starting at
location a to single precision (32-bit) floating-point elements, and store the results in dst.
_mm256_cvtneps_avx_pbhConvert packed single precision (32-bit) floating-point elements in a to packed BF16 (16-bit) floating-point
elements, and store the results in dst.
_mm_bcstnebf16_psConvert scalar BF16 (16-bit) floating point element stored at memory locations starting at location
a to single precision (32-bit) floating-point, broadcast it to packed single precision (32-bit)
floating-point elements, and store the results in dst.
_mm_bcstnesh_psConvert scalar half-precision (16-bit) floating-point element stored at memory locations starting
at location a to a single-precision (32-bit) floating-point, broadcast it to packed single-precision
(32-bit) floating-point elements, and store the results in dst.
_mm_cvtneebf16_psConvert packed BF16 (16-bit) floating-point even-indexed elements stored at memory locations starting at
location a to single precision (32-bit) floating-point elements, and store the results in dst.
_mm_cvtneeph_psConvert packed half-precision (16-bit) floating-point even-indexed elements stored at memory locations starting at
location a to single precision (32-bit) floating-point elements, and store the results in dst.
_mm_cvtneobf16_psConvert packed BF16 (16-bit) floating-point odd-indexed elements stored at memory locations starting at
location a to single precision (32-bit) floating-point elements, and store the results in dst.
_mm_cvtneoph_psConvert packed half-precision (16-bit) floating-point odd-indexed elements stored at memory locations starting at
location a to single precision (32-bit) floating-point elements, and store the results in dst.
_mm_cvtneps_avx_pbhConvert packed single precision (32-bit) floating-point elements in a to packed BF16 (16-bit) floating-point
elements, and store the results in dst.