Function _mm_maskz_fnmsub_round_sh
#[target_feature(enable = "avx512fp16")]
pub fn _mm_maskz_fnmsub_round_sh(k: __mmask8, a: __m128h, b: __m128h, c: __m128h, ROUNDING: i32) -> __m128h
Multiply the lower half-precision (16-bit) floating-point elements in a and b, and subtract the intermediate result from the lower element in c. Store the result in the lower element of dst using zeromask k (the element is zeroed out when the mask bit 0 is not set), and copy the upper 7 packed elements from a to the upper elements of dst.
Rounding is done according to the rounding parameter, which can be one of:
_MM_FROUND_TO_NEAREST_INT|_MM_FROUND_NO_EXC: round to nearest and suppress exceptions_MM_FROUND_TO_NEG_INF|_MM_FROUND_NO_EXC: round down and suppress exceptions_MM_FROUND_TO_POS_INF|_MM_FROUND_NO_EXC: round up and suppress exceptions_MM_FROUND_TO_ZERO|_MM_FROUND_NO_EXC: truncate and suppress exceptions_MM_FROUND_CUR_DIRECTION: useMXCSR.RC- see_MM_SET_ROUNDING_MODE